ocr
157 个项目 · ⭐ 880.8kRuVector is a High Performance, Real-Time, Self-Learning Ai, Vector GNN, Memory DB built in Rust.
The media player for language learning, with dual subtitles, AI-generated subtitles, real-time translation, and more!
Text recognition (optical character recognition) with deep learning methods, ICCV 2019
A tensorflow implementation of EAST text detector
A wrapper to work with Tesseract OCR inside PHP.
Enhances Tesseract OCR output using LLMs (local or API) for error correction, smart chunking, and markdown formatting of scanned PDFs
A Model Context Protocol server for converting almost anything to Markdown
Optical character recognition for Japanese text, with the main focus being Japanese manga
Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI
基于 manga-image-translator 的开源漫画翻译工具。支持日/韩/美漫自动翻译,内置 OpenAI、Gemini 等 5 种翻译引擎,并提供可视化编辑器自由调整文本样式。一键安装,开箱即用。如果喜欢,欢迎点亮 ⭐ Star 支持!
LLM Agent Framework in ComfyUI includes MCP sever, Omost,GPT-sovits, ChatTTS,GOT-OCR2.0, and FLUX prompt nodes,access to Feishu,discord,and adapts to all llms with similar openai / aisuite interfaces, such as o1,ollama, gemini, grok, qwen, GLM, deepseek, kimi,doubao. Adapted to local llms, vlm, gguf such as llama-3.3 Janus-Pro, Linkage graphRAG
Handwritten Text Recognition (HTR) system implemented with TensorFlow.
Convolutional Recurrent Neural Network (CRNN) for image-based sequence recognition.
An on-premises, OCR-free unstructured data extraction, markdown conversion and benchmarking toolkit. (https://idp-leaderboard.org/)
共 157 条 · 第 3 / 8 页