📄 Awesome OCR multiple programing languages toolkits based on ONNX Runtime, OpenVINO, MNN, PaddlePaddle, TensorRT and PyTorch.
☆7,621Aug 30, 2026Updated this week
Alternatives and similar repositories for RapidOCR
Users that are interested in RapidOCR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/…☆88,484Jul 22, 2026Updated last month
- rapidocr onnx cpp☆371Mar 25, 2025Updated last year
- 基于PaddleOCR重构,并且脱离PaddlePaddle深度学习训练框架的轻量级OCR,推理速度超快 —— A lightweight OCR system based on PaddleOCR, decoupled from the PaddlePaddle d…☆1,863Jun 11, 2026Updated 2 months ago
- Official code implementation of General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model☆8,220Feb 10, 2025Updated last year
- OCR software, free and offline. 开源、免费的离线OCR软件。支持截屏/批量导入图片,PDF文档识别,排除水印/页眉页脚,扫描/生成二维码。内置多国语言库。☆46,992Nov 20, 2025Updated 9 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- 🔥🔥🔥Java代码实现调用RapidOCR(基于PaddleOCR),适配Mac、Win、Linux,支持最新PP-OCRv4☆588Jun 5, 2024Updated 2 years ago
- OCR, layout analysis, reading order, table recognition in 90+ languages☆21,333Aug 21, 2026Updated last week
- 超轻量级中文ocr,支持竖排文字识别, 支持ncnn、mnn、tnn推理 ( dbnet(1.8M) + crnn(2.5M) + anglenet(378KB)) 总模型仅4.7M☆12,339May 18, 2026Updated 3 months ago
- 整理目前开源的最优表格识别模型,完善前后处理,模型转换为ONNX | Organize the currently open-source optimal table recognition models, improve pre-processing and post-…☆962Aug 3, 2025Updated last year
- Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.☆78,796Updated this week
- 文档方向分类☆223Feb 3, 2026Updated 6 months ago
- Convert the model in PaddleOCR to ONNX format☆120Jul 15, 2025Updated last year
- Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and …☆29,953Dec 5, 2025Updated 8 months ago
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆20,081Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Based on RapidOCR, extract the PDF content☆193Mar 6, 2026Updated 5 months ago
- Toolkit for linearizing PDFs for LLM datasets/training☆19,402Mar 25, 2026Updated 5 months ago
- OCR离线图片文字识别命令行windows程序,以JSON字符串形式输出结果,方便别的程序调用。提供各种语言API。由 PaddleOCR C++ 编译。☆1,543Apr 7, 2025Updated last year
- PaddleOCR inference in PyTorch. Converted from [PaddleOCR](https://github.com/PaddlePaddle/PaddleOCR)☆1,205Jul 10, 2026Updated last month
- DocLayout-YOLO: Enhancing Document Layout Analysis through Diverse Synthetic Data and Global-to-Local Adaptive Perception☆2,259Apr 14, 2025Updated last year
- 基于序列表格识别算法推理库,集成PP-Structure和modelscope等表格识别算法。☆436Apr 23, 2026Updated 4 months ago
- A lightweight LMM-based Document Parsing Model☆6,636Jul 20, 2026Updated last month
- Analysis of Chinese and English layouts 中英文版面分析☆277Mar 24, 2026Updated 5 months ago
- CnOCR: Awesome Chinese/English OCR Python toolkits based on PyTorch. It comes with 20+ well-trained models for different application scen…☆3,772Jul 5, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- OCRFlux is a lightweight yet powerful multimodal toolkit that significantly advances PDF-to-Markdown conversion, excelling in complex lay…☆2,533Apr 14, 2026Updated 4 months ago
- Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio…☆9,179Updated this week
- Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain…☆38,594Nov 10, 2025Updated 9 months ago
- FastGPT is a knowledge-based platform built on the LLMs, offers a comprehensive suite of out-of-the-box capabilities such as data process…☆29,499Updated this week
- RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to creat…☆89,648Updated this week
- Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self…☆153,903Updated this week
- A Comprehensive Toolkit for High-Quality PDF Content Extraction☆10,000Jan 3, 2025Updated last year
- SOTA Open Source TTS☆32,470Aug 22, 2026Updated last week
- 基于Pytorch的OCR工具库,支持常用的文字检测和识别算法☆1,523Jan 4, 2026Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 开源易用的中文离线OCR,识别率媲美大厂,并且提供了易用的web页面及web的接口,方便人类日常工作使用或者其他程序来调用~☆2,877Jun 14, 2023Updated 3 years ago
- Question and Answer based on Anything.☆14,081Mar 24, 2025Updated last year
- High-performance Inference and Deployment Toolkit for LLMs and VLMs based on PaddlePaddle☆3,711Updated this week
- Multilingual Document Layout Parsing in a Single Vision-Language Model☆9,094Mar 24, 2026Updated 5 months ago
- A collection of original, innovative ideas and algorithms towards Advanced Literate Machinery. This project is maintained by the OCR Team…☆1,834Mar 17, 2026Updated 5 months ago
- Convert PDF to markdown + JSON quickly with high accuracy☆39,382Updated this week
- A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone☆26,260Updated this week