📄 Awesome OCR multiple programing languages toolkits based on ONNX Runtime, OpenVINO, MNN, PaddlePaddle, TensorRT and PyTorch.
☆7,877Sep 19, 2026Updated this week
Alternatives and similar repositories for RapidOCR
Users that are interested in RapidOCR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/…☆89,827Updated this week
- rapidocr onnx cpp☆374Mar 25, 2025Updated last year
- 基于PaddleOCR重构,并且脱离PaddlePaddle深度学习训练框架的轻量级OCR,推理速度超快 —— A lightweight OCR system based on PaddleOCR, decoupled from the PaddlePaddle d…☆1,868Jun 11, 2026Updated 3 months ago
- Official code implementation of General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model☆8,220Feb 10, 2025Updated last year
- OCR software, free and offline. 开源、免费的离线OCR软件。支持截屏/批量导入图片,PDF文档识别,排除水印/页眉页脚,扫描/生成二维码。内置多国语言库。☆47,394Nov 20, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 🔥🔥🔥Java代码实现调用RapidOCR(基于PaddleOCR),适配Mac、Win、Linux,支持最新PP-OCRv4☆593Jun 5, 2024Updated 2 years ago
- OCR, layout analysis, reading order, table recognition in 90+ languages☆21,399Sep 11, 2026Updated last week
- 超轻量级中文ocr,支持竖排文字识别, 支持ncnn、mnn、tnn推理 ( dbnet(1.8M) + crnn(2.5M) + anglenet(378KB)) 总模型仅4.7M☆12,346May 18, 2026Updated 4 months ago
- 整理目前开源的最优表格识别模型,完善前后处理,模型转换为ONNX | Organize the currently open-source optimal table recognition models, improve pre-processing and post-…☆964Aug 3, 2025Updated last year
- Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.☆80,251Updated this week
- 文档方向分类☆224Feb 3, 2026Updated 7 months ago
- Convert the model in PaddleOCR to ONNX format☆120Jul 15, 2025Updated last year
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆20,430Updated this week
- Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and …☆30,005Dec 5, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Based on RapidOCR, extract the PDF content☆195Mar 6, 2026Updated 6 months ago
- Toolkit for linearizing PDFs for LLM datasets/training☆19,630Mar 25, 2026Updated 5 months ago
- OCR离线图片文字识别命令行windows程序,以JSON字符串形式输出结果,方便别的程序调用。提供各种语言API。由 PaddleOCR C++ 编译。☆1,549Apr 7, 2025Updated last year
- PaddleOCR inference in PyTorch. Converted from [PaddleOCR](https://github.com/PaddlePaddle/PaddleOCR)☆1,210Jul 10, 2026Updated 2 months ago
- DocLayout-YOLO: Enhancing Document Layout Analysis through Diverse Synthetic Data and Global-to-Local Adaptive Perception☆2,275Apr 14, 2025Updated last year
- 基于序列表格识别算法推理库,集成PP-Structure和modelscope等表格识别算法。☆438Apr 23, 2026Updated 4 months ago
- A lightweight LMM-based Document Parsing Model☆6,646Jul 20, 2026Updated 2 months ago
- Analysis of Chinese and English layouts 中英文版面分析☆279Mar 24, 2026Updated 5 months ago
- CnOCR: Awesome Chinese/English OCR Python toolkits based on PyTorch. It comes with 20+ well-trained models for different application scen…☆3,769Jul 5, 2026Updated 2 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio…☆9,329Sep 10, 2026Updated last week
- FastGPT is a knowledge-based platform built on the LLMs, offers a comprehensive suite of out-of-the-box capabilities such as data process…☆29,693Updated this week
- OCRFlux is a lightweight yet powerful multimodal toolkit that significantly advances PDF-to-Markdown conversion, excelling in complex lay…☆2,530Apr 14, 2026Updated 5 months ago
- Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain…☆38,643Nov 10, 2025Updated 10 months ago
- RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to creat…☆90,993Updated this week
- Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self…☆156,438Updated this week
- A Comprehensive Toolkit for High-Quality PDF Content Extraction☆10,022Jan 3, 2025Updated last year
- SOTA Open Source TTS☆32,745Updated this week
- 基于Pytorch的OCR工具库,支持常用的文字检测和识别算法☆1,526Jan 4, 2026Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 开源易用的中文离线OCR,识别率媲美大厂,并且提供了易用的web页面及web的接口,方便人类日常工作使用或者其他程序来调用~☆2,880Jun 14, 2023Updated 3 years ago
- Question and Answer based on Anything.☆14,177Sep 3, 2026Updated 2 weeks ago
- High-performance Inference and Deployment Toolkit for LLMs and VLMs based on PaddlePaddle☆3,716Aug 26, 2026Updated 3 weeks ago
- Multilingual Document Layout Parsing in a Single Vision-Language Model☆9,122Mar 24, 2026Updated 5 months ago
- Convert PDF to markdown + JSON quickly with high accuracy☆39,845Sep 13, 2026Updated last week
- A collection of original, innovative ideas and algorithms towards Advanced Literate Machinery. This project is maintained by the OCR Team…☆1,835Mar 17, 2026Updated 6 months ago
- A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone☆26,411Sep 8, 2026Updated last week