Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
📊 Project Info
- Language
- Python
- Stars
- ⭐ 78,961
- Forks
- 10,519
- Today
- +118
- Ranking
- #10
- Collection
- Language
- Trending Date
- May 29, 2026
- Last Push
- 5/29/2026
🏷️ Topics
ai4sciencechineseocrdocument-parsingdocument-translationkieocrpaddleocr-vlpdf-extractor-ragpdf-parserpdf2markdownpp-ocrpp-structurerag


