Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
OCR for 100+ languages with layout and table recognition, as a Python library and CLI.
- 言語
- Python
- ライセンス
- Apache-2.0
- スター
- 91k
- 最新リリース
- v3.7.0
- 最終プッシュ
- 2026-09-16