compare
OCRmyPDF vs Surya
The same facts for both, read from GitHub every night, and the relation a person reviewed.
| Fact | OCRmyPDF | Surya |
|---|---|---|
| Language | Python | Python |
| Licence | MPL-2.0 | Apache-2.0 |
| Stars | 35k | 21k |
| Latest | v17.13.0 | v0.22.1 |
| Last push | 2026-10-06 | 2026-09-11 |
| Release cadence | about 15 days between releases | about 3 days between releases |
| Active contributors | 8 commit authors on the default branch in the last 90 days | 1 commit author on the default branch in the last 90 days |
| Flags | none | none |
How they relate
Both replace ABBYY FineReader. Alternatives to ABBYY FineReader →
PartialOCRmyPDFAdds a searchable text layer to scanned PDFs from the command line.
PartialSuryaOCR, layout and reading-order detection in 90+ languages, from Python or a CLI.
OCRmyPDF
- v17.13.02026-09-28PDF/A made without Ghostscript ("speculative" conversion) is now validated
- v17.12.12026-09-16Fixes
- v17.12.02026-09-16OCRmyPDF now requires pikepdf 10.2 or later, up from pikepdf 10. This is the
- v17.11.02026-08-28Enhancements
- v17.10.02026-08-05The watcher.py watched-folder helper (the watcher extra) has been
OCRmyPDF
unsigned The latest release, v17.13.0, carries no signature GitHub could verify.
Loading the security report
Surya
✓ signed The latest release, v0.22.1, carries a signature GitHub verified.
Loading the security report