comparar
OCRmyPDF vs Tesseract
Os mesmos fatos para os dois, lidos do GitHub toda noite, e a relação que uma pessoa revisou.
| Fato | OCRmyPDF | Tesseract |
|---|---|---|
| Linguagem | Python | C++ |
| Licença | MPL-2.0 | Apache-2.0 |
| Estrelas | 35k | 77k |
| Última release | v17.13.0 | 5.5.3 |
| Último push | 2026-10-06 | 2026-09-28 |
| Ritmo de releases | cerca de 15 dias entre releases | cerca de 203 dias entre releases |
| Contribuidores ativos | 8 autores de commits na branch padrão nos últimos 90 dias | 12 autores de commits na branch padrão nos últimos 90 dias |
| Alertas | nenhum | nenhum |
Como se relacionam
Os dois substituem ABBYY FineReader. Alternativas a ABBYY FineReader →
ParcialOCRmyPDFAdds a searchable text layer to scanned PDFs from the command line.
ParcialTesseractThe OCR engine and a command line, without a document editor.
OCRmyPDF
- v17.13.02026-09-28PDF/A made without Ghostscript ("speculative" conversion) is now validated
- v17.12.12026-09-16Fixes
- v17.12.02026-09-16OCRmyPDF now requires pikepdf 10.2 or later, up from pikepdf 10. This is the
- v17.11.02026-08-28Enhancements
- v17.10.02026-08-05The watcher.py watched-folder helper (the watcher extra) has been
OCRmyPDF
não assinada A última release, v17.13.0, não tem nenhuma assinatura que o GitHub consiga verificar.
Carregando o relatório de segurança
Tesseract
✓ assinada A última release, 5.5.3, tem uma assinatura verificada pelo GitHub.
Carregando o relatório de segurança