ocrmypdf/OCRmyPDF
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
A Python wrapper for Google Tesseract
Appears on
Quick read
Latest capture 2026-08-23 03:06
0 paths
Agent instructions and tool configuration found in this repository.
No config files detected.
9 observed captures since 2026-05-22. Observed captures are shown by default.
Stars from first capture +41
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
Models, data loaders and abstractions for language processing, powered by PyTorch
PyTorch Geometric Temporal: Spatiotemporal Signal Processing with Neural Machine Learning Models (CIKM 2021)
A multi-voice TTS system trained with an emphasis on quality
Python library and CLI tool to interface with Google Translate's text-to-speech API
Multi-modal OCR pipeline optimized for ML training (text, figure, math, tables, diagrams)