pdfOCR is an iText add-on to recognize and extract text in scanned documents and images. It can also convert them into fully ISO-compliant PDF or PDF/A-3u files that are accessible, searchable, and suitable for archiving
-
Updated
Aug 26, 2026 - C#
pdfOCR is an iText add-on to recognize and extract text in scanned documents and images. It can also convert them into fully ISO-compliant PDF or PDF/A-3u files that are accessible, searchable, and suitable for archiving
pdfOCR is an iText add-on to recognize and extract text in scanned documents and images. It can also convert them into fully ISO-compliant PDF or PDF/A-3u files that are accessible, searchable, and suitable for archiving
Find what is extractable from your code — scans any GitHub repo, scores self-contained components worth shipping as standalone packages. Tested on Twitter, Flask, React, Kubernetes, CPython. 58K+ files, zero crashes.
To associate your repository with the extractable topic, visit your repo's landing page and select "manage topics."