pdfOCR is an iText add-on for Java to recognize and extract text in scanned documents and images. It can also convert them into fully ISO-compliant PDF or PDF/A-3u files that are accessible, searchable, and suitable for archiving

Artifacts using PdfOCR API (7)
Sort by:Popular

Pdf2data
Last Release on Jul 8, 2024
pdfOCR-Tesseract4 is an iText add-on for Java to recognize and extract text in scanned documents and images. It can also convert them into fully ISO-compliant PDF or PDF/A-3u files that are accessible, searchable, and suitable for archiving
Last Release on Jul 9, 2026
pdfOCR-Onnx is an iText add-on for Java to recognize and extract text in scanned documents and images. It can also convert them into fully ISO-compliant PDF or PDF/A-3u files that are accessible, searchable, and suitable for archiving
Last Release on Jul 9, 2026
pdfOCR-OnnxTR is an iText add-on for Java to recognize and extract text in scanned documents and images. It can also convert them into fully ISO-compliant PDF or PDF/A-3u files that are accessible, searchable, and suitable for archiving
Last Release on Dec 29, 2025
IText 7 Functional Tests
Last Release on Feb 12, 2024
Pdf2data Default OCR Engine
Last Release on Jul 8, 2024
IText Functional Tests
Last Release on Jul 6, 2026
  • Prev
  • 1
  • Next