Pre-OCR Image Normalization — perspective correction, deskew, denoise for medical scans
-
Updated
Jul 7, 2026 - Python
Pre-OCR Image Normalization — perspective correction, deskew, denoise for medical scans
OpenCV-based document scanner with perspective correction, adaptive binarization, de-shadowing, batch processing and scan quality assessment.
Preflight checks for document extraction pipelines — validate, render, and screen PDFs before they reach your LLM. Pure-Python wheel, in-memory only.
Convert multi-page PDFs into clean, OCR-ready images... CLI, watch folder, web UI, and optional Google Drive sync.
Segment medical documents from smartphone photos using U²-Net, with tools for annotation, fine-tuning, and OCR-ready preprocessing.
Add a description, image, and links to the ocr-preprocessing topic page so that developers can more easily learn about it.
To associate your repository with the ocr-preprocessing topic, visit your repo's landing page and select "manage topics."