Skip to main content
medical-ocr is a multi-engine OCR pipeline for medical and legal documents. It extracts structured data — ICD codes, CPT codes, medications, timelines, impairment ratings — from PDFs and scanned documents.

GitHub

nometria/medical-ocr

PyPI

medical-ocr on PyPI

Install

Usage

Extraction pipeline

Supported document types

Output format