LightOnOCR-3 turns complex documents into structured content that applications can use directly.

Building on the success of LightOnOCR, with more than 4.5 million downloads on Hugging Face, this new family of end-to-end OCR models goes beyond accurate text transcription: it identifies and locates layout elements, describes images, and extracts data from charts and scientific figures.

- Leading transcription accuracy in compact models: Our models have leading results across three open benchmarks: first on FRBench-pdf2md, with particular strengths in handwriting, second on OlmOCR-Bench, and first among open-weight models on ParseBench.
- Document structure and content-aware chunking: The model groups content into logical paragraphs, headings, tables and other document elements, each with a label and bounding box. These spatially grounded blocks support downstream chunking, retrieval and links back to the original page.
- Image descriptions and structured chart data: LightOnOCR-3 models describe images and extract data from charts, plots and scientific figures, making visual information accessible to search and analysis alongside the document’s text.
- Single models that can be prompted in two modes: A simple prompt change switches between plain transcription, matching LightOnOCR-2 output, and richer grounded output.


