Chandra OCR 2 Chandra 2 is a state of the art OCR model from Datalab that outputs markdown, HTML, and JSON. It is highly accurate at extracting text from images and PDFs, while preserving layout information.

Try Chandra in the free playground, or use the hosted API for higher accuracy and speed.

What's New in Chandra 2 85.8% olmocr bench score (sota), 77.8% multilingual bench score (12% improvement over Chandra 1) Significant improvements to math, tables, complex layouts Improved layout, especially on wider documents Significantly better image captioning 90+ language support with major accuracy gains Features Convert documents to markdown, HTML, or JSON with detailed layout information Excellent handwriting support Reconstructs forms accurately, including checkboxes Strong performance with tables, math, and complex layouts Extracts images and diagrams, with captions and structured data Support for 90+ languages

Downloads last month
251
GGUF
Model size
5B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

2-bit

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Yashp2003/chandra-ocr-2-ggufs

Quantized
(38)
this model