← all releases
Jina AI·Sep 18, 2026·v1·open source
jina-ocr-v1
jina-ocr-v1 is an efficient, end-to-end document parsing model designed to convert images and PDFs into structured Markdown in a single pass. It utilizes a 3.4B parameter mixture-of-experts architecture with 570M active parameters to deliver high-performance document transcription on low-budget hardware.
OCRDocument ParsingVision-Language ModelSpeculative Decoding
overview
jina-ocr-v1 is an efficient, end-to-end document parsing model designed to convert images and PDFs into structured Markdown in a single pass. It utilizes a 3.4B parameter mixture-of-experts architecture with 570M active parameters to deliver high-performance document transcription on low-budget hardware.
key features
- 01Lossless speculative decoding via FastMTP for accelerated inference
- 02Efficient 3.4B parameter MoE architecture with 570M active parameters
- 03Native Markdown output for tables, formulas, and text
- 04High-throughput performance reaching 2.57 pages per second
- 05Dynamic resolution support for varied document layouts
related productsbrowse all →