Jina AI Releases jina-ocr-v1: A 3.4B MoE Document Parser With Built-In Speculative Decoding for Low-Budget GPUs
MarkTechPost
Read Full Article at MarkTechPost →Ad Slot — In-Article (728x90)
Jina AI has released jina-ocr-v1, a visual document parser that converts PDFs, scans, tables, charts and invoices into Markdown. The model has 3. 4B total parameters, with about 570M active per token, and builds on DeepSeek-OCR.
A built-in FastMTP speculative decoding head drafts 3 tokens per step while keeping output lossless. It scores 91. 14 on OmniDocBench v1. 6 and 83. 4 on olmOCR-Bench, and parses 2. 57 pages per second on 1 A100. Weights are on Hugging Face under CC BY-NC 4. 0, with hosted access through Jina Reader.
This is a summary. For the full story, read the original article at MarkTechPost.
Original source: MarkTechPost