OCR

5 pieces on OCR.

News & Analysis

news
Shan2026-09-08
Document AIOCRModel ArchitectureInference CostEnterprise AI

Reducto r-1 Cuts Document Parsing Cost to 1¢ Per Page With Single-Pass Model

Reducto r-1 replaces a 4-stage agentic OCR pipeline with one full-page pass, cutting price from 3–6¢ to 1¢ and reporting 20% fewer errors.

Read more
news
Shan2026-09-03
CohereDocument AIVision Language ModelOCRRAGOpen Weights

Cohere Parse 5: 2.3B-Parameter VLM Beats Gemini 3 Flash on ParseBench

Cohere's Parse 5 scores 79.2 on ParseBench, outperforming Mistral OCR and Gemini 3 Flash with a 2.3B-parameter VLM that outputs Markdown plus bounding boxes.

Read more
news
Shan2026-08-27
CohereVision Language ModelsDocument ParsingEnterprise AIOCR

Cohere Parse 5: 2.3B Model Converts Enterprise Docs to Markdown at $1.50/1K Pages

Cohere's parse-v5.0 replaces OCR stacks with a single 2.3B vision-language model call — but its 79.2 ParseBench score covers only 3 of 5 dimensions.

Read more
How Baidu Unlimited-OCR Solves Long-Document Transcription
news
Shan2026-08-14
Computer VisionOCRBaiduOpen WeightsVision-Language Models

How Baidu Unlimited-OCR Solves Long-Document Transcription

Baidu's Unlimited-OCR fixes the output-side KV cache bottleneck in long-document OCR using Reference Sliding Window Attention, keeping memory constant at m+n tokens.

Read more
articles
Shan2026-06-21
Open SourceOCRDocument AI

Chandra OCR 2 Shows How Fast Open-Source Document AI Is Catching Up

Datalab's Chandra OCR 2 is pushing open-source OCR past legacy parsers with stronger layout, math, table, and multilingual performance.

Read more