Tag: #donut
Writing on GPUs, LLMs, MLOps, Kubernetes — and mindset · 2 posts
Beyond OCR — OCR-free Document Understanding and Unified Models
A traditional OCR pipeline splits into detection, recognition, and layout stages, but errors accumulate. We organize the shift in document AI: Donut-style and VLM-based OCR-free document understanding, high-resolution an
2026-06-26 · 20 min read #ai-papers#ocr-free#document-understanding#multimodal#donutDocument AI / OCR in 2026 — Mistral OCR / Marker / Surya / LlamaParse / Docling / OlmoOCR Deep Dive
Document AI in 2026 is no longer "extract text with Tesseract." Purpose-built APIs like Mistral OCR (March 2025), open-source PDF-to-Markdown engines like Marker / Surya / Docling / OlmoOCR, pretrained document models li
2026-05-15 · 19 min read #ocr#document-ai#pdf#mistral-ocr#marker