AI Pulse by Inblix

Topic: ocr

5 articles

Explore our coverage of ocr — 5 curated articles, summaries, and related resources from the Inblix archive.

PixelRAG ditches HTML parsing, retrieves documents as screenshots with 7-query Recall@k eval — Inblix summary
AI News

PixelRAG ditches HTML parsing, retrieves documents as screenshots with 7-query Recall@k eval

MarkTechPost · Aug 4, 2026 · 2 min read

Most retrieval pipelines treat every web page as a bag of parsed text, but that assumption breaks the moment you encoun...

Baidu's Unlimited-OCR model parses entire pages in one shot, skipping the layout analysis stage — Inblix summary
AI News

Baidu's Unlimited-OCR model parses entire pages in one shot, skipping the layout analysis stage

MarkTechPost · Jul 24, 2026 · 2 min read

Baidu has released a new vision-language model called Unlimited-OCR that takes a fundamentally different approach to do...

PaddleOCR v6 drops: tiny models hit 86.2% detection, run on CPU — Inblix summary
Research

PaddleOCR v6 drops: tiny models hit 86.2% detection, run on CPU

Hugging Face Blog · Jun 22, 2026 · 2 min read

PaddleOCR just shipped its sixth-generation model family, and the headline isn't just about accuracy—it's about where t...

PaddleOCR 3.5 now runs on a Transformers backend, cutting a major friction point for RAG devs — Inblix summary
Research

PaddleOCR 3.5 now runs on a Transformers backend, cutting a major friction point for RAG devs

Hugging Face Blog · May 18, 2026 · 2 min read

The team behind PaddleOCR just shipped version 3.5, and the real news isn't a new model—it's a new inference engine opt...

Albumentations adds text-aware augmentation that rewrites document images without wrecking OCR — Inblix summary
Research

Albumentations adds text-aware augmentation that rewrites document images without wrecking OCR

Hugging Face Blog · Aug 6, 2024 · 2 min read

Fine-tuning vision language models on document images has always been a headache. You need the model to actually read t...