Digital PDFs or warped phone photos, TeleOCR parses them with one lightweight 1.2B…
This is a AI post classified by Jev as RAG & memory (a model release), kept by the AI Radar because it carries real work, not commentary.
Digital PDFs or warped phone photos, TeleOCR parses them with one lightweight 1.2B vision-language model. 📜 Apache 2.0 License. 🤖 https://modelscope.ai/models/XingChen-AGI/TeleOCR 📃 https://modelscope.ai/papers/2608.12898 🏆 Scores 96.87 overall on OmniDocBench v1.6, the highest among the listed specialized VLMs, and ranks #1 in the ICDAR 2026 Sci-ImageMiner Challenge. 📷 Handles digital, photographed, curved, and degraded documents directly, without a separate dewarping model. 🧠 Combines geometry-aware synthesis, consensus-generated labels, image-based self-verification, and progressive traini
Posted by ModelScope (15.9k followers) 2 h ago · 34 likes · 1.8k views · view the original post on X. Kept by the AI Radar as RAG & memory. Tools mentioned: Ling-3.0-flash-Fin.
More AI work like this
- Interesting paper on agent memory stored as a linked markdown wiki. — @omarsar0
- D-RAC: Universal Retrieval-Aware Ingestion — @HuggingPapers
- liteparse is the fastest pdf parser on the planet — @jerryjliu0
- DocJev is the fastest way to classify and split complex document packets ⚡️. (and the… — @jerryjliu0
- Congrats to @CalebPeffer, @ericciarla, @nickscamara_, and @firecrawl on their $75M… — @ycombinator
- The fastest PDF-to-Markdown parser just got faster. ⚡️ — @llama_index
- Today we are releasing DolphinBench: Mapping the Pareto frontier of agent memory. — @mem0ai
- Keenable is now a web-fetch backend in @cognee_ — @KeenableAI
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 44.3k posts from 5k X accounts over the last 14 days, 1.8k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-23 12:57 UTC. Full method.