Dev Radar
Support
LiveUpdated 2026-09-23 12:49 UTC

Digital PDFs or warped phone photos, TeleOCR parses them with one lightweight 1.2B…

Digital PDFs or warped phone photos, TeleOCR parses them with one lightweight 1.2B vision-language model. 📜 Apache…

This is a dev post classified by Jev as Other (a tool drop), kept by the Dev Radar because it carries real work, not commentary.

Digital PDFs or warped phone photos, TeleOCR parses them with one lightweight 1.2B vision-language model. 📜 Apache 2.0 License. 🤖 https://modelscope.ai/models/XingChen-AGI/TeleOCR 📃 https://modelscope.ai/papers/2608.12898 🏆 Scores 96.87 overall on OmniDocBench v1.6, the highest among the listed specialized VLMs, and ranks #1 in the ICDAR 2026 Sci-ImageMiner Challenge. 📷 Handles digital, photographed, curved, and degraded documents directly, without a separate dewarping model. 🧠 Combines geometry-aware synthesis, consensus-generated labels, image-based self-verification, and progressive traini

Posted by ModelScope (15.9k followers) 2 h ago · 20 likes · 933 views · view the original post on X. Kept by the Dev Radar as Other. Tools mentioned: Ling-3.0-flash-Fin.

More dev work like this

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 22.4k posts from 5k X accounts over the last 21 days, 2.6k tools, 12 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-23 12:49 UTC. Full method.