Register and share your invite link to earn from video plays and referrals.

ModelScope
@ModelScope2022
Driving innovations with open communities. đŸ’Ŧ Join our Discord:
Joined April 2024
183 Following    16K Followers
Digital PDFs or warped phone photos, TeleOCR parses them with one lightweight 1.2B vision-language model. 📜 Apache 2.0 License. 🤖 📃 🏆 Scores 96.87 overall on OmniDocBench v1.6, the highest among the listed specialized VLMs, and ranks #1# in the ICDAR 2026 Sci-ImageMiner Challenge. 📷 Handles digital, photographed, curved, and degraded documents directly, without a separate dewarping model. 🧠 Combines geometry-aware synthesis, consensus-generated labels, image-based self-verification, and progressive training from vision-language alignment to reinforcement learning. ⚡ Supports structured parsing of text, tables, formulas, layouts, and reading order, with synchronous or asynchronous vLLM inference.
Show more