TECH NEWS
TeleOCR: A 1.2B Vision-Language Model for Structured Document Parsing
TeleOCR is an open-source, approximately 1.2-billion-parameter vision-language model for parsing both born-digital documents and camera-captured pages. It is maintained by XingChen-AGI, uses the Transformers library, and is loaded through AutoProcessor and AutoModel. Its central distinction is that it targets geometric distortion and structured content—especially tables and formulas—as