TeleOCR (ex NaviDC-OCR)
Released Aug 17, 2026 as NaviDC-OCR (StarDoc-AI), renamed TeleOCR on Sept 10. ~1,280 HF likes and ~41k downloads by Oct 6, 2026; GGUF/llama.cpp ports exist. Benchmarks are self-reported. The README links TeleAI's (China Telecom) document-parsing web app; the HF org is XingChen-AGI.
- Input
- image, pdf
- Output
- text
- License
- apache-2.0
- Verified
- 2026-10-06
How to call it
| Provider | Model id | Endpoint / URL | Docs |
|---|---|---|---|
| Hugging Face | — | huggingface.co/XingChen-AGI/TeleOCR | — |
| GitHub | — | github.com/caipeng328/TeleOCR | — |
| Web app | — | www.teleai.com.cn/docparse/DocumentParsing | — |
Notable capabilities (2)
- Camera-captured document parsing without dewarping: One ~1.4B-parameter Qwen2.5-VL-based model parses both digital and photographed or curved documents directly, with no separate rectification step (DocUNet/DIR300 demos). source
- Dr.DocBench Challenge lead (found after launch): Self-reported 67.96 overall on the EMNLP 2026 Dr.DocBench Challenge vs MinerU 2.5 Pro 62.26 and PaddleOCR-VL 1.6 55.11. source
Lightweight open document-parsing VLM. Technical report: arXiv 2608.12898.