arXiv cs.CV
· Papers
HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better
arXiv:2607.04884v2 Announce Type: replace Abstract: We present HunyuanOCR-1.5, a lightweight end-to-end OCR-specialized vision-language model. HunyuanOCR unifies document parsing, text spotting, information extraction, text-image translation, and multi-image document understanding within a single end-to-end VLM. Buildi