HunyuanOCR
HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better
About
HunyuanOCR-1.5 is a lightweight, end-to-end OCR-specialized vision-language model. It targets a broad range of text-centric visual tasks and unifies document parsing, text spotting, information extraction, text-image translation within a single end-to-end VLM.
Open Source Health
- Stars
- 1,997
- Forks
- 159
- License
- Not stated
- Last commit
- 28 days ago
Resources & Links
Related Categories
Vendor
Tencent-Hunyuan
Open-source projects on GitHub: HunyuanVideo, HunyuanVideo-1.5 and HunyuanOCR
Quick Links
Open Source
More by Tencent-Hunyuan
Related Products
OCRmyPDF
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
Top category match
ZenlessZoneZero-Auto
绝区零 | ZenlessZoneZero | 零号空洞 | 自动战斗 | 自动化 | 图片分类 | OCR识别
Top category match
PaddleOCR2Pytorch
PaddleOCR inference in PyTorch. Converted from [PaddleOCR](https://github.com/PaddlePaddle/PaddleOCR)
Top category match
MinerU
A practical document parsing tool for converting PDF, images, DOCX, PPTX, and XLSX into Markdown and JSON
Top category match
Deepseek OCR App
A quick vibe coded app for deepseek OCR
Top category match