Franken OCR
Pure-Rust, CPU-hyper-optimized runner for the Baidu Unlimited-OCR model (single-binary CLI: focr)
About
A pure-Rust, memory-safe, CPU-only OCR engine for a small family of hand-ported vision-language models. Baidu Unlimited-OCR is the fast default for document OCR, GOT-OCR2 handles specialized structured formats, SmolVLM2 handles image description and VQA, OneChart extracts chart data, and Polyphonic-TrOMR turns full scanned sheet-music pages or staff crops into MusicXML through --task music. All five runtime models are available through focr pull; TrOMR publishes both the 61 MB int8 default artifact and an 86 MB f32 reference artifact. The v0.7.0 Unlimited-OCR artifact uses the conservative exact recipe and passes the hard-page termination and complete 20-page corpus budget. The models run through model-specific Rust kernels and need no general ML framework, Python, CUDA, FFI at inference, or GPU.
Open Source Health
- Stars
- 320
- Forks
- 38
- License
- Not stated
- Last commit
- 26 days ago
Resources & Links
Related Categories
Vendor
Jeff Emanuel
Building in NY
Quick Links
Open Source
More by Jeff Emanuel
Related Products
OCRmyPDF
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
Top category match
ZenlessZoneZero-Auto
绝区零 | ZenlessZoneZero | 零号空洞 | 自动战斗 | 自动化 | 图片分类 | OCR识别
Top category match
PaddleOCR2Pytorch
PaddleOCR inference in PyTorch. Converted from [PaddleOCR](https://github.com/PaddlePaddle/PaddleOCR)
Top category match
MinerU
A practical document parsing tool for converting PDF, images, DOCX, PPTX, and XLSX into Markdown and JSON
Top category match
Deepseek OCR App
A quick vibe coded app for deepseek OCR
Top category match