China open-sourced a peanut-sized OCR that parses entire 100-page PDFs in one shot.. It's called Unlimited-OCR. Only…
Summary
China released Unlimited-OCR, a 3-billion-parameter open-source optical character recognition model that can process entire 100-page PDFs in a single pass with 32K context window, eliminating the page-by-page limitations of traditional OCR tools. The model achieves 93% accuracy on parsing benchmarks, runs entirely locally for free, and supports multiple languages, positioning it as a free alternative to expensive cloud-based OCR services like AWS Textract and Google Vision.
Summarized by ThreadOut AI from the full thread. May miss nuance — read the thread below.
- #1
- #2