Unlimited OCR is a good example of a low-parameter OCR model that can process any document extremely quickly.
It's good at reading order and layout, and decent over tables. It is slightly better than PaddleOCR on tables and slightly worse on layout.
There are tradeoffs though - it ignores visual elements, formatting, and the most complex tables. It scores ~46.2% and ranks #115 on ParseBench.
There's plenty of opportunities to improve this frontier, but Unlimited OCR is still a good milestone!
Check out ParseBench: https://www.parsebench.ai/
