Loading video...
Video Failed to Load
A peanut-sized Chinese model just dethroned Gemini at reading documents. GLM-OCR is a 0.9B parameter vision-language model. It scores 94.62 on OmniDocBench V1.5, ranking #1 overall. For context, it outperforms models 100x its size. 100% open-source. It works in two stages. 1. A layout engine detects every region in... show more
13,630 views • 2 months ago •via X (Twitter)
0 Comments
No comments available
Comments from the original post will appear here
