正在加载视频...
视频加载失败
Check out our demos using LFM2.5-VL-3B, our latest lightweight, vision-language model that reads screens, documents, and the physical world. First up: LFM2.5-VL-3B running fully on-device in the browser with WebGPU to understand a document page. The model parses the entire layout in one pass and returns regions and labels... show more
12,613 次观看 • 1 个月前 •via X (Twitter)
0 条评论
暂无评论
原始帖子的评论将显示在这里
