正在加载视频...

视频加载失败

Florence-2, the new vision foundation model by Microsoft, can now run 100% locally in your browser on WebGPU, thanks to Transformers.js! 🤗🤯 It supports tasks like image captioning, optical character recognition, object detection, and many more! 😍 WOW! Demo (+ source code) 👇

88,762 次观看 • 2 年前 •via X (Twitter)

9 条评论

Xenova 的头像
Xenova2 年前

ONNX models: Source code: Demo:

nickmystic 的头像
nickmystic2 年前

amazing work!

Samuel Tallet 的头像
Samuel Tallet2 年前

It's awesome, thank you Xenova! Does the "Florence-2-large" model can also run on the client-side with WebGPU acceleration?

snats 的头像
snats2 年前

When is v3 coming out? I love this!

Aiflowly.com 的头像
Aiflowly.com2 年前

Amazing! We should add this to the list of our integrations.

Aaron Planell 的头像
Aaron Planell2 年前

@daviddincognit Maybe this can be interesting for you

PDS_B2BMGMT 的头像
PDS_B2BMGMT2 年前

I tried twice! I only got this

Thomas Hill 的头像
Thomas Hill2 年前

🔥

Gather Grove 的头像
Gather Grove2 年前

This actually works great even on my lame GPU, but it's accuracy is kinda random. The next version will be awesome.

相关视频