Video yükleniyor...
Video Yüklenemedi
Florence-2, the new vision foundation model by Microsoft, can now run 100% locally in your browser on WebGPU, thanks to Transformers.js! 🤗🤯 It supports tasks like image captioning, optical character recognition, object detection, and many more! 😍 WOW! Demo (+ source code) 👇
88,762 görüntüleme • 2 yıl önce •via X (Twitter)
9 Yorum

Xenova2 yıl önce
ONNX models: Source code: Demo:

nickmystic2 yıl önce
amazing work!

Samuel Tallet2 yıl önce
It's awesome, thank you Xenova! Does the "Florence-2-large" model can also run on the client-side with WebGPU acceleration?

snats2 yıl önce
When is v3 coming out? I love this!

Aiflowly.com2 yıl önce
Amazing! We should add this to the list of our integrations.

Aaron Planell2 yıl önce
@daviddincognit Maybe this can be interesting for you

PDS_B2BMGMT2 yıl önce
I tried twice! I only got this

Thomas Hill2 yıl önce
🔥

Gather Grove2 yıl önce
This actually works great even on my lame GPU, but it's accuracy is kinda random. The next version will be awesome.
