正在加载视频...

视频加载失败

This is HUGE, Hugging Face just shipped Inference Providers on the Hub Partnering with Together AI, @FAL, Replicate and SambaNova! Starting today you can access thousands of Models like DeepSeek R1, Llama, Flux, Whisper, and more Directly from Hugging Face!

177,126 次观看 • 1 年前 •via X (Twitter)

11 条评论

AK 的头像
AK1 年前

blog:

APILayer 的头像
APILayer1 年前

🔓 Unlock new possibilities for your business with APIs! 🚀 Discover our extensive API Marketplace, featuring top-notch, cutting-edge solutions tailored just for you. Elevate your company and drive innovation! Get started with apilayer today:

TheHeroShep 的头像
TheHeroShep1 年前

@togethercompute @FAL @replicate @SambaNovaAI We love when an epic team comes together… Now who’s Hawkeye? 🤣

Sourabh 的头像
Sourabh1 年前

@togethercompute @FAL @replicate @SambaNovaAI ai game just changed fr

Cédric Limousin 的头像
Cédric Limousin1 年前

@togethercompute @FAL @replicate @SambaNovaAI So like a competitor to Fal and Replicate ?

Rahul Asthana 的头像
Rahul Asthana1 年前

@togethercompute @FAL @replicate @SambaNovaAI so cool

Amoussouvi 的头像
Amoussouvi1 年前

@togethercompute @FAL @replicate @SambaNovaAI Is it free or do you need a pro account ?

Victor M 的头像
Victor M1 年前

This is wild 🤯 We finally have an open-source music generation model that rocks 🚀

Unsloth AI 的头像
Unsloth AI1 年前

Introducing 1.58bit DeepSeek-R1 GGUFs! 🐋 DeepSeek-R1 can now run in 1.58-bit, while being fully functional. We shrank the 671B parameter model from 720GB to just 131GB - a 80% size reduction. Naively quantizing all layers breaks the model entirely, causing endless loops & gibberish outputs. Our dynamic quants solve this. The 1.58-bit quant fits in 160GB VRAM (2x H100 80GB) for fast inference at ~140 tokens/sec. By studying DeepSeek-R1’s architecture, we selectively quantized certain layers to higher bits (like 4-bit), and leave most MoE layers to 1.5-bit. Benchmarks + Blog: Dynamic GGUFs (131GB–212GB) on Hugging Face:

Qwen 的头像
Qwen1 年前

🎉 恭喜发财🧧🐍 As we welcome the Chinese New Year, we're thrilled to announce the launch of Qwen2.5-VL , our latest flagship vision-language model! 🚀 💗 Qwen Chat: 📖 Blog: 🤗 Hugging Face: 🤖 ModelScope: 🌟 Key Highlights: * Visual Understanding : From flowers to complex charts, Qwen2.5-VL sees it all! * Agentic Capabilities : It’s a visual agent that can reason and interact with tools like computers & phones. * Long Video Comprehension : Captures events in videos over 1 hour long! ⏳🎥 * Precise Localization : Generates bounding boxes & JSON outputs for accurate object detection. * Structured Data Outputs : Perfect for finance & commerce, handling invoices, forms & more! 💼📊 Try Qwen2.5-VL now at Qwen Chat or explore models on Hugging Face & ModelScope . 🌐

SkalskiP 的头像
SkalskiP1 年前

Alibaba dropped QWEN2.5VL yesterday; I spend all night working on fine-tuning tutorial this notebook covers - dataset preparation - tokenization - training with LoRA/QLoRA (for max performance on low-power devices) - fine-tuned model evaluation link:

相关视频