Video wird geladen...
Video konnte nicht geladen werden
This is HUGE, Hugging Face just shipped Inference Providers on the Hub Partnering with Together AI, @FAL, Replicate and SambaNova! Starting today you can access thousands of Models like DeepSeek R1, Llama, Flux, Whisper, and more Directly from Hugging Face!
177,126 Aufrufe • vor 1 Jahr •via X (Twitter)
11 Kommentare

blog:

🔓 Unlock new possibilities for your business with APIs! 🚀 Discover our extensive API Marketplace, featuring top-notch, cutting-edge solutions tailored just for you. Elevate your company and drive innovation! Get started with apilayer today:

@togethercompute @FAL @replicate @SambaNovaAI We love when an epic team comes together… Now who’s Hawkeye? 🤣

@togethercompute @FAL @replicate @SambaNovaAI ai game just changed fr

@togethercompute @FAL @replicate @SambaNovaAI So like a competitor to Fal and Replicate ?

@togethercompute @FAL @replicate @SambaNovaAI so cool

@togethercompute @FAL @replicate @SambaNovaAI Is it free or do you need a pro account ?

This is wild 🤯 We finally have an open-source music generation model that rocks 🚀

Introducing 1.58bit DeepSeek-R1 GGUFs! 🐋 DeepSeek-R1 can now run in 1.58-bit, while being fully functional. We shrank the 671B parameter model from 720GB to just 131GB - a 80% size reduction. Naively quantizing all layers breaks the model entirely, causing endless loops & gibberish outputs. Our dynamic quants solve this. The 1.58-bit quant fits in 160GB VRAM (2x H100 80GB) for fast inference at ~140 tokens/sec. By studying DeepSeek-R1’s architecture, we selectively quantized certain layers to higher bits (like 4-bit), and leave most MoE layers to 1.5-bit. Benchmarks + Blog: Dynamic GGUFs (131GB–212GB) on Hugging Face:

🎉 恭喜发财🧧🐍 As we welcome the Chinese New Year, we're thrilled to announce the launch of Qwen2.5-VL , our latest flagship vision-language model! 🚀 💗 Qwen Chat: 📖 Blog: 🤗 Hugging Face: 🤖 ModelScope: 🌟 Key Highlights: * Visual Understanding : From flowers to complex charts, Qwen2.5-VL sees it all! * Agentic Capabilities : It’s a visual agent that can reason and interact with tools like computers & phones. * Long Video Comprehension : Captures events in videos over 1 hour long! ⏳🎥 * Precise Localization : Generates bounding boxes & JSON outputs for accurate object detection. * Structured Data Outputs : Perfect for finance & commerce, handling invoices, forms & more! 💼📊 Try Qwen2.5-VL now at Qwen Chat or explore models on Hugging Face & ModelScope . 🌐

Alibaba dropped QWEN2.5VL yesterday; I spend all night working on fine-tuning tutorial this notebook covers - dataset preparation - tokenization - training with LoRA/QLoRA (for max performance on low-power devices) - fine-tuned model evaluation link:

