Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

WF's Model Comparison is a game-changing tool optimized for web AND mobile use! 📱⚡ 🔍 View all 26 weather models side by side 📈 Spot trends & forecast shifts in seconds ⛈️ Track the latest model runs anytime, anywhere 👀 What if we told you this was recorded on...

66,399 Aufrufe • vor 1 Monat •via X (Twitter)

0 Kommentare

Keine Kommentare verfügbar

Kommentare vom Original-Post werden hier angezeigt

Ähnliche Videos

Small Language Models (SML) are the future of AI. "Small" (SML) instead of "Large" (LLM). These small models are highly specialized models with superhuman abilities on specific tasks. Here are two techniques to build these models: • Spectrum • Model Merging I give you a short introduction in the attached video, but here is a quick summary: Spectrum helps us identify the most relevant layers to solve one specific task. We can ignore everything else and focus on fine-tuning these layers. Using Spectrum, we can fine-tune models in a heartbeat. Model Merging combines multiple models into a unique, much better model than any of the individual input models. You can also combine models specialized in different tasks and get a model with multiple abilities. This is the state of the art of productizing models. It's what Arcee.ai's platform does behind the scenes. Arcee collaborated with me on this post and is sponsoring it. There are three main steps to produce a model for your particular use case: 1. You create a dataset by uploading your data. 2. You train a model. At this step, Arcee uses Spectrum and Model Merging to produce a highly specialized model for your task. 3. You can deploy that model to any environment you want. Three important notes: • Training process is 2x faster and 2x cheaper than regular fine-tuning. • Resultant models are smaller and have higher accuracy. • They create these specialized models from open-source models. Check this site so you can fully appreciate how this works: If you want to fine-tune an open-source model, consider Arcee's platform. This is the state of the art.

Santiago

164,162 Aufrufe • vor 2 Jahren

Introducing PhoneLLM, an open model for voice agents. GPT 5.6 Terra performance on typical voice agent tasks at 1/3 the latency and 1/18 the cost. For voice agents, we need models that are both very low latency and very good at tool calling and instruction following. There's a trade-off here, and we often have to compromise on either latency or capability when building voice agents. With PhoneLLM (and the training and data stack that made this model possible) we're fixing this problem. For the last couple of years, most of the effort in frontier model development has gone towards leveraging test-time compute. Which is awesome! Models of all shapes and sizes are available that perform really, really well ... if you have "thinking" turned on for your model. But if you need your agent to respond at voice conversation speed, you can't use thinking models. PhoneLLM is a full-weights fine-tune of NVIDIA Nemotron Nano 30B. We trained on a wide range of real-world telephone and customer support use cases. The training focused on taking the excellent Nano 30B base capabilities and teaching the model to do typical voice agent tasks with thinking disabled. The results are really good: accurate tool calling and concise, on-topic responses in long conversations. And fast: TTFAT measured server-side is <100ms if you run PhoneLLM on a lightly loaded B200. :-) But seriously, when we characterize model latency, we do it with full, end-to-end, batched request simulations using real Pipecat voice agent pipelines. You can serve more than 80 concurrent agents on a single B200 with P95 end-to-end TTFAT <600ms. Including network overhead. That's an LLM cost-per-minute around $0.0025. (1/4 of a cent.) At a latency lower than any third-party API offers today. More details about this model, including weights on Hugging Face, how to spin it up with one click on Modal, and a starter project repo you can clone, are in the thread ...

kwindla

325,296 Aufrufe • vor 10 Tagen

🎉🔵ONE MONTH of WeatherFront🔴🎉 Thanks to everyone who has helped make WF a success since our launch on May 4th! What a fun journey it has been. To operational meteorologists - we hope WF has made it easier to interrogate models, construct forecasts, and quickly find the weather data you need as you serve the public and your partners/customers. To emergency managers - we hope WF has enhanced your ability to quickly respond to hazardous weather with more detailed basemaps to pinpoint affected locations and a wealth of weather data catered to your mobile device. To broadcast meteorologists - we hope WF has provided high-quality data visualization for use on social media and given you a mobile resource for radar, satellite, and model data when you’re doing live coverage on air during severe weather. To weather hobbyists - we hope WF has helped you learn more about different weather data types and equipped you with an expanded set of tools to track storms as you build a deeper passion for meteorology and stay informed about what is headed your way. To storm chasers and spotters - we hope WF has increased your situational awareness on the road or while watching from home - thanks for taking WF to all corners of the US! We’ve enjoyed seeing your content and congratulate you on a great spring season so far. To our international customers - we hope WF has provided a unique perspective to view global weather model data. We appreciate your support and look forward to bringing you additional data sources in the future. To all current and future WF users - thanks for joining us on the journey! We’re just getting started. We hope you’ll invite your friends and colleagues to give WeatherFront a try as we continue to innovate and bring even more products to your mobile devices!

WeatherFront

33,176 Aufrufe • vor 1 Jahr