Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Okay this is kinda wild 👀 NVIDIA is basically handing out FREE API keys to 100+ top AI models - GLM 5.2, DeepSeek V4, Kimi K2.6, MiniMax M3, their own Nemotron and a ton more. it's called NVIDIA NIM and I had no idea it existed. it's rate-limited, so...

26,272 görüntüleme • 1 ay önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

Big win for open-source LLMs! DeepSeek V4 Pro holds the top open-weights score on SWE-bench Verified, in the GPT-5.5 range. GLM 5.2 leads the open-weight intelligence index and sits near the closed frontier on long-horizon coding. But this leaderboard number is a weak proxy for real performance. It comes from one task set, run through one harness, served at one precision. The same weights can even score differently across providers, since many hosts quantize activations to fp8 and drift the model off its reference weights. Real performance is determined based on whether a model can read a repo, make coordinated edits across files, run the tests, and recover when one breaks. By that measure, the top open models hold up, but only inside the right harness. The teams that actually put DeepSeek V4 into production pipelines as a frontier substitute got there through the harness they built around the model, not by picking a stronger model. If you want to see this in practice, Cline (64k+ stars) has actually built that harness around open models, tuned so they run at production quality. And it's tuned so that these LLMs can run at production quality, with plan and act modes, checkpoints, and terminal feedback. ClinePass is the new access layer on top of it. It runs a curated set of those models inside Cline, narrowed to the ones tested for coding-agent use, with 2 to 5x the standard rate limits and no separate provider accounts, keys, or billing to track. The video below shows the setup, and I worked with the team to put this together. It runs alongside custom keys and local models as well, not in place of them.

Avi Chawla

44,124 görüntüleme • 1 ay önce

New open-source agent harness just landed! I got early access to TrueForge by TrueFoundry and have been running it locally for the past few days. The harness layer deserves as much attention as the model, and open source matters here because you can inspect the loop, run it on your own infrastructure, and swap to the latest or cheaper models. TrueForge handles the runtime work that makes an agent reliable. It drives the tool-calling loop, manages context, coordinates subagents, and executes code in a sandbox, with any model you choose. Every tool call re-sends the growing context to the model, so in practice the harness controls most of what an agent costs to run. A few things stood out from my testing and their published benchmarks. Vendor-Neutral by design. It runs OpenAI, Anthropic, and Google models alongside open-weight models like Kimi, GLM, and DeepSeek. Model routing is a setting, and you can send each task to the model that fits it. On a 14-task enterprise agent benchmark, it matched the accuracy of Claude Managed Agents running the same Opus 4.8 model at roughly 30% lower cost per run (3.8M tokens vs 10M for the same answers). Routing the same tasks to GLM-5.2 held accuracy and brought cost down by about 75%, around $3 per run instead of $12. Fully self-hosted and Open Source (MIT License). I had it running locally with one command, with sandboxed code execution working out of the box. It's time to own your agent harness. Thanks to TrueFoundry for partnering on this post.

elvis

11,303 görüntüleme • 4 gün önce

✨ Grok Imagine Video is now live on Photo AI It's hard to explain how impressive this is because of the speed that xAI got itself from literally nothing to the top of the leaderboards Six months ago Grok's video model was a joke, it wasn't even close to any of the video models out there, it looked cartoony and wasn't there and nobody took it seriously Now it's here and it's instantly the #1 video model out there now, it shot above Kling (which I used before on Photo AI and usually my favorite) and above Runway Gen 4.5 which was just launched 6 days ago! Mmore importantly it's now above xAI's biggest competitors' models: Google's Veo 3 and OpenAI's Sora 2 Being the best video model doesn't mean it's flawless: video is incredibly hard and actually because it looks so realistic now when it does make a mistakes it's even funnier One thing I noticed is that it still has a hard time with is voice, it does it well for a majority of the video but then slips up and produces unintelligible blabbering (which is really funny to hear) in both English (video 1: "it's where I find my naim", what's a "naim"?), and tested it in Portuguese too (video 3 at the end is unintelligible Portuguese I believe) In many ways Grok Imagine Video also reminds me of Sora, it has that weird but funny Sora conversation style But guys it's REALLY really really really close to getting perfect, we're so close to having full video productions being to be able done in AI, actually you already can if you just cut out the bad parts already Very exciting and I'm grateful I can experience this

@levelsio

195,023 görüntüleme • 6 ay önce