Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Introducing Lucy: 1.7B model that Google for you It's an agentic‑search model that can even run on your phone. - Agentic search on tap - Lucy calls tools ( ‑aware) - Fits in your pocket - runs on CPU or mobile Under the hood: - Built on Qwen's Qwen3‑1.7B...

20,205 görüntüleme • 1 yıl önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

J-Cal Explains Why Google is UNDERRATED in AI 👀 On E227, the besties discussed Google's value in a post-search world if AI replaces traditional search. @jason broke down why he thinks Google is being slept on: "I think there's a chance that we're underestimating the power of Google's ad network right now." "They have four or five products that are one or two billion users per month. You have YouTube, Google Docs, Android." "They have such a data advantage and such a deep integration into people's lives because they use three or four services, I think Google's gonna figure this out." "It's quite possible that knowing your queries in Gemini, knowing what you're doing in Calendar, knowing what you're watching on YouTube could lead to a stream of more targeted ads that do better and are more valuable." "We've been seeing a number of startups that are figuring out how to use your queries and what you're doing in AI to present to you search results." "So imagine you're doing a Gemini search and on the side of it, it's giving you a rolling list of ads or offers that you might be more interested in." "That could be a better advertising product than even search itself." "I think YouTube search is the place to go all-in." "Right now, when you do a YouTube search, it just gives you 10 links, right? It just gives you that rolling thing." "You should be able to ask a question to YouTube, and you should be able to ask questions to your calendar." "You should be able to say, who have I met with over the last 10 years? Who I'm no longer in touch with and what are they up to?" "And it should do a Gemini search inside of Google Calendar. It's very light right now." "And then if you did that on YouTube, this would train people at the point of pain in a very deep way without sacrificing Google Search queries too aggressively."

The All-In Podcast

58,275 görüntüleme • 1 yıl önce

Introducing Workshop: cloud + on-device agentic AI. And to celebrate, we're giving away $250k in Google Gemini AI credits. (details below). The future of AI work is neither cloud-based nor local. It's both. In Workshop Cloud, you can use agents powered by frontier models like Claude and/or open source models like Z.ai's GLM-5 to build internal tools, dashboards, and AI web apps. Or, breeze through tasks like managing your Google and Meta Ads. In Workshop Desktop, you can do all the same right on your computer, plus make desktop apps, mobile apps, and 3D creations. Our favorite part? You can power the full agent experience with local models like Qwen 3.5 family on your computer. Fully offline. 2026 is the year in which local models for agentic tasks will become viable for mainstream use. But the setup for tools like OpenClaw is like setting up Linux from scratch on your computer. Workshop Desktop is one-click to install on Windows, Mac, and Linux. It recommends which open source model you should use for your hardware and lets you download and run it right in the app. And its agent harness allows you to chat, create websites, build personal utilities, and analyze data. 100% offline. Or multitask with AI models in the cloud while running other agent threads locally. Start in Workshop Cloud when you want flexibility and speed. Download your project and continue in Workshop Desktop when you want local files, privacy, and/or better performance on large code bases. Publish from either. The agent tooling space is maturing and discerning users have come to expect a lot from their tools. We've packed Workshop with features to help you 10x your productivity. - Native support for skills - Autocompaction for seamless context management - Built-in AI for your apps - Dozens of connectors, like Google Drive, Big Query, and Supabase - dbt integration to ground your dashboards in your semantic layer - Native Github integration - Private app deployment - ... and more (+ we're shipping super fast) To access the free credit offer, RT this post and reply with "Workshop". Make sure you are following us so we can DM you the instructions to redeem. - First 100 to RT + comment get $500 in credits. - Everyone else gets up to $250 And thanks to our partners Modal, Google Gemini, and Z.ai!

Workshop AI

28,745 görüntüleme • 4 ay önce

We’re launching Optima. Now anyone can create a custom benchmark for their use case, leveraging Artificial Analysis’ leading research and platform Building and running benchmarks is difficult. We have distilled Artificial Analysis’ research and experience developing benchmarks into Optima, a new platform for benchmarking models on your own workloads and comparing performance, speed and cost efficiency. Optima allows you to find the best model for your task, or an equally performant alternative to your current setup at 10x lower cost or time per task. We’ve integrated Artificial Analysis' research and experience in benchmarks across the Optima workflow: ➤ Build benchmarks based on your own data and use cases: There are three ways to build a benchmark with Optima. Upload an existing evaluation dataset from your own files or Hugging Face, or import agent traces from platforms including Arize AI, Braintrust and langfuse.com. Install the Optima skill to build a benchmark using context from your coding environment and previous sessions. Or simply describe your use case and provide example inputs and outputs, and Optima will build the benchmark for you ➤ Run across the latest models: Run the same benchmark across leading models in a single click, and keep your leaderboard up to date as soon as new models are released ➤ Bring Artificial Analysis grading to your own benchmark: Evaluate responses against objective rubric criteria or using the same pairwise judging approach used for Artificial Analysis benchmarks including GDPval-AA and AA-Briefcase. For pairwise judging, select your preferred responses from a sample and Optima uses those preferences to rank models across your test set ➤ Compare performance, cost and time efficiency: Optima measures more than model performance. Cost per Task and Time per Task are tracked alongside benchmark scores, with category-level results and support for custom metrics, allowing you to compare the tradeoffs between models for your specific use case Ahead of launch, here are examples questions our beta testers answered with Optima: ➤ Which model can save me 10x the cost without a meaningful decrease in quality for my finance & accounting agent? ➤ Which model best matches the writing style of lawyers for my legal agent? ➤ Which model can best identify different elements in my custom image dataset? Optima is available today. Build your own benchmark at

Artificial Analysis

106,425 görüntüleme • 1 gün önce