Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Validating your backend infrastructure requires traffic that mimics user behavior, but writing custom load-testing scripts that handle concurrency and specific user journeys is time consuming. See how these tasks are well suited for Gemini 3 Flash →

18,097 Aufrufe • vor 7 Monaten •via X (Twitter)

0 Kommentare

Keine Kommentare verfügbar

Kommentare vom Original-Post werden hier angezeigt

Ähnliche Videos

"The future of AI is agentic. That includes browsers!" Imagine having an AI agent in your browser that can help you complete complex tasks, answer your questions, and streamline your workflow. Today I'm thrilled to share a sneak peek at Project Mariner, a cutting-edge research collaboration between Chrome and Google DeepMind, exploring the future of agentic AI within the browser! Building on the power of Gemini 2.0, Mariner envisions AI agents seamlessly guiding users through online tasks, streamlining workflows and enriching browsing experiences. Imagine having an intelligent co-pilot in your browser, anticipating your needs and proactively offering assistance. We're in the early stages of experimentation, focusing on core functionalities like understanding user intent, automating actions, and providing personalized recommendations. This prototype leverages Gemini's advanced natural language understanding and reasoning capabilities to interpret user requests, both typed and spoken. Mariner can then interact with web pages, retrieve information, and even perform actions like filling out forms or navigating to specific sites. For example, a user could simply ask "Find me a job near me," and Mariner would understand the request, navigate to a relevant job search site, and tailor the search based on the user's location and preferences. This is just one example of how we're exploring Gemini 2.0's potential to unlock agentic experiences through a series of prototypes, including: 1. Agents with multimodal reasoning: Project Astra, our research prototype exploring the capabilities of a universal AI assistant, is enhanced by Gemini 2.0. 2. Agents that can help you accomplish complex tasks: Project Mariner itself focuses on the future of human-agent interaction within the browser. 3. Agents for developers: Jules is an experimental AI-powered coding agent that integrates directly into a GitHub workflow. 4. Agents applied across domains: We're exploring agents for navigating video games and even applying Gemini 2.0's spatial reasoning to robotics. We believe that integrating AI agents directly into the browser has the potential to revolutionize how we interact with the web. Project Mariner aims to make browsing more intuitive, efficient, and personalized. By understanding user context and proactively offering assistance, Mariner can simplify complex tasks, save users time, and empower them to achieve more online. This aligns perfectly with the vision of Gemini 2.0 to create more helpful and intuitive AI experiences. We’re currently testing Mariner with a small group of trusted users to gather feedback and refine the user experience. We believe that this technology holds immense potential to transform the way we browse and interact with information online.

Addy Osmani

29,518 Aufrufe • vor 1 Jahr

NORAH O'DONNELL: So why do we need to test our nuclear weapons? US PRESIDENT TRUMP: Well, because you have to see how they work, you know. You do have to, and the reason I'm saying testing is because Russia announced that they're gonna be doing a test. If you noticed, North Korea is testing constantly, other countries are testing. We're the only country that doesn't test, and I don't wanna be the only country that doesn't test. NORAH O'DONNELL: Are you saying that after more than 30 years, the United States is going to start detonating nuclear weapons? US PRESIDENT TRUMP: I'm saying that we're going to test nuclear weapons like other countries do, yes. NORAH O'DONNELL: But the only country that is testing nuclear weapons is North Korea. China and Russia are not... US PRESIDENT TRUMP: No, no. Russia is testing nuclear weapons. And China is testing them too. You just don't know about it. NORAH O'DONNELL: That would be certainly newsworthy. My understanding is that what Russia did recently was test essentially the delivery system for nuclear weapons. Essentially - missiles. Which we can do that, but not with nuclear warheads. US PRESIDENT TRUMP: Russia is testing and China is testing, but they don't talk about it, you know. We're an open society, we're different. We talk about it, we have to talk about it. Because otherwise you people gonna report. They don't have reporters that gonna be writing about it, we do.

Status-6 (Military & Conflict News)

19,263 Aufrufe • vor 10 Monaten

Learning from Human Demonstrations: Show the Robot How to Act! The pipeline is very similar to older experiments using Gemini & pi0 with LeRobot. Pi-zero runs locally, while Gemini Flash generates the affordances and the high-level task. (More details are in the thread.) The new component is learning from demonstrations via Gemini 2.5 Pro. I capture a video while demoing & take one of the last frames. Gemini 2.5 Pro then extracts the instructions & passes them to Gemini Flash to process the scene. The fun part is that there's no fancy insight that came from me; other than the days spent figuring out the right prompts. It's the bitter lesson hitting you in the face -> Enhanced Gemini capabilities make this possible. For example, Gemini Flash cannot do Russian doll stacking, but Gemini 2.5 Pro can do it consistently. The current limitation is low-level manipulation: - As you can see, I'm aligning the objects so they are easy to grasp using the same technique from the training data. I couldn't get Gemini Flash to consistently output an accurate grasping angle, and Gemini 1.5 Pro was too expensive and slow for real-time deployment. - Getting a symmetrical gripper should also help a lot. Adding rubber to the tips would probably also help prevent objects from slipping. Collecting & curating the data was the most time consuming & labor intensive part. Next, to improve low-level manipulation and make the system more real-time, I'm shifting to focus more on sims & synthetic data. This aligns better with my core competence. I'm open to tips and suggestions.

Shreyas Gite

22,555 Aufrufe • vor 1 Jahr

Introducing Kiro, an all-new agentic IDE that has a chance to transform how developers build software. Let me highlight three key innovations that make Kiro special: 1 - Kiro introduces spec-driven development, helping developers express their intent clearly through natural language specifications and architecture diagrams for complex features. This comprehensive context helps Kiro’s AI agents deliver better results with fewer iterations. 2 - Kiro features intelligent agent hooks that automatically handle critical but time-consuming tasks like generating documentation, writing tests, and optimizing performance. These hooks work in the background, triggered by events like saving files or making commits. It’s like having an experienced developer constantly reviewing your work and handling the maintenance tasks that often get delayed. 3 - Kiro provides a purpose-built interface that adapts to how developers work. Whether you prefer chat interactions or working with specifications, Kiro supports your workflow while keeping you in control of the development process. Kiro is really good at "vibe coding" but goes well beyond that. While other AI coding assistants might help you prototype quickly, Kiro helps you take those prototypes all the way to production by following a mature, structured development process out of the box. This means developers can spend less time on boilerplate code and more time where it matters most – innovating and building solutions that customers will love. Starting today, Kiro is available for free during preview and supports most popular programming languages. Here’s how to get started with Kiro today: Excited to see how developers use Kiro, and to work with the developer community to continue to shape Kiro moving forward.

Andy Jassy

668,263 Aufrufe • vor 1 Jahr

Kled Version 3 is coming. Over $20M+ in rewards will be paid directly to users from leading AI labs across robotics, legal services, image and video generation, world modeling, and more. In the last seven days, we’ve received inbound data requests from several decacorn AI labs and enterprises for datasets our human data marketplace is uniquely positioned to provide. Since receiving the specs for these requests, we now have a much better picture and understanding of how to reshape the systems that collect this data, so here’s what’s coming: 1. A fully redesigned home experience: The home feed is being rebuilt to surface the highest-value, most relevant tasks for each user, similar to how Uber Eats surfaces top restaurants. The goal is to turn every user into their most effective version as a data contributor. 2. Automated quality enforcement at scale: New ML systems are being built to evaluate task-specific requirements in real time. For example, if a task requires “two hands visible on camera at all times,” any video that fails that spec will be automatically rejected. This logic will apply across thousands of tasks and specifications using a general ML. 3. Kled Shop: Some tasks require better capture hardware. We’re introducing Kled Shop, where users can redeem points or tokens for equipment like Meta glasses, drones, and other tools. Points and tokens can be converted directly from payouts. 4. Partner-run data labeling and evaluation work: Some of our partners operate high-paying data labeling and model evaluation programs. We’re integrating their workflows directly into Kled so qualified users can access these roles in one place. These jobs are owned and managed by our partners. Kled’s role is to route the right people to the right work. Some opportunities pay $50–$1,000 per hour depending on expertise. 5. Global payouts and localization: We’re partnering with a major payment processor to enable cashouts in users’ native currencies. This unlocks broader global participation. Multi-language support is also coming to accelerate user growth. This full suite of tools will be rolling out soon, directly to Kled users. Top earners are currently making ~$7,000 per month. With this update, we should see the first ~$10,000 per month earner.

Avi Patel

124,728 Aufrufe • vor 7 Monaten