Loading video...

Video Failed to Load

Go Home

The Physical AI data space is exploding, but most vendors still share samples through Google Drive or Hugging Face links that are slow and painful to evaluate. We built the Human Archive Catalog so researchers can quickly browse and filter our off-the-shelf datasets by environment, hardware configuration, modality, annotations,...

12,484 views • 1 month ago •via X (Twitter)

0 Comments

No comments available

Comments from the original post will appear here

Related Videos

🚀 We just raised $40 million to build infrastructure for Physical AI! 🦾 AI is rapidly transforming critical industries like manufacturing, logistics, transportation, agriculture, construction, aerospace, and defense. Teams that win in the physical world are those who can create a data flywheel, leveraging infrastructure to capture, ingest, analyze, and evaluate the vast quantities of data generated by real-world systems. Robotics data is multimodal, time-synchronized, and bandwidth‑constrained at the edge. Traditional data and observability platforms were only designed to store and query text and time-series data, not petabyte-scale 3D, video, audio, GNSS, and proprioceptive data. The ability to efficiently capture, ingest, search, visualize, and evaluate multimodal data is critical to Physical AI development. Foxglove is a modern data engine for Physical AI, enabling you to record logs or capture demonstrations at the edge, sync recordings to the cloud or on-premises storage, find critical events across petabytes of data, evaluate robot performance, and watch a 3D frame-by-frame replay using our advanced visualization tool. 👉 Today is still Day 1 for Physical AI, and we're hiring for dozens of roles to assemble the best team in the industry. If you've built ML platforms, data infrastructure, dataset curation, evaluation and validation, or visualization tools at a leading robotics or autonomous vehicle company, let's chat – drop me a note or tag a friend below and I'll follow up personally! Thank you to Alexandra Sukin and Jeremy Levine at Bessemer, Seth Winterroth 🤖 at Eclipse, David Beyer and Sunil Dhaliwal at Amplify Partners, and Icehouse Ventures for joining us on this mission. Also a special shoutout to our angels tobi lutke Alex Kendall Kyle Vogt Milan Kovac Hussein Mehanna Pieter Abbeel Brad Porter Boris Sofman Kevin Peterson Chris Walti Lindon Gao Daniel Kan Adam Draper ⏻ Fred Ehrsam and Karri Saarinen!

Adrian Macneil

46,639 views • 9 months ago

🚨: HARDWARE FOUNDERS this one is for you! If you’re building: > machines - from CNCs to circuits, from ships (space or sea - doesn't matter) to satellites > robots - arms, quadrupeds, humanoids, semi-humanoids, exoskeletons > drones or UAVs or things that fly > wearables rings or pendants or smart glasses or .... basically anything in hardware, manufacturing, industrial automation, or embodied AI ... we’ve built a Home for you! Whether you’re in Tokyo or Tennessee, Melbourne or Madrid, Singapore or San Francisco, Beijing or Berlin, Shenzhen or Stockholm, Bangalore or Boston, or Lima or London — if you’re building the physical future, you belong here⚒️ MOSAIX is a global network for hardware x AI founders — the first of its kind — designed to help you solve, scale, and sustain your vision. This is born out of years of working with hardware founders across the world and seeing how hard it is to build in hardware alone. Founders like you have a unique journey. You deal with machines, with grease, with soldering irons and sensors, with components and gears. Your challenges don’t look like those in the software world, and that’s exactly why MOSAIX exists. Our team spans China, India, Europe, and the U.S. and so does our purpose. We’re independent. We’re not backed by anyone. We don’t carry anyone’s agenda. Just a shared belief that hardware founders deserve a network of their own — a safe space, and a safety net. MOSAIX is invite-only. For now, joining is free. Welcome to the Age of Hardware. Welcome to the MOSAIX Network. More details in the tweet below.

AML

63,316 views • 10 months ago

Exciting updates on Project GR00T! We discover a systematic way to scale up robot data, tackling the most painful pain point in robotics. The idea is simple: human collects demonstration on a real robot, and we multiply that data 1000x or more in simulation. Let’s break it down: 1. We use Apple Vision Pro (yes!!) to give the human operator first person control of the humanoid. Vision Pro parses human hand pose and retargets the motion to the robot hand, all in real time. From the human’s point of view, they are immersed in another body like the Avatar. Teleoperation is slow and time-consuming, but we can afford to collect a small amount of data. 2. We use RoboCasa, a generative simulation framework, to multiply the demonstration data by varying the visual appearance and layout of the environment. In Jensen’s keynote video below, the humanoid is now placing the cup in hundreds of kitchens with a huge diversity of textures, furniture, and object placement. We only have 1 physical kitchen at the GEAR Lab in NVIDIA HQ, but we can conjure up infinite ones in simulation. 3. Finally, we apply MimicGen, a technique to multiply the above data even more by varying the *motion* of the robot. MimicGen generates vast number of new action trajectories based on the original human data, and filters out failed ones (e.g. those that drop the cup) to form a much larger dataset. To sum up, given 1 human trajectory with Vision Pro -> RoboCasa produces N (varying visuals) -> MimicGen further augments to NxM (varying motions). This is the way to trade compute for expensive human data by GPU-accelerated simulation. A while ago, I mentioned that teleoperation is fundamentally not scalable, because we are always limited by 24 hrs/robot/day in the world of atoms. Our new GR00T synthetic data pipeline breaks this barrier in the world of bits. Scaling has been so much fun for LLMs, and it's finally our turn to have fun in robotics! We are building tools to enable everyone in the ecosystem to scale up with us. Links in thread:

Jim Fan

364,565 views • 2 years ago

Agentic AI will transform every enterprise–but only if agents are trusted experts. The key: Evaluation & tuning on specialized, expert data. I’m excited to announce two new products to support this–Snorkel AI Evaluate & Expert Data-as-a-Service–along w/ our $100M Series D! --- Snorkel Evaluate is our new data-centric agentic AI evaluation platform for specialized, mission-critical enterprise settings where vibe checks and out-of-the-box metrics driven by simple LLM prompts are not enough. Snorkel Expert Data-as-a-Service is our white glove service for expert-level AI datasets, powering frontier LLM developers in areas like expert knowledge, reasoning, agentic action and tool use, and more! Both built on top of Snorkel AI’s Data Development Platform, using our programmatic technology to drive higher-quality expert data, faster– for getting specialized AI to real production value. If you’re building enterprise AI and want to partner around the key ingredient in AI today–the data–book a demo and let's talk! Finally, see thread for details on 🧵👇 - 📽️ A walkthrough of Snorkel Evaluate and Expert Data-as-a-Service on an agentic AI enterprise task - 📅 An upcoming event on Enterprise Agentic AI with innovators from Accenture @BNY Comcast Stanford University QBE & others - 📊 An upcoming series of benchmark datasets and model artifact releases 👀 Want early access to the full agentic AI dataset? Retweet this post and we'll send you the link!

Alex Ratner

50,043 views • 1 year ago