Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Gpt-6 astra can do in-context learning on mobile manipulation! • Different environment • Different camera angle • Different layout No text prompt, it infers from video It even chooses when to use end-effector or joint space Trace recording after ⬇️

67,278 Aufrufe • vor 15 Tagen •via X (Twitter)

26 Kommentare

Profilbild von Axel
Axelvor 15 Tagen

The agent is asked to establish a plan based on the demonstration only It self-corrects and retries if needed Honestly it's pretty mesmerizing to see, and it will keep getting faster

Profilbild von Axel
Axelvor 15 Tagen

The code is available here: We haven't completely cleaned it yet but your favorite agent will be able to understand our approach

Profilbild von Dhruv Diddi
Dhruv Diddivor 15 Tagen

Nicely done! 💯🦾👏

Profilbild von David Dobáš
David Dobášvor 15 Tagen

Look at the beautiful phone teleop interface

Profilbild von Axel
Axelvor 15 Tagen

woah who built that

Profilbild von Robert Scoble
Robert Scoblevor 15 Tagen

It is getting faster! Congrats.

Profilbild von Axel
Axelvor 15 Tagen

🫡

Profilbild von Andrew Lyubovsky
Andrew Lyubovskyvor 15 Tagen

Pretty Cool !!

Profilbild von Miguel Gregori
Miguel Gregorivor 14 Tagen

Cuando el entorno cambia y el robot no pide un manual nuevo, ahí hay diseño.

Profilbild von NAMAN RAJ
NAMAN RAJvor 15 Tagen

Mind blown! 🤯 That's some advanced AI magic right there! 🧙‍♂️ What do you guys think, is this the future of human-AI collaboration?

Profilbild von AIwithMinal
AIwithMinalvor 15 Tagen

Pure artistic brilliance.

Profilbild von SEAR
SEARvor 15 Tagen

this is exactly the kind of learning Sear is built around

Profilbild von ethereagle · building
ethereagle · buildingvor 15 Tagen

no text, video only. are those frames dumped straight into Astra's context, or is there a separate encoder in front? that's copy-with-the-API vs needing your stack.

Profilbild von Axel
Axelvor 15 Tagen

dumped straight in

Profilbild von atharva ☆
atharva ☆vor 15 Tagen

woah incredible

Profilbild von Axel
Axelvor 15 Tagen

yeah am pretty stoked

Profilbild von Aaron
Aaronvor 15 Tagen

so cool! 😎

Profilbild von Axel
Axelvor 15 Tagen

couldn't believe it at first tbh

Profilbild von Tepulous
Tepulousvor 15 Tagen

$20/min .....

Profilbild von Axel
Axelvor 15 Tagen

not if you use kv-caching

Profilbild von RealMan Robotics
RealMan Roboticsvor 15 Tagen

The shift from simulation to real-world manipulation is the part that really matters. Curious to see how far this generalization can scale across tasks and environments.

Profilbild von AI Quanting
AI Quantingvor 14 Tagen

The control space choice is the bit Id want to see stressed. Does it switch to joint space when the demo path is actually constrained, or does it settle per task type regardless?

Profilbild von Hibrinix
Hibrinixvor 15 Tagen

Meanwhile your future mobile phone

Profilbild von Amogh Shrivastava
Amogh Shrivastavavor 15 Tagen

that looks peak

Profilbild von Axel
Axelvor 15 Tagen

That's because it is

Profilbild von Degenpark
Degenparkvor 15 Tagen

end-effector choice is huge

Ähnliche Videos