Загрузка видео...

Не удалось загрузить видео

На главную

Gpt-6 astra can do in-context learning on mobile manipulation! • Different environment • Different camera angle • Different layout No text prompt, it infers from video It even chooses when to use end-effector or joint space Trace recording after ⬇️

67,278 просмотров • 15 дней назад •via X (Twitter)

Комментарии: 26

Фото профиля Axel
Axel15 дней назад

The agent is asked to establish a plan based on the demonstration only It self-corrects and retries if needed Honestly it's pretty mesmerizing to see, and it will keep getting faster

Фото профиля Axel
Axel15 дней назад

The code is available here: We haven't completely cleaned it yet but your favorite agent will be able to understand our approach

Фото профиля Dhruv Diddi
Dhruv Diddi15 дней назад

Nicely done! 💯🦾👏

Фото профиля David Dobáš
David Dobáš15 дней назад

Look at the beautiful phone teleop interface

Фото профиля Axel
Axel15 дней назад

woah who built that

Фото профиля Robert Scoble
Robert Scoble15 дней назад

It is getting faster! Congrats.

Фото профиля Axel
Axel14 дней назад

🫡

Фото профиля Andrew Lyubovsky
Andrew Lyubovsky15 дней назад

Pretty Cool !!

Фото профиля Miguel Gregori
Miguel Gregori14 дней назад

Cuando el entorno cambia y el robot no pide un manual nuevo, ahí hay diseño.

Фото профиля NAMAN RAJ
NAMAN RAJ15 дней назад

Mind blown! 🤯 That's some advanced AI magic right there! 🧙‍♂️ What do you guys think, is this the future of human-AI collaboration?

Фото профиля AIwithMinal
AIwithMinal15 дней назад

Pure artistic brilliance.

Фото профиля SEAR
SEAR14 дней назад

this is exactly the kind of learning Sear is built around

Фото профиля ethereagle · building
ethereagle · building15 дней назад

no text, video only. are those frames dumped straight into Astra's context, or is there a separate encoder in front? that's copy-with-the-API vs needing your stack.

Фото профиля Axel
Axel15 дней назад

dumped straight in

Фото профиля atharva ☆
atharva ☆15 дней назад

woah incredible

Фото профиля Axel
Axel15 дней назад

yeah am pretty stoked

Фото профиля Aaron
Aaron15 дней назад

so cool! 😎

Фото профиля Axel
Axel15 дней назад

couldn't believe it at first tbh

Фото профиля Tepulous
Tepulous15 дней назад

$20/min .....

Фото профиля Axel
Axel15 дней назад

not if you use kv-caching

Фото профиля RealMan Robotics
RealMan Robotics15 дней назад

The shift from simulation to real-world manipulation is the part that really matters. Curious to see how far this generalization can scale across tasks and environments.

Фото профиля AI Quanting
AI Quanting14 дней назад

The control space choice is the bit Id want to see stressed. Does it switch to joint space when the demo path is actually constrained, or does it settle per task type regardless?

Фото профиля Hibrinix
Hibrinix14 дней назад

Meanwhile your future mobile phone

Фото профиля Amogh Shrivastava
Amogh Shrivastava14 дней назад

that looks peak

Фото профиля Axel
Axel14 дней назад

That's because it is

Фото профиля Degenpark
Degenpark14 дней назад

end-effector choice is huge

Похожие видео