Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

"run the model on the robot: cloud is too slow." been experimenting. Same VLA, same arms, same task. our custom cloud engine: 5.1x faster compute. 1.7x faster end-to-end round-trip (more in thread)

22,318 Aufrufe • vor 3 Monaten •via X (Twitter)

25 Kommentare

Profilbild von Pabs • Robot Everest
Pabs • Robot Everestvor 3 Monaten

when experimenting with MolmoAct2 on bimanual i2rt YAM arms we got 18Hz to the arms on edge box spends 87% of every cycle just thinking kinda crazy

Profilbild von Pabs • Robot Everest
Pabs • Robot Everestvor 3 Monaten

if we pull inference off the robot it gets lighter, cooler, more efficient and can run big server-side models that hold up in complex terrain only bottleneck is latency + live comms. that's the infra we need to build if we want robots in space and we're doing it.

Profilbild von amv
amvvor 3 Monaten

this is great man! i tried solving the latency issue by temporal ensembling but the next chunk kept landing slower than I could roll the previous one out, so there was no overlap window left to average over lol interesting to see the 1.7x number (will be experimenting with a gpu cluster closer to home now)

Profilbild von Pabs • Robot Everest
Pabs • Robot Everestvor 3 Monaten

yeah that’s the failure mode. for us, overlap goes to 0 the second round-trip beats your chunk horizon. cloud is great for that we have lower latency end-to-end with only 1-3ms std dev. local jitter was all over the place, and that variance is really what kills the overlap imo

Profilbild von sergeycrypto.base.eth
sergeycrypto.base.ethvor 3 Monaten

@pabloberlangab I'm not sure about internet access in the real “cloud” when you're climbing Everest, but I suppose there can be network issues. If so, I don't think cloud computing is a great choice :)

Profilbild von Pabs • Robot Everest
Pabs • Robot Everestvor 3 Monaten

starlink for the win video on how that works soon

Profilbild von sergeycrypto.base.eth
sergeycrypto.base.ethvor 3 Monaten

@pabloberlangab Elon should repost this 😁

Profilbild von Pabs • Robot Everest
Pabs • Robot Everestvor 3 Monaten

yes he should my tweets are bangers

Profilbild von Milan Lustig
Milan Lustigvor 3 Monaten

@pabloberlangab I would bet all of my life savings that the Thor deployment is horribly unoptimized and latency could easily match cloud with better (specialized towards edge/vla) software.

Profilbild von Frame
Framevor 3 Monaten

@pabloberlangab cloud still adds network jitter in real deployments

Profilbild von Pabs • Robot Everest
Pabs • Robot Everestvor 3 Monaten

yes you need good event orchestration when operating with starlink to account for obstructions and change in satellites more on that soon

Profilbild von Samuel Tampubolon
Samuel Tampubolonvor 3 Monaten

Been thinking about this there is so much compute limitation on local robotics. Even if we are to increase the local modal there would be a wattage and battery limitation as well. Wouldn’t be surprised that there is a hybrid solution to this. Let’s say in a factory or a house there is a compute infrastructure(server of sort) to handle the heavy compute processing. Something I do wonder is the speed for vision processing. Will think about it more as I work on it.

Profilbild von 𝕍𝕍αḹḍɛ𝕄Ɔr¡𝕩
𝕍𝕍αḹḍɛ𝕄Ɔr¡𝕩vor 3 Monaten

@pabloberlangab Locally run models will always be better. As we are seeing now, the new models are shifting away from needing massive amounts of computational power

Profilbild von Pabs • Robot Everest
Pabs • Robot Everestvor 3 Monaten

completely disagree, if you look at Claude for instance. moment you could run sonnet locally, fable arrived. you will always have better bigger models if you can tap in the cloud.

Profilbild von Will Hughes
Will Hughesvor 3 Monaten

@pabloberlangab 1.7x faster cloud round-trip is surprising. Latency breakdown?

Profilbild von Pabs • Robot Everest
Pabs • Robot Everestvor 3 Monaten

about 100ms for inference + 200ms for communication when doing cloud vs 500ms when doing edge, graphs available in second video in thread!

Profilbild von 风雪漫千山
风雪漫千山vor 3 Monaten

@pabloberlangab 👍 当本地推理耗时大于网络延时时,云端模型就具备理论优势了 一般本地算力都不会太高,所以我认为端云结合,端侧采集云端推理,是很好的方案

Profilbild von Manu Botija
Manu Botijavor 3 Monaten

@pabloberlangab Cost difference?

Profilbild von Pabs • Robot Everest
Pabs • Robot Everestvor 3 Monaten

cloud is pay as you go and available to everyone, edge you might have to spend upwards of $3000 just to get access to it

Profilbild von Paolo AI
Paolo AIvor 3 Monaten

@pabloberlangab @batsuev_es

Profilbild von Nabilfa-jr
Nabilfa-jrvor 3 Monaten

@pabloberlangab Hey chief..I can help to improve your market,let chat briefly..i dm you

Profilbild von Eliott (r/acc 🤖)
Eliott (r/acc 🤖)vor 3 Monaten

@pabloberlangab The problem is not quite speed as much as it is at the intersection of fault tolerance, expeditionary mission support (what if my robot is under a tunnel? In that one room with bad WiFi?), privacy, and recurring costs

Profilbild von KuphDev
KuphDevvor 3 Monaten

@pabloberlangab That sounds like something I gotta try 👀

Profilbild von Pabs • Robot Everest
Pabs • Robot Everestvor 3 Monaten

for sure movements are sooo clean on cloud

Profilbild von KuphDev
KuphDevvor 3 Monaten

@pabloberlangab Haha dooope I really wanna try a more advanced model but my old ass local GPU is soooo weak…

Ähnliche Videos