正在加载视频...

视频加载失败

🚦 Excited introducing Urban-Sim — our new simulator presented at #CVPR2025 as a highlight paper! ⚡️ Fast training with IsaacSim backend 🏙️ Diverse 3D assets for rich urban scenes 🤖 Towards generalizable robots in dynamic urban environments. Webpage:

18,280 次观看 • 1 年前 •via X (Twitter)

3 条评论

Arrogant Bill 的头像
Arrogant Bill1 年前

Really impressed by Urban-Sim! Fast, diverse simulations like this will push next-gen robotics and real-world applications further. Looking forward to seeing what emerges from this platform.

Page to Pixel Publishing 的头像
Page to Pixel Publishing2 年前

The Art of Flight is a homage to 80s/90s arcade action shmups with a fresh twist on the genre. Pilot multiple ships at the same time to take on oncoming waves of enemies in this fast paced space shooter. Wishlist on Steam today!

Daniel Ortega 的头像
Daniel Ortega1 年前

Esto sí es dar un salto real: simuladores que acercan la robótica urbana a la calle, no solo a pruebas académicas. Felicidades, así se innova de verdad.

相关视频

Autonomous driving through tight, dynamic, stochastic, and adversarial traffic-dynamics on sub-urban roads in India, as well as through partially unstructured environments. This demos showcases the robustness of our motion planning and decision making algorithmic frameworks in enabling #autonomousdriving through seamlessly through such traffic and environmental scenarios. The vehicle starts from a generic open environment at the temple, where there are no traffic-rules to abide by. It then exits the region and assumes a generic autonomous navigation behaviour, negotiating complex traffic scenes. At various points it can be seen that the other vehicles (bikes, autos, bicyclists, and cars) didn't abide by any traffic-rule and moved in crisscross fashion, presenting adversarial scenarios, challenging our autonomous vehicle at Swaayatt Robots to take care of the collision avoidance. This classical motion planning and decision making algorithmic framework is being further scaled up with deep #reinforcementlearning, which will practically solve the sub-urban traffic-dynamics and environment negotiation for #autonomousvehicles in India and throughout the world as well. This demo was done at the Kankali Kali Mata mandir in the city of Bhopal. This demo was a culmination of our prior works and demos: off-roads, on-roads, bidirectional traffic negotiation in single lane roads, and toll-plaza navigation. We have taken up the arduous task of solving the Level-4 autonomous driving by the end of 2024, globally. #machinelearning #deeplearning

Sanjeev Sharma

186,291 次观看 • 2 年前

3D-LLM: Injecting the 3D World into Large Language Models paper page: Large language models (LLMs) and Vision-Language Models (VLMs) have been proven to excel at multiple tasks, such as commonsense reasoning. Powerful as these models can be, they are not grounded in the 3D physical world, which involves richer concepts such as spatial relationships, affordances, physics, layout, and so on. In this work, we propose to inject the 3D world into large language models and introduce a whole new family of 3D-LLMs. Specifically, 3D-LLMs can take 3D point clouds and their features as input and perform a diverse set of 3D-related tasks, including captioning, dense captioning, 3D question answering, task decomposition, 3D grounding, 3D-assisted dialog, navigation, and so on. Using three types of prompting mechanisms that we design, we are able to collect over 300k 3D-language data covering these tasks. To efficiently train 3D-LLMs, we first utilize a 3D feature extractor that obtains 3D features from rendered multi- view images. Then, we use 2D VLMs as our backbones to train our 3D-LLMs. By introducing a 3D localization mechanism, 3D-LLMs can better capture 3D spatial information. Experiments on ScanQA show that our model outperforms state-of-the-art baselines by a large margin (e.g., the BLEU-1 score surpasses state-of-the-art score by 9%). Furthermore, experiments on our held-in datasets for 3D captioning, task composition, and 3D-assisted dialogue show that our model outperforms 2D VLMs. Qualitative examples also show that our model could perform more tasks beyond the scope of existing LLMs and VLMs.

AK

249,798 次观看 • 3 年前