Bilawal Sidhu's banner
Bilawal Sidhu's profile picture

Bilawal Sidhu

@bilawalsidhu110,320 subscribers

Spatial intelligence. World models. Visual effects. Creator w/ 2.1M+ audience. Tech Curator @ TED. A16z Scout. Ex-Google PM (AR/VR & 3D Maps) https://t.co/fysPkbPoQ2

Shorts

2d moments frozen in time. Turned into 3d holograms. Left exactly where they happened in the real world. The geospatial memory palace coming to life. Excited baby deer startles mama deer momentarily, who then chills out after :) Can you imagine when your entire camera roll comes to life like this?

2d moments frozen in time. Turned into 3d holograms. Left exactly where they happened in the real world. The geospatial memory palace coming to life. Excited baby deer startles mama deer momentarily, who then chills out after :) Can you imagine when your entire camera roll comes to life like this?

343,600 görüntüleme

peak internet: ai generated cctv footage of police arresting ppl for wearing huge boots video credit: u/Qemmish

peak internet: ai generated cctv footage of police arresting ppl for wearing huge boots video credit: u/Qemmish

25,872,865 görüntüleme

What if every video you ever shot could turn into a permanent 3D hologram in the exact place it happened? I did it with photos first. Now I can do it with any video. Here’s a mama deer and her baby deer, filmed with my iphone a few days ago -- now floating as geo anchored holograms exactly where I captured them.

What if every video you ever shot could turn into a permanent 3D hologram in the exact place it happened? I did it with photos first. Now I can do it with any video. Here’s a mama deer and her baby deer, filmed with my iphone a few days ago -- now floating as geo anchored holograms exactly where I captured them.

132,809 görüntüleme

Progress update: figuring out the 3D location of every photo taken at TED with attribution. It's starting to feel magical -- makes me wonder what experiences are possible if we could do this for every live event ever.

Progress update: figuring out the 3D location of every photo taken at TED with attribution. It's starting to feel magical -- makes me wonder what experiences are possible if we could do this for every live event ever.

100,640 görüntüleme

OpenAI just dropped their Sora research paper. As expected, the video-to-video results are flipping spectacular 🪄 A few other gems:

OpenAI just dropped their Sora research paper. As expected, the video-to-video results are flipping spectacular 🪄 A few other gems:

1,873,706 görüntüleme

Nano Banana Pro is a really good cartographer. Used it to turn low res satellite imagery into a detailed hand drawn map and vector HD map. Pretty wild how well it segments everything and even recovers paths/roads hidden under tree cover. Looks way more detailed than the current google basemap which is pretty sparse in countries like India. Included both in video for comparison.

Nano Banana Pro is a really good cartographer. Used it to turn low res satellite imagery into a detailed hand drawn map and vector HD map. Pretty wild how well it segments everything and even recovers paths/roads hidden under tree cover. Looks way more detailed than the current google basemap which is pretty sparse in countries like India. Included both in video for comparison.

554,024 görüntüleme

Generative AI is super cool… BUT I’m continually blown away with the work happening in ‘procedural’ 3D modelling. Especially given these plugins are for a free (!) 3D tool like Blender. We needed a fancy Houdini license to do this just a few years ago 🤯

Generative AI is super cool… BUT I’m continually blown away with the work happening in ‘procedural’ 3D modelling. Especially given these plugins are for a free (!) 3D tool like Blender. We needed a fancy Houdini license to do this just a few years ago 🤯

1,365,963 görüntüleme

Damn it worked! Genie 3 world --> inpaint UI --> 4x topaz AI upscale --> train 3d gaussian splat You can step inside a painting of Socrates from 1787. Better than any image-to-3d model I've seen. I think Google has stumbled upon the killer app for VR -- the literal holodeck.

Damn it worked! Genie 3 world --> inpaint UI --> 4x topaz AI upscale --> train 3d gaussian splat You can step inside a painting of Socrates from 1787. Better than any image-to-3d model I've seen. I think Google has stumbled upon the killer app for VR -- the literal holodeck.

658,620 görüntüleme

This is IronSight - a 4D reconstruction built by fusing footage from two pairs of meta ray-bans. A google research buddy saw the prototype and joked this would've been a siggraph paper a few years ago. Today it's a weekend build with fable. Stay tuned for more mad science. Happy 4th!

This is IronSight - a 4D reconstruction built by fusing footage from two pairs of meta ray-bans. A google research buddy saw the prototype and joked this would've been a siggraph paper a few years ago. Today it's a weekend build with fable. Stay tuned for more mad science. Happy 4th!

95,105 görüntüleme

This tech maps the physical world in 3d and snaps it perfectly to my camera feed at 60 fps. Visual Positioning Systems (VPS) are the under hyped backbone of spatial computing. This is how we connect the world of bits & atoms.

This tech maps the physical world in 3d and snaps it perfectly to my camera feed at 60 fps. Visual Positioning Systems (VPS) are the under hyped backbone of spatial computing. This is how we connect the world of bits & atoms.

93,660 görüntüleme

One of the wildest emergent capabilities of Genie 3 is that maps actually work. As I walk around the forest, the GPS display updates its heading in real time. Remember. There is no game engine here. This is an AI hallucinating a working navigational instrument purely from next frame prediction. 🤯

One of the wildest emergent capabilities of Genie 3 is that maps actually work. As I walk around the forest, the GPS display updates its heading in real time. Remember. There is no game engine here. This is an AI hallucinating a working navigational instrument purely from next frame prediction. 🤯

247,184 görüntüleme

Lmao. What niche even is this — grassroots dirt track racing meets google maps nerds? Veo 3 videos are seriously ridiculous and fun. Turn audio on for max enjoyment.

Lmao. What niche even is this — grassroots dirt track racing meets google maps nerds? Veo 3 videos are seriously ridiculous and fun. Turn audio on for max enjoyment.

452,046 görüntüleme

I love 3d maps like this because they’re inherently an abstraction of reality. Unlike 3d scans, the goal isn’t to create a 1:1 mirror world. The goal is to create a stylized distillation down to its visual essence, so you can recognize it effortlessly.

I love 3d maps like this because they’re inherently an abstraction of reality. Unlike 3d scans, the goal isn’t to create a 1:1 mirror world. The goal is to create a stylized distillation down to its visual essence, so you can recognize it effortlessly.

249,351 görüntüleme

The lines between code & content are blurring. I made this 3d city block animation in claude 3.7. Then I used runway gen-3's video-to-video to style it like a lego city at night. At the rate things are going, this'll be a shader running in real-time.

The lines between code & content are blurring. I made this 3d city block animation in claude 3.7. Then I used runway gen-3's video-to-video to style it like a lego city at night. At the rate things are going, this'll be a shader running in real-time.

487,987 görüntüleme

AI stitching together multiple video feeds into one omniscient traffic god. This is what happens when cameras start talking to each other -- mapping the trajectory of every vehicle and pedestrian seamlessly across cameras. Spatial intelligence is coming to a city near you.

AI stitching together multiple video feeds into one omniscient traffic god. This is what happens when cameras start talking to each other -- mapping the trajectory of every vehicle and pedestrian seamlessly across cameras. Spatial intelligence is coming to a city near you.

329,235 görüntüleme

Nano is a depth-aware atmospheric haze plugin that uses ML depth estimation to add physically accurate fog and light scattering to your footage. Works *best* on log footage with visible light sources - it analyzes scene highlights then creates airlight (atmospheric scatter) and halation (light bloom) that responds to actual depth in the scene. Pretty clever approach to getting that cinematic haze look without having to pump a fog machine on set. Makes the OG Trapcode Shine look extremely dated (basically 2D light streaks masked by luminance values), and is yet way more controllable than the current crop of generative AI video-to-video tools.

Nano is a depth-aware atmospheric haze plugin that uses ML depth estimation to add physically accurate fog and light scattering to your footage. Works *best* on log footage with visible light sources - it analyzes scene highlights then creates airlight (atmospheric scatter) and halation (light bloom) that responds to actual depth in the scene. Pretty clever approach to getting that cinematic haze look without having to pump a fog machine on set. Makes the OG Trapcode Shine look extremely dated (basically 2D light streaks masked by luminance values), and is yet way more controllable than the current crop of generative AI video-to-video tools.

275,436 görüntüleme

Generative AI is really cool but sometimes you want to drift your car through an intersection in a super specific way. Draw a spline then get 3D onion skinning so you can adjust the curves with a clear spatial reference. iCars plugin for Blender:

Generative AI is really cool but sometimes you want to drift your car through an intersection in a super specific way. Draw a spline then get 3D onion skinning so you can adjust the curves with a clear spatial reference. iCars plugin for Blender:

183,667 görüntüleme

Heads up! Mosaic dropped a pretty wild dataset of 1.26 million 360° images of Prague 🤯 If you're a researcher, creator or developer into 3D/AI/Geo, I think you're gonna wanna play with this Here's the scoop on this 15 TERAPIXEL dataset & the crazy things you can do with it 🧵

Heads up! Mosaic dropped a pretty wild dataset of 1.26 million 360° images of Prague 🤯 If you're a researcher, creator or developer into 3D/AI/Geo, I think you're gonna wanna play with this Here's the scoop on this 15 TERAPIXEL dataset & the crazy things you can do with it 🧵

295,885 görüntüleme

I can no longer walk around a city without seeing a machine-readable 3D model in my head. This is a geometric and semantic 3D model of San Francisco. These maps connect the world of bits and atoms, enabling visual positioning, 3D navigation and geospatial intelligence.

I can no longer walk around a city without seeing a machine-readable 3D model in my head. This is a geometric and semantic 3D model of San Francisco. These maps connect the world of bits and atoms, enabling visual positioning, 3D navigation and geospatial intelligence.

260,241 görüntüleme

Videos