Bilawal Sidhu's banner
Bilawal Sidhu's profile picture

Bilawal Sidhu

@bilawalsidhu110,320 subscribers

Spatial intelligence. World models. Visual effects. Creator w/ 2.1M+ audience. Tech Curator @ TED. A16z Scout. Ex-Google PM (AR/VR & 3D Maps) https://t.co/fysPkbPoQ2

Shorts

2d moments frozen in time. Turned into 3d holograms. Left exactly where they happened in the real world. The geospatial memory palace coming to life. Excited baby deer startles mama deer momentarily, who then chills out after :) Can you imagine when your entire camera roll comes to life like this?

2d moments frozen in time. Turned into 3d holograms. Left exactly where they happened in the real world. The geospatial memory palace coming to life. Excited baby deer startles mama deer momentarily, who then chills out after :) Can you imagine when your entire camera roll comes to life like this?

343,600 次观看

peak internet: ai generated cctv footage of police arresting ppl for wearing huge boots video credit: u/Qemmish

peak internet: ai generated cctv footage of police arresting ppl for wearing huge boots video credit: u/Qemmish

25,872,865 次观看

What if every video you ever shot could turn into a permanent 3D hologram in the exact place it happened? I did it with photos first. Now I can do it with any video. Here’s a mama deer and her baby deer, filmed with my iphone a few days ago -- now floating as geo anchored holograms exactly where I captured them.

What if every video you ever shot could turn into a permanent 3D hologram in the exact place it happened? I did it with photos first. Now I can do it with any video. Here’s a mama deer and her baby deer, filmed with my iphone a few days ago -- now floating as geo anchored holograms exactly where I captured them.

132,809 次观看

Progress update: figuring out the 3D location of every photo taken at TED with attribution. It's starting to feel magical -- makes me wonder what experiences are possible if we could do this for every live event ever.

Progress update: figuring out the 3D location of every photo taken at TED with attribution. It's starting to feel magical -- makes me wonder what experiences are possible if we could do this for every live event ever.

100,640 次观看

OpenAI just dropped their Sora research paper. As expected, the video-to-video results are flipping spectacular 🪄 A few other gems:

OpenAI just dropped their Sora research paper. As expected, the video-to-video results are flipping spectacular 🪄 A few other gems:

1,873,706 次观看

Nano Banana Pro is a really good cartographer. Used it to turn low res satellite imagery into a detailed hand drawn map and vector HD map. Pretty wild how well it segments everything and even recovers paths/roads hidden under tree cover. Looks way more detailed than the current google basemap which is pretty sparse in countries like India. Included both in video for comparison.

Nano Banana Pro is a really good cartographer. Used it to turn low res satellite imagery into a detailed hand drawn map and vector HD map. Pretty wild how well it segments everything and even recovers paths/roads hidden under tree cover. Looks way more detailed than the current google basemap which is pretty sparse in countries like India. Included both in video for comparison.

554,024 次观看

Generative AI is super cool… BUT I’m continually blown away with the work happening in ‘procedural’ 3D modelling. Especially given these plugins are for a free (!) 3D tool like Blender. We needed a fancy Houdini license to do this just a few years ago 🤯

Generative AI is super cool… BUT I’m continually blown away with the work happening in ‘procedural’ 3D modelling. Especially given these plugins are for a free (!) 3D tool like Blender. We needed a fancy Houdini license to do this just a few years ago 🤯

1,365,963 次观看

Damn it worked! Genie 3 world --> inpaint UI --> 4x topaz AI upscale --> train 3d gaussian splat You can step inside a painting of Socrates from 1787. Better than any image-to-3d model I've seen. I think Google has stumbled upon the killer app for VR -- the literal holodeck.

Damn it worked! Genie 3 world --> inpaint UI --> 4x topaz AI upscale --> train 3d gaussian splat You can step inside a painting of Socrates from 1787. Better than any image-to-3d model I've seen. I think Google has stumbled upon the killer app for VR -- the literal holodeck.

658,620 次观看

This is IronSight - a 4D reconstruction built by fusing footage from two pairs of meta ray-bans. A google research buddy saw the prototype and joked this would've been a siggraph paper a few years ago. Today it's a weekend build with fable. Stay tuned for more mad science. Happy 4th!

This is IronSight - a 4D reconstruction built by fusing footage from two pairs of meta ray-bans. A google research buddy saw the prototype and joked this would've been a siggraph paper a few years ago. Today it's a weekend build with fable. Stay tuned for more mad science. Happy 4th!

95,105 次观看

This tech maps the physical world in 3d and snaps it perfectly to my camera feed at 60 fps. Visual Positioning Systems (VPS) are the under hyped backbone of spatial computing. This is how we connect the world of bits & atoms.

This tech maps the physical world in 3d and snaps it perfectly to my camera feed at 60 fps. Visual Positioning Systems (VPS) are the under hyped backbone of spatial computing. This is how we connect the world of bits & atoms.

93,660 次观看

One of the wildest emergent capabilities of Genie 3 is that maps actually work. As I walk around the forest, the GPS display updates its heading in real time. Remember. There is no game engine here. This is an AI hallucinating a working navigational instrument purely from next frame prediction. 🤯

One of the wildest emergent capabilities of Genie 3 is that maps actually work. As I walk around the forest, the GPS display updates its heading in real time. Remember. There is no game engine here. This is an AI hallucinating a working navigational instrument purely from next frame prediction. 🤯

247,184 次观看

Lmao. What niche even is this — grassroots dirt track racing meets google maps nerds? Veo 3 videos are seriously ridiculous and fun. Turn audio on for max enjoyment.

Lmao. What niche even is this — grassroots dirt track racing meets google maps nerds? Veo 3 videos are seriously ridiculous and fun. Turn audio on for max enjoyment.

452,046 次观看

I love 3d maps like this because they’re inherently an abstraction of reality. Unlike 3d scans, the goal isn’t to create a 1:1 mirror world. The goal is to create a stylized distillation down to its visual essence, so you can recognize it effortlessly.

I love 3d maps like this because they’re inherently an abstraction of reality. Unlike 3d scans, the goal isn’t to create a 1:1 mirror world. The goal is to create a stylized distillation down to its visual essence, so you can recognize it effortlessly.

249,351 次观看

The lines between code & content are blurring. I made this 3d city block animation in claude 3.7. Then I used runway gen-3's video-to-video to style it like a lego city at night. At the rate things are going, this'll be a shader running in real-time.

The lines between code & content are blurring. I made this 3d city block animation in claude 3.7. Then I used runway gen-3's video-to-video to style it like a lego city at night. At the rate things are going, this'll be a shader running in real-time.

487,987 次观看

AI stitching together multiple video feeds into one omniscient traffic god. This is what happens when cameras start talking to each other -- mapping the trajectory of every vehicle and pedestrian seamlessly across cameras. Spatial intelligence is coming to a city near you.

AI stitching together multiple video feeds into one omniscient traffic god. This is what happens when cameras start talking to each other -- mapping the trajectory of every vehicle and pedestrian seamlessly across cameras. Spatial intelligence is coming to a city near you.

329,235 次观看

Nano is a depth-aware atmospheric haze plugin that uses ML depth estimation to add physically accurate fog and light scattering to your footage. Works *best* on log footage with visible light sources - it analyzes scene highlights then creates airlight (atmospheric scatter) and halation (light bloom) that responds to actual depth in the scene. Pretty clever approach to getting that cinematic haze look without having to pump a fog machine on set. Makes the OG Trapcode Shine look extremely dated (basically 2D light streaks masked by luminance values), and is yet way more controllable than the current crop of generative AI video-to-video tools.

Nano is a depth-aware atmospheric haze plugin that uses ML depth estimation to add physically accurate fog and light scattering to your footage. Works *best* on log footage with visible light sources - it analyzes scene highlights then creates airlight (atmospheric scatter) and halation (light bloom) that responds to actual depth in the scene. Pretty clever approach to getting that cinematic haze look without having to pump a fog machine on set. Makes the OG Trapcode Shine look extremely dated (basically 2D light streaks masked by luminance values), and is yet way more controllable than the current crop of generative AI video-to-video tools.

275,436 次观看

Generative AI is really cool but sometimes you want to drift your car through an intersection in a super specific way. Draw a spline then get 3D onion skinning so you can adjust the curves with a clear spatial reference. iCars plugin for Blender:

Generative AI is really cool but sometimes you want to drift your car through an intersection in a super specific way. Draw a spline then get 3D onion skinning so you can adjust the curves with a clear spatial reference. iCars plugin for Blender:

183,667 次观看

Heads up! Mosaic dropped a pretty wild dataset of 1.26 million 360° images of Prague 🤯 If you're a researcher, creator or developer into 3D/AI/Geo, I think you're gonna wanna play with this Here's the scoop on this 15 TERAPIXEL dataset & the crazy things you can do with it 🧵

Heads up! Mosaic dropped a pretty wild dataset of 1.26 million 360° images of Prague 🤯 If you're a researcher, creator or developer into 3D/AI/Geo, I think you're gonna wanna play with this Here's the scoop on this 15 TERAPIXEL dataset & the crazy things you can do with it 🧵

295,885 次观看

I can no longer walk around a city without seeing a machine-readable 3D model in my head. This is a geometric and semantic 3D model of San Francisco. These maps connect the world of bits and atoms, enabling visual positioning, 3D navigation and geospatial intelligence.

I can no longer walk around a city without seeing a machine-readable 3D model in my head. This is a geometric and semantic 3D model of San Francisco. These maps connect the world of bits and atoms, enabling visual positioning, 3D navigation and geospatial intelligence.

260,241 次观看

Videos