Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

🙏Thanks I'm overwhelmed by all the response!😍 ✨ Quick test with Pano360Sharp Just playing around with the output... 4M splats Added my Unity re-lighting and that dreamy tilt-shift blur effect 🎬 There's something magical about turning a single 360° photo into a little 3D world you can fly through...

25,097 görüntüleme • 8 ay önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

3D scanning and rendering is moving so fast - got my splats up and running and I'm mind blown getting ~100fps for this complex 3D scene ⬇️ 🤯 1. WAY faster than NeRF: For comparison, NeRFs would takes around 10 seconds per frame (!) Instead I'm zipping around with FPV controls without breaking a sweat - though I do crash a few times towards the end of the video lol 2. Old Meets New: Gaussian Splatting is cool in that it fuses classical graphics and deep learning techniques. Like NeRFs, this is still a radiance field - just without the slower (ne)ural rendering part. 3. Explicit Representation: Instead you represent a 3D scene as a collection of ellipsoidal "splats" called gaussians. Each gaussian has a position, size, and color. Rendering in real-time is done by projecting into the image plane and alpha blending. 4. Photorealistic Effects: Gaussian splatting use spherical harmonics to represent the view-dependent effects and lighting - allowing surfaces to change color when viewed from different angles, enabling greater photorealism. It doesn't use a neural network, but the training loop is similar to deep learning. 5. Enables Direct Editing: But it's not just speed - with Gaussian Splatting you also get 3D editing support! So you can select, move, and delete stuff, even relight stuff. This type of editing has been more tedious to do with NeRFs and their implicit black box representations. 📲 More tests cooking! Much more to unpack here including simpler explanations. If you enjoyed this post, you might enjoy my feed: Bilawal Sidhu

Bilawal Sidhu

337,090 görüntüleme • 2 yıl önce

✨ I made a 100% AI video of a single night in 🇩🇪 Berlin's underground techno scene Doing an AI shoot is very much like a real life shoot, you have to prepare locations, settings, outfits and models, you can do the same: 1) First I generated photos with the [ 🇩🇪 Berlin nightlife ] photo pack on they're all non-existing AI people except me and my gf 2) Then I put the pics I liked in Kling AI to make them into 10 second videos (with Professional Mode), soon I'll have API access and it'll be one-click inside Photo AI to turn a photo into a 10-sec video, it then takes about 5 minutes to make a video. It's not always great, especially dancing is hard to get right! 3) Then I wanted to have a German narrator talking about Berlin's techno nights in a poetic way like a documentary, so I asked ChatGPT 4o: "write in the style of a cult German novelist about the Berlin techno scene" I then copied that text into Eleven Labs and selected a German AI voice 4) I then collected all the videos, the music and the audio narrator and edited them together in Final Cut Pro adding music from HÖR and adding audio effects 5) Then I added English captions with CapCut Time to make it: 3 hours for 1 minute of video I hope you like it! My dream is to have this entire pipeline in Photo AI at some point. Connecting to Kling AI's video API is the first step to that. Imagine just writing a prompt and you end up with a video like this without all the work Credits: Music: Asquith - Let Me (Rave Mix) [ASQ004] Barbax - Si Vis Pacem Para Bellum (Original Mix) Mixed by Ellen Allien @ HÖR Video: 100% AI characters and video by Photo AI Edited by me Narrator voice: 100% written by ChatGPT 4o 100% narrated by Eleven Labs AI voice

@levelsio

676,094 görüntüleme • 1 yıl önce

"Hah - generative ai can't even make an image of a hand with the right number of fingers.." "Stop pushing this slop" It's way past the point now where it must be clear to everyone, that generative ai is here to stay AND that the quality will continue to increase. I've been talking about this trajectory for years now, and I've been working towards finding ways to combine the strength of these models, with the best of what I love about "old-school" creation. Building with my hands, moving a pencil across the paper and seeing shapes emerge, moving a building slightly to the right to get just that composition I had in mind. Being fully immersed in a scene I'm building in VR, being inspired by the immersion to take the story in a new direction. For years I've been talking about how powerful the combination of 3d and generative ai is, be it traditional 3d, SDF volumes in Dreams or Gaussian splats - with experiments around using V2V as a "render pass" or with experiments around realtime ai. Enough talk you might think, where's the proof? It's all around us these days honestly and here's a small test I did during some OOO. Blender MPC + Fable - a pretty powerful combination! With a bit of Google Omni Fast on top as a "render" pass. What do you think of where this is heading? Hopeful, disheartened, inspired or the opposite? Can you imagine working with tools like this in a way where we still retain the human "spark" and the creative nerve that makes each persons creation unique?

Martin Nebelong

49,395 görüntüleme • 1 ay önce

I asked Garry Tan how to use meta prompting to get better at AI: "My partners at YC Jared Friedman and Pete Koomen showed me how to do this. You can take almost anything that you do all the time and just drop it into a context window. And then say, “Here’s a bunch of inputs and outputs." And maybe you also add a bunch of notes. And then you tell it, “Write me a prompt that can act as an agent that takes this input and makes this output over here.” You can do this for almost any type of knowledge work. And you can even introspect. "What are things you notice that I did to convert this from the input to the output?”. And then you can just start using the prompt. Initially, it’s going to suck. Because it’s just not that smart yet. But what’s funny is now, I also use it to Iterate my writing. You can be very direct, "I would never say that", "Don’t say it like this", or "Oh, you used the long word there, use the short word". Just speak to it conversationally. And then when you're happy with the output, you can use that new output to make a new prompt. "Based on this conversation, give me a better initial prompt that incorporates all the things we talked about." And you can do this with literally everything. And in theory, there’s so much it applies to that people do day-to-day. You could use it for tweets. You could use it for editing podcasts. You can use it for pretty much everything. I have a folder of prompts that I use all the time. My YouTube prompt is on v27 or something. I'll go through this process with all the different max models. I'll use GPT 5.2 Pro. I’ll use Grok. I'll use Claude. Then, I’ll take all the outputs from all the models and put them into Claude and say "Here’s my prompt, here’s the output from four LLMs, including yourself. Rate each response and tell me what the pros and cons of each approach are." And I usually say "give it to me in numbered form". And then you can agree with one, disagree with two, tell it three is this or that. And then after that, you say given all of this, synthesize it."

The Peel

51,632 görüntüleme • 6 ay önce