.Tencent Hy HY-World 2.0 has landed,fully OPEN SOURCE !... We kept hearing the same thing: world models look great, but you can't use the output. So we made one you can, real 3D assets you can edit, rearrange, and ship into production. Prompt to engine-ready 3D. For real this time: ✨Directly outputs editable, production-ready 3D assets ✨Seamlessly plugs into existing game pipelines ✨Multimodal input: text, image, videos 🪄Character mode: freely explore streets, buildings, and open worlds with realistic physics, collisions, and no time limits 🔗GitHub:show more

Tencent AI
116,648 Aufrufe • vor 4 Monaten
Seedance 2.0 - Cinematic Summoning VFX Prompt This prompt... generates a multi-cut VFX showcase with a clear progression. It’s designed for realistic 3D with grounded physics, natural character motion and readable, production-style VFX. You don’t need to define the summon, but you can specify what is being summoned if you want control. If you leave it open, the model decides based on the character sheet. You can find gpt image 2 character sheet prompt in the quoted post. Also, this isn’t limited to summoning. You can reuse the same VFX structure for anything, just swap the effect logic. Just change the parts about summoning. Prompt in the replies 👇show more

Kōda
28,400 Aufrufe • vor 3 Monaten
With Hunyuan3D World Model 1.0 now released and open-sourced,... we're excited to showcase the technical highlights behind this impressive innovation: ✅360° Panoramic Generation: Creates complete, immersive “world scenes”, far beyond localized views. ✅Explorable 3D Scene Generation: Generates diverse, spatially consistent 3D worlds from text/image for truly immersive exploration. ✅Interactive/Editable: Achieves separation of foreground objects, background terrain, ground, and sky, for seamless secondary editing. ✅Exportable Mesh: Generated scenes can be exported as 3D meshes for direct import into mainstream game engines and modeling software. ✅Industry-Leading SOTA Evaluation: Surpasses state-of-the-art open-source models in generation quality. As the industry's first open-source model for physical simulation and explorable world generation, Hunyuan3D World Model 1.0 aims to foster a collaborative community ecosystem with developers and enthusiasts. ✨ Try it now: 🤗 Hugging Face:show more

Tencent Hy
23,193 Aufrufe • vor 1 Jahr
2023 was the year of AI avatars 2024 was... the year of AI photos 2025 was the year of AI videos And I think it's becoming clear now that 2026 will be the year of AI world models Fully interactive explorable 3d worlds generated from one or multiple 2d images or a prompt In turn these 2d images can then be generated by AI too So soon you can generate fully explorable virtual 3d worlds based on your own imagination Next will be figuring out how to make those worlds interactive This is World Labs (unaffiliated, but I like it) As always a lot of big AI model companies are now working on the same thing: 3d world models, only World Labs has a real properly working demo (for now) Very exciting time again!show more

@levelsio
584,511 Aufrufe • vor 11 Monaten
Today, we released Lyra 2.0, a framework for generating... persistent, explorable 3D worlds at scale, from NVIDIA Research. Generating large-scale, complex environments is difficult for AI models. Current models often “forget” what spaces look like and lose track of movement over time, causing objects to shift, blur, or appear inconsistent. This prevents them from creating the reliable 3D environments required for downstream simulations. Lyra 2.0 solves these issues by: ✅ Maintaining per-frame 3D geometry to retrieve past frames and establish spatial correspondences ✅ Using self-augmented training to correct its own temporal drifting. Lyra 2.0 turns an image into a 3D world you can walk through, look back, and drop a robot into for real-time rendering, simulation, and immersive applications. ➡️ Learn more: 📄 Read the paper:show more

NVIDIA AI Developer
437,047 Aufrufe • vor 4 Monaten
What if you could turn a single 360° photo... into a production-ready Isaac Sim environment in minutes? That's exactly what we did here. Using World Labs' Marble and an Insta360 X5 capture (rotating on top), we generated a complete navigable 3D environment and populated it with Lightwheel Sim Ready assets (bottom view). The result? A fully interactive scene in Isaac Sim, ready for sim2real testing,. Navigation, manipulation, or any robotics task you need to validate. What used to take weeks of manual 3D modeling and asset placement now takes minutes. Capture once in the real world, simulate everywhere in your training pipeline. This is the future of robotics development with world models. NVIDIA Robotics NVIDIA Omniverse #Sim2Real #Robotics #Simulationshow more

Jonathan Stephens
46,643 Aufrufe • vor 7 Monaten
World Model is trending— let's revisit our HunyuanWorld journey.... We’ve been pioneering open-source 3D world generation in the past two months, and this ride’s only getting started. 🌍 📅 July: HunyuanWorld 1.0 📌 First open-source 3D world model compatible with CG pipelines (Unity/Unreal/Blender) 📌 Hit 2K+ GitHub stars in just two months ⭐—thank you for the love! 📅 August: 1.0-Lite 📌Same top-tier quality, running on consumer GPUs! 📅 September: 1.0-Voyager 📌 Direct 3D output + world memory—taking exploration further! Seamlessly integrated into CG pipelines with layered 3D modeling (assets, terrain, skybox) and fully open-sourced.. we’re fully committed to building open-source spatial intelligence for all! 🚀 💡 Why it matters? ✅ Seamless CG Pipeline Integration: Export generated 3D scenes as standard mesh formats, effortlessly integrating into industry-standard tools like Blender, Unity, and Unreal Engine for direct editing, animation, and physical simulation. ✅ Hierarchical Scene Editing: Deconstruct scenes into semantic layers (sky, background, foreground objects) via instance recognition and layer decomposition, allowing for atomic-level control—independently modify, relocate, or replace objects without rebuilding the entire world. Project page: Github: Amazing creations by Stijn Spanhove camenduru GENEL | AIを用いた動画制作 apolinario 🌐 とりにく Directive Creator 🪥 👇 #AI #3DGeneration #OpenSource #WorldModels #Hunyuan3D #HunyuanWorldshow more

Tencent HY
20,178 Aufrufe • vor 11 Monaten
Can Fable write good TypeGPU code? I’d say so,... but make sure to use the TypeGPU skill, written by yours truly: It made a pretty nice generic horde game from scratch, no engine involved. Assets by the amazing Kay Lousberg: I paid for one of the packs and I encourage you to do the same. No better time to support 3D artists :Dshow more

Konrad Reczko
22,291 Aufrufe • vor 1 Monat
✨ 3d models are now LIVE on Photo AI... 😊 You can now turn any AI photo you make into a 3d model by pressing [ 📦 Make 3d model ] And then you can view it inside Photo AI or download it as a .GLB 3d model file It's still very early in AI generated 3d model world but it's nice to have this feature working already As always, the models will keep improving, so this feature will keep getting better (like it did with video, it sucked before, now it's getting passable) Next would be nice to switch to .USDZ so you can load it straight into your iPhone with ARKit and put it in your room Available now for everyone on the Premium and Ultra planshow more

@levelsio
112,930 Aufrufe • vor 1 Jahr
One trick we discovered for avoiding realistic face moderation... issues in Seedance 2.0 is using character turnaround sheets (front / side / back views). The first video is one of our experiment results — and it runs successfully. We’ve now integrated character turnarounds directly into our workflow + canvas system: 1. If your artwork was generated on our site, you can drag the image into the canvas directly from the Assets tab 2. Click the “Character Turnaround” button above the image to automatically generate a 3-view turnaround sheet 3. Create a new video node and use the turnaround sheet directly with Seedance 2.0 inside the workflow I’ve shared the workflow link in the comments if you want to explore the exact prompts, setup, and workflow details.show more

underwood
32,883 Aufrufe • vor 2 Monaten
Competition 2: Mesh Generation 🦾 Until now, our subnet... has focused on Gaussian splats. That work led to integrations with Unity, building the world’s largest 3D dataset, and early application deployments with real users and organic traffic growth. 🕹️ The next step is clear. Developers want native mesh generation that drops straight into existing game pipelines and unlocks real production use today. 🤖 Our competition format lets us move flexibly between goals, optimizing for near term commercial utility with meshes and balancing longer term research with Gaussian splat world models and digital twins. 📈 With significant open source mesh models releasing in the past weeks, the timing is right and the bar for miner innovation has never been higher. Mesh competition launches in 2 weeks!show more

404
33,090 Aufrufe • vor 7 Monaten
✨ I can now generate 3d assets for my... drone sim at directly from Cursor (sponsor of #vibejam) I need buildings that you'd see in a war torn city, like warehouses in ruins, broken down abandoned houses, bombed out bridges etc. Nano Banana Pro or 2 can generate them really well and then you can put them in an image-to-3d model and you get a GLB or FBX That one you can then import into your Three.js game, the models might be big though, in my case like 16MB, so I ask it to compress it and make it more low poly so it loads fast ThreeJS then loads the individual GLBs on page load and puts them in my drone sim somewhere randomly, I think I should remove some of the grass and match the sandy color of the ruins though to make it fit in moreshow more

@levelsio
134,777 Aufrufe • vor 4 Monaten
This is some quietly impressive work on making video... world models actually controllable in 4D space. VerseCrafter lets you take an input image, use something like Blender to animate the 3D camera path and object trajectories, then uses that to condition generation. Scribbling in 2D feels so crude in comparison. The authors represent everything in a shared 4D world state - static background as a point cloud, moving objects as 3D gaussian trajectories. The gaussians are an interesting choice because they capture position, shape, and orientation probabilistically rather than forcing rigid bounding boxes or category specific models like SMPL-X for human bodies. They bolt this onto frozen Wan2.1 with a lightweight adapter, so they get a strong video prior. They also built a pipeline to auto extract 4D annotations from real world videos to train this puppy. It doesn't look sexy yet, but IMO this is the interface video world models need - actual 3D authoring tools to exert control rather than crude scribbles and prompt incantations.show more

Bilawal Sidhu
25,802 Aufrufe • vor 7 Monaten
This week is already so hot. 🔥 Massive release... from Decart : Lucy 2.0 a World Editing Model running at 1080p, 30FPS in realtime. This is truly exciting, the era of real-time generative reality is here. We are moving from watching AI video to living inside AI video. A breakthrough model capable of transforming the visual world in real-time. Moving beyond offline rendering, Lucy 2.0 delivers high-fidelity 1080p video generation with near-zero latency. Lucy 2.0 literally "redraws" the entire world pixel-by-pixel, while you are watching it. e.g. If you want to be an anime character, it doesn't just put a mask on you. It turns your skin into anime skin, your hair into anime hair, and the lighting in your room into anime lighting. Lucy 2.0 is also trained to stop the generated video from slowly falling apart over time, so the same stream can run much longer without faces and details drifting. So why is this a "Massive Deal"? Traditional AI video-generation model takes a prompt, you wait 10–20 minutes, and the computer "bakes" a video for you. You couldn't touch it or change it while it was happening. But Lucy 2.0 works like a mirror. It happens in real-time (30 frames per second). There is no waiting. You move your hand, the AI character moves its hand instantly. The craziest part isn't the visuals; it's the physics. Usually, AI hallucinations are glitchy—hands merge into faces, walls melt. Lucy 2.0 understands how the world works without being told. It knows that if you take off a helmet, there is hair underneath. It knows that if you splash water, droplets fly. It learned "physics" just by watching millions of videos. The physical behavior you see emerges from learned visual dynamics, not from engineered geometry or explicit physics engines. Their official technical report explicitly states that the model does not use traditional 3D engines, depth maps, or wireframes. It is a "pure diffusion model."show more

Rohan Paul
12,761 Aufrufe • vor 6 Monaten
So, today we have fast SDF sculpting + real-time... AI in Unbound Loop. Old news🥱 Coming up next: -quad-view generation from sculpted geometry -image tweaks via nano🍌 -tripo HD & Low-Poly 3D generation There's more in the upcoming release, but these three deserve a closer look: Quad-View Generation Most platforms offer some version of this, but Loop has a key advantage, your sculpted model is the reference. That means less guessing from the generator. Though it’s not 100% foolproof, like any AI I guess? (I should stop stating the obvious every time). Image Tweaks via Chat Select any generated image and ask for fixes or changes in real time. Works great on unintended quad-view hallucinations, but also handy for quick iterations. Swapping colors, tweaking details, removing elements. Tripo 3D generator Especially the low-poly model, it consistently delivered fantastic game-ready topology when we tried it with our real-time AI output. And it's super fast.show more

Andrea Intg.
15,132 Aufrufe • vor 20 Tagen
Blender just became a prompt box. With Kimi K3... connected through Blender MCP, you can describe a scene like this in plain English and let the model handle the ugly part: terrain, buildings, lighting, materials, camera movement, animation, and the Python scripts holding everything together. The interesting part is not the first render. Kimi can inspect the result, notice that the camera clips through a tree or the city looks suspiciously like plastic, then edit the actual Blender scene and render it again instead of restarting from zero. You still need taste, because “make it cinematic” remains one of humanity’s least useful instructions. But the distance between an empty Blender file and a fully editable 3D world just got embarrassingly small. A few years ago, creating this meant weeks of modeling, scripting, lighting, and animation. Now you can describe the world, watch Kimi build it, and spend your time fixing the final 10% instead of manually constructing the first 90%. Kimi K3 + Blender MCP is basically text-to-3D without trapping the result inside a useless generated video. Every object, material, light, camera, and keyframe stays editable.show more

Rina
264,903 Aufrufe • vor 25 Tagen
If you think OpenAI Sora is a creative toy... like DALLE, ... think again. Sora is a data-driven physics engine. It is a simulation of many worlds, real or fantastical. The simulator learns intricate rendering, "intuitive" physics, long-horizon reasoning, and semantic grounding, all by some denoising and gradient maths. I won't be surprised if Sora is trained on lots of synthetic data using Unreal Engine 5. It has to be! Let's breakdown the following video. Prompt: "Photorealistic closeup video of two pirate ships battling each other as they sail inside a cup of coffee." - The simulator instantiates two exquisite 3D assets: pirate ships with different decorations. Sora has to solve text-to-3D implicitly in its latent space. - The 3D objects are consistently animated as they sail and avoid each other's paths. - Fluid dynamics of the coffee, even the foams that form around the ships. Fluid simulation is an entire sub-field of computer graphics, which traditionally requires very complex algorithms and equations. - Photorealism, almost like rendering with raytracing. - The simulator takes into account the small size of the cup compared to oceans, and applies tilt-shift photography to give a "minuscule" vibe. - The semantics of the scene does not exist in the real world, but the engine still implements the correct physical rules that we expect. Next up: add more modalities and conditioning, then we have a full data-driven UE that will replace all the hand-engineered graphics pipelines.show more

Jim Fan
6,182,991 Aufrufe • vor 2 Jahren
Big win for open-source LLMs! DeepSeek V4 Pro holds... the top open-weights score on SWE-bench Verified, in the GPT-5.5 range. GLM 5.2 leads the open-weight intelligence index and sits near the closed frontier on long-horizon coding. But this leaderboard number is a weak proxy for real performance. It comes from one task set, run through one harness, served at one precision. The same weights can even score differently across providers, since many hosts quantize activations to fp8 and drift the model off its reference weights. Real performance is determined based on whether a model can read a repo, make coordinated edits across files, run the tests, and recover when one breaks. By that measure, the top open models hold up, but only inside the right harness. The teams that actually put DeepSeek V4 into production pipelines as a frontier substitute got there through the harness they built around the model, not by picking a stronger model. If you want to see this in practice, Cline (64k+ stars) has actually built that harness around open models, tuned so they run at production quality. And it's tuned so that these LLMs can run at production quality, with plan and act modes, checkpoints, and terminal feedback. ClinePass is the new access layer on top of it. It runs a curated set of those models inside Cline, narrowed to the ones tested for coding-agent use, with 2 to 5x the standard rate limits and no separate provider accounts, keys, or billing to track. The video below shows the setup, and I worked with the team to put this together. It runs alongside custom keys and local models as well, not in place of them.show more

Avi Chawla
44,124 Aufrufe • vor 1 Monat