We are excited to introduce Stable Fast 3D, Stability... AI’s latest breakthrough in 3D asset generation technology. This innovative model transforms a single input image into a detailed 3D asset in just 0.5 seconds, setting a new standard for speed and quality in the field of 3D reconstruction! Alongside this release, we’ve also published a technical report that highlights how we achieve fast inference speeds with reduced baked illumination and material parameters. 👾You can learn more and access the report here:show more

Stability AI
438,350 次观看 • 2 年前
Wonderland: Navigating 3D Scenes from a Single Image Contributions:... • First, we introduce a representation for controllable 3D generation by leveraging the generative priors from camera-guided video diffusion models. Unlike image models, video diffusion models are trained on extensive video datasets. This enables them to capture comprehensive spatial relationships within scenes across multiple views and embed a form of "3D awareness" in their latent space, which allows us to maintain 3D consistency in novel view synthesis. • Second, to achieve controllable novel view generation, we empower video models with precise control over specified camera motions. We introduce a novel dual-branch conditioning mechanism that effectively incorporates desired diverse camera trajectories into the video diffusion model. This enables expansion of a single image into a multi-view consistent capture of a 3D scene with precise pose control. • Third, to achieve efficient 3D reconstruction, we directly transform video latents into 3DGS. We propose a novel latent-based large reconstruction model (LaLRM) that lifts video latents to 3D in a feed-forward manner. With this design, during inference, our model directly predicts 3DGS from a single input image, effectively aligning the generation and reconstruction tasks—and bridging image space and 3D space—through the video latent space. Compared with reconstructing scenes from images, the video latent space offers a 256× spatial-temporal reduction while retaining essential and consistent 3D structural details. Such a high degree of compression is crucial, as it allows the LaLRM to handle a wider range of 3D scenes within the reconstruction framework, with the same memory constraints.show more

MrNeRF
52,849 次观看 • 1 年前
🔥Meshy Image to 3D has achieved a major breakthrough... once again! We've boosted the generation speed by 20x, making it effortless to obtain a 3D model in just about 60 seconds. No more wait, Image-to-model transformation has never been this lightning-fast! 🔥🚀 Try Meshy image to 3D on Discord now: #3DModeling #GenerativeAI #MeshyAIshow more

MeshyAI
35,533 次观看 • 2 年前
the fact that i can take an image of... a room and turn it into a 3d model in one shot is actually insane this took like 30 seconds from image to 3d modelshow more

Jan
131,884 次观看 • 8 个月前
As we move more and more into 3D, texturing... 3D assets will be essential ☝️ Repainting 3D Assets is a new AI method that can take any 3D asset and paint it with a given text prompt. The results, while low-resolution, are pretty impressive.show more

Dreaming Tulpa 🥓👑
16,291 次观看 • 2 年前
✨ Made a new mini feature on Photo AI:... [ Grab from 3d model ] So the problem is we're at that stage in time (typical for AI) where image-to-3d models are not good enough but are fun to play with, but we know they'll be good enough in 1-2 years With [ Make 3d model ] you already can turn any Photo AI pic into a 3d model but it still looks hyper clunky and deformed, but it works! One cool idea I had to make that more useful and made now: Let people make a 3d model then change the view of the it with the 3d viewer, then press [ o ] and it grabs a frame of the 3d That image you can then [ Remix ] (img2img), and it becomes a real photo again and that in turn you can then turn into a video again with [ Make video ] So that essentially gives you a fully freeform camera position control to take photos with One thing I need to fix is the background/skybox, I kinda need to take the original photo and remove the person and just get the background for the 3d model viewer, in this case it should be white, but it's a start!show more

@levelsio
119,210 次观看 • 1 年前
The new Krea 3D feature is insanely good. You... can turn an image into a 3D model to move & rotate in space - and it guides real-time scene generation. Watch me move this couch around a living room ⬇️show more

Justine Moore
32,708 次观看 • 1 年前
Get sculpting-level details in seconds. Meshy 6 Preview is... out now for both Text-to-3D and Image-to-3D, delivering a significant leap in mesh quality. This is just the start.show more

MeshyAI
202,413 次观看 • 10 个月前
Create a 3D model from a single image, set... of images or a text prompt in < 1 minute 😮💨 This new AI paper called CAT3D shows us that it’ll keep getting easier to produce 3D models from 2D images — whether it’s a sparser real world 3D scan (a few photos instead of hundreds) or your favorite 2D image generator like Midjourney (just an image). How does this magic work? “This architecture is similar to video diffusion models, but with camera pose embeddings for each image instead of time embeddings. The generated views are passed into a robust 3D reconstruction pipeline to create the 3D representation (Zip-NeRF or 3DGS)”show more

Bilawal Sidhu
92,792 次观看 • 2 年前
the fact that i can take a random car... image from google and turn it into a 3d asset in one go with ai is completely nuts to meshow more

Jan
72,355 次观看 • 8 个月前
Try new poses and angles! 3D drawing figures are... ideal for when you have a pose in mind but can't find that perfect reference image. Learn how to apply poses and other fun additions like BVH files to your 3D model in this tutorial: #clipstudioshow more

CLIP STUDIO PAINT
165,643 次观看 • 1 年前
DroneSplat: 3D Gaussian Splatting for Robust 3D Reconstruction from... In-the-Wild Drone Imagery Abstract: Drones have become essential tools for reconstructing wild scenes due to their outstanding maneuverability. Recent advances in radiance field methods have achieved remarkable rendering quality, providing a new avenue for 3D reconstruction from drone imagery. However, dynamic distractors in wild environments challenge the static scene assumption in radiance fields, while limited view constraints hinder the accurate capture of underlying scene geometry. To address these challenges, we introduce DroneSplat, a novel framework designed for robust 3D reconstruction from in-the-wild drone imagery. Our method adaptively adjusts masking thresholds by integrating local-global segmentation heuristics with statistical approaches, enabling precise identification and elimination of dynamic distractors in static scenes. We enhance 3D Gaussian Splatting with multi-view stereo predictions and a voxel-guided optimization strategy, supporting high-quality rendering under limited view constraints. For comprehensive evaluation, we provide a drone-captured 3D reconstruction dataset encompassing both dynamic and static scenes. Extensive experiments demonstrate that DroneSplat outperforms both 3DGS and NeRF baselines in handling in-the-wild drone imagery.show more

MrNeRF
21,346 次观看 • 1 年前
Introducing 📦𝗔𝗿𝘁𝗶𝗟𝗮𝘁𝗲𝗻𝘁🔧 (SIGGRAPH Asia 2025) — a high-quality 3D... diffusion model that explicitly models object articulation, paving the way for richer, more realistic assets in embodied AI and simulation: – Generates fully articulated 3D objects – Physically plausible joints & motion – High-fidelity 3D Gaussian appearance – Supports generation from a single real image arXiv: Project: Code (coming soon):show more

Xingang Pan
11,517 次观看 • 8 个月前
Segment Any 3D Gaussians paper page: Interactive 3D segmentation... in radiance fields is an appealing task since its importance in 3D scene understanding and manipulation. However, existing methods face challenges in either achieving fine-grained, multi-granularity segmentation or contending with substantial computational overhead, inhibiting real-time interaction. In this paper, we introduce Segment Any 3D GAussians (SAGA), a novel 3D interactive segmentation approach that seamlessly blends a 2D segmentation foundation model with 3D Gaussian Splatting (3DGS), a recent breakthrough of radiance fields. SAGA efficiently embeds multi-granularity 2D segmentation results generated by the segmentation foundation model into 3D Gaussian point features through well-designed contrastive training. Evaluation on existing benchmarks demonstrates that SAGA can achieve competitive performance with state-of-the-art methods. Moreover, SAGA achieves multi-granularity segmentation and accommodates various prompts, including points, scribbles, and 2D masks. Notably, SAGA can finish the 3D segmentation within milliseconds, achieving nearly 1000x acceleration compared to previous SOTA.show more

AK
69,542 次观看 • 2 年前
DimensionX: Create Any 3D and 4D Scenes from a... Single Image with Controllable Video Diffusion TL;DR: Create 3/4DGS from Video Diffusion Note: Some first inference code released (not all yet). Contributions (cited): • We present DimensionX, a novel framework for generating photorealistic 3D and 4D scenes from only a single image using controllable video diffusion. • We propose ST-Director, which decouples the spatial and temporal priors in video diffusion models by learning (spatial and temporal) dimension-aware modules with our curated datasets. We further enhance the hybriddimension control with a training-free composition approach according to the essence of video diffusion denoising process. • To bridge the gap between video diffusion and real-world scenes, we design a trajectory-aware mechanism for 3D generation and an identity-preserving denoising approach for 4D generation, enabling more realistic and controllable scene synthesis. • Extensive experiments manifest that our DimensionX delivers superior performance in video, 3D, and 4D generation compared with baseline methods.show more

MrNeRF
17,052 次观看 • 1 年前
Humans are social creatures and physical interaction with other... humans plays an important role in our everyday life. In this thread, we explain how such interactions can be reconstructed in 3D from a single image utilising a denoising diffusion probabilistic model. (1/n)show more

Lea Wilken/Müller
76,563 次观看 • 2 年前
A quick test of using 3d drawing with 6DOF... controllers as an "instructor" for a generative AI process. There's so many powerful and fun new ways to create just around the corner.. Here I'm using Dreams, Krea and 3daistudio. The 3d model at the end of the video was generated from the Dreams+Krea output in just around 15 seconds. Only the model on the left is a "true" 3d model. #ai #madeindreamsshow more

Martin Nebelong
132,864 次观看 • 2 年前
Huawei's new 3D lock screen wallpaper is actually pretty... interesting. Unlike iPhone's Spatial Photo, which creates a fixed stereoscopic depth effect from a photo, this lets you scan an object 360°, generate a full 3D Gaussian Splatting asset, and use that as your lock screen. Tilt the phone and you're moving through a reconstructed 3D scene, not a fixed stereo image. The part worth paying attention to: 3D capture moving out of creative pipelines and into everyday phone personalization. That's a different kind of adoption curve.show more

KIRI Engine - 3D Scanner App
185,002 次观看 • 1 个月前
🌍 As some of you might know, last year... we started building an app that required a 3D Map, and we were taken aback by the lack of good SDKs. They’re all clunky, slow, and unstable.🤔 Today, we're thrilled to introduce Cartes - a fast, easy-to-use, and visually appealing 3D Map SDK for Unity. This tool is built to empower your creativity in developing delightful apps and games for the Real World Metaverse. 📱🤳 ⚡️ With blazing-fast performance, our SDK offers a seamless integration for your (modern) Unity projects. We provide default navigation features, intuitive gestures, and clustering capabilities that we meticulously refined over hundreds of hours and proof-tested in guerilla tests. 🏞️ We wanted to build a genuinely 3D map, with 3D terrain, monuments and decorations, able to transition smoothly from a global view down to the human eye level. 🔍 #Unity #AR #Maps #LBEshow more

Tina Debove ᯅ
15,210 次观看 • 3 年前