Given a monocular video as input, #HOLD reconstructs 3D... hand and object surfaces for every frame without assuming a known object template. Our key insight is that interacting hands and objects provide complementary cues about each other's shape and pose. 1/4show more

Michael Black
21,606 次观看 • 2 年前
As a Product Designer with an educational and professional... background in stop motion and frame-to-frame animation, creating these loaders was seamless. I carefully considered the shape of each frame for the desired animation. Here's my process: 👇👇 ◉ ○ ○ ○ - 1/4show more

Muhammed Adepoju - Design Yoda
41,672 次观看 • 2 年前
GSTAR: Gaussian Surface Tracking and Reconstruction Contributions: • A... new framework for tracking and reconstructing dynamic scenes, combining 3D Gaussians and meshes to effectively manage changes in topology. • A method for Gaussian unbinding and surface re-meshing, allowing for the generation of new surfaces as topologies evolve. • A method for handling large or fast deformations of surfaces between frames using scene flow warping. Abstract (excerpt): However, tracking dynamic surfaces with 3D Gaussians remains challenging due to complex topology changes, such as surfaces appearing, disappearing, or splitting. To address these challenges, we propose GSTAR, a novel method that achieves photo-realistic rendering, accurate surface reconstruction, and reliable 3D tracking for general dynamic scenes with changing topology. Given multi-view captures as input, GSTAR binds Gaussians to mesh faces to represent dynamic objects. For surfaces with consistent topology, GSTAR maintains the mesh topology and tracks the meshes using Gaussians.show more

MrNeRF
22,698 次观看 • 1 年前
Another explanation for Oumuamua's unusual shape has been proposed... The asteroid Oumuamua, which flew through the Solar System in 2017, amazed scientists with its unprecedentedly elongated shape. It is 230 meters long and about 35 meters wide. Astronomers have never seen such objects before and tried to explain its origin. Oumuamua is the first known object to arrive from interstellar space. Its trajectory and speed left no doubt that it does not belong to the Solar System. The object rotated around its axis, changing its brightness, which made it possible to determine its size.show more

Black Hole
237,067 次观看 • 1 年前
NVIDIA finally released Neuralangelo's source code! The model can... turn videos from any device into detailed 3D structures, fully replicating buildings, sculptures, or other real aworld objects or spaces virtually. Here's how it works: A model utilizes a 2D video with multiple angles of an object or scene. I selects frames from different viewpoints to understand depth, size, and shape. The AI creates an initial 3D representation, similar to a sculptor shaping a subject. The render is optimized to enhance details, like a sculptor refining texture. The outcome is a 3D object or scene suitable for virtual reality, digital twins, or robotics.show more

Lior Alexander
478,069 次观看 • 3 年前
Another SUPER useful prototype app to make using visionOS... 27’s High Frame Rate Object Tracking and Low-Latency Video (via LowLevelDeviceResource) Capturing Spatial Video on my Stereo Lens is annoying. Viewfinder is split in half! Lets make a 3D digital one that tracks on top!show more

Brad Lynch
53,941 次观看 • 2 个月前
Gemma 4 just dropped. I had it captioning video... in real-time within an hour. Running locally on a MacBook. No cloud. No API. Real-time scene understanding. Oh and SAM3 is segmenting every object in the same frame. Same laptop.show more

Maziyar PANAHI
196,846 次观看 • 5 个月前
Here are more results from #RigidFormer: predicting physical dynamics... with purely neural simulators — an attempt to learn physical dynamics in a scalable manner. 🤖 1) Controllable Articulated Body Simulation — More Results Additional Unitree G1 humanoid rollouts under controlled motion. Each sample uses a different initial state and control signal (direction and velocity). 🏺 2) Object Fragmentation Simulating the cracking and fragmentation process of objects. Thanks Žiga Kovačič for suggesting this experiment! 🎬 3) Combining Rigidformer with Diffusion-as-Shader for controllable video generation. Note: the meshes shown here are only for visualization — the network takes point clouds as input and predicts the updated state of each point.show more

Zhiyang (Frank) Dou
21,768 次观看 • 3 个月前
Massive performance improvement. This is a bit more of... technical post, but man do I love this stuff! Units navigate the map using a 'Navigation Mesh'. Before, I was using one giant nav mesh that spanned the entire map. The more objects that were placed (especially on a large map such as this 'RadarAttack' map designed by Syphotic | Steel Command), the larger the 'lag' would be after placement. You can see here that there is a massive frame drop and the navmesh doesnt update for almost 5 seconds. Now, there are a ton of tiny navmeshes that connect to one another, and together they cover the whole map. Now, when an object is placed, the navmesh will update instantly, because it no longer needs to parse through every object on the map (potentially thousands!!!). It only needs to parse through the objects that exist in the mini navmesh that the object was placed in (probably only 1-5 objects now!). Performance XP Boost +100! Charles Horwood You might appreciate this one :) #Rts #RTSGame #IndieGameshow more

Smitty | Steel Command
60,072 次观看 • 8 个月前
Static 3D generation isn't enough. We need assets ready... for animation. Our new #SIGGRAPH work, AniGen, takes a single image and generates the 3D shape, skeleton, and skinning weights all at once. Code is fully open-sourced! Kudos to Yihua and VAST AI Research 🧵(1/4)show more

Yanpei Cao
145,188 次观看 • 4 个月前
Gravitational lensing, also known as an Einstein ring This... occurs when the strong gravitational pull of a massive object warps spacetime around it so much that light and other forms of electromagnetic radiation are deflected from a straight path. For example, animations show how a black hole passing in front of a galaxy distorts its visible image. An object with such a powerful gravitational pull is called a gravitational lens.show more

Black Hole
36,586 次观看 • 6 个月前
Gravitational lensing, also known as an Einstein ring This... occurs when the strong gravitational pull of a massive object warps spacetime around it so much that light and other forms of electromagnetic radiation are deflected from a straight path. For example, animations show how a black hole passing in front of a galaxy distorts its visible image. An object with such a powerful gravitational pull is called a gravitational lens.show more

Black Hole
11,651 次观看 • 6 个月前
This is some quietly impressive work on making video... world models actually controllable in 4D space. VerseCrafter lets you take an input image, use something like Blender to animate the 3D camera path and object trajectories, then uses that to condition generation. Scribbling in 2D feels so crude in comparison. The authors represent everything in a shared 4D world state - static background as a point cloud, moving objects as 3D gaussian trajectories. The gaussians are an interesting choice because they capture position, shape, and orientation probabilistically rather than forcing rigid bounding boxes or category specific models like SMPL-X for human bodies. They bolt this onto frozen Wan2.1 with a lightweight adapter, so they get a strong video prior. They also built a pipeline to auto extract 4D annotations from real world videos to train this puppy. It doesn't look sexy yet, but IMO this is the interface video world models need - actual 3D authoring tools to exert control rather than crude scribbles and prompt incantations.show more

Bilawal Sidhu
26,017 次观看 • 7 个月前
a Vietnamese AI engineer just built a real-time system... that counts objects moving down a conveyor belt. it runs on ultralytics' objectcounter. train a detector, define the region you care about, and it handles the rest. works for anything on a line, parts, packages, produce. the clever bit: instead of hand-labeling thousands of frames, he took a single frame, pre-annotated it with meta's segment anything (sam-2, built into ultralytics), and trained yolo11-nano on just that. one annotated frame, and it tracks objects accurately across the entire video.show more

Oliver Prompts
29,631 次观看 • 25 天前
AI that watches your work! 👀 Codya built AI... that watches assembly lines and verifies every step got done right. Engine being assembled, bolt by bolt. The model tracks each one, inserted vs fastened. Standard computer vision would struggle here. It's about occlusion. A worker's hands constantly cover the bolts during assembly. Hand moves in, bolt disappears. Hand moves away, bolt reappears. Without smart tracking, the model forgets which bolt is which every time it's hidden. They solved it with BoT-SORT, a tracking algorithm that keeps object identity through occlusions. When a hand covers a bolt and moves away, the tracker knows it's still the same bolt, same state, same position. SkalskiP looking forward to see more and more industrial use cases with your Roboflow models :) ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →show more

Lukas Ziegler
39,467 次观看 • 1 个月前
Introducing Attio Objects 🚀 We know how hard... it is to find a CRM that fits your unique business model. That's why we built Attio Objects – our powerful data model with custom objects that gives you complete flexibility to structure your CRM exactly how you need it. Along with custom objects, we've also introduced new standard objects: - Workspaces and Users objects for PLG businesses. - A robust Deals object for sales-driven companies. This is the culmination of a 4-year effort, with 3 years of work put in even before launching Attio. Since day one, we've been determined to solve the fundamental problem in the CRM space: the trade-off between power and time-to-value. If you wanted power and flexibility, your CRM would take forever to build and not work well with your stack. If you wanted speed, you'd need to use highly opinionated, inflexible software that doesn't really work for your business. That ends today. With Attio, you no longer have to compromise. Build your CRM your way, fast. Iterate as you grow. High-growth startups like Replicate, , and Modal and more are already using Attio's object architecture to perfectly match their businesses and accelerate their growth. To get all the details, check out our blog post 👇 show more

Attio
26,821 次观看 • 2 年前
FILIPINA BEAUTY MEETS BRITISH ICON! Liza Soberano met Hollywood... icon Tilda Swinton at a luxury brand event in Singapore. In an Instagram video, Liza and Tilda were seen interacting with each other. They held hands and laughed as they talked. Tilda is a top British actress best known for her work in "We Need to Talk About Kevin," "Suspiria," and "Snowpiercer." She won the Oscar for Best Performance by an Actress in a Supporting Role for her portrayal of Karen Crowder in "Michael Clayton." COURTESY: Liza Soberano/Instagram Read more:show more

GMA News
147,372 次观看 • 10 个月前
You can't 3D reconstruct glass from images... ...WRONG! Thanks... for video diffusion, now just about anything is possible! Introducing...Diffusion Knows Transparency (DKT) Transparent and reflective objects usually break robot vision and photogrammetry pipelines because they don't follow the "solid object" rules standard cameras expect. DKT is a new AI model that repurposes the "internal physics engine" found in video generation models to solve this problem. Researchers took a massive video diffusion model (WAN) and fine-tuned it using a custom-built synthetic dataset to turn it into a high-precision depth sensor. To train the AI, they built the first massive synthetic video library of transparent objects, 1.32 million frames of perfectly labeled glass and metal objects in motion. Without ever seeing a "real" labeled video of glass during training, the model (DKT) outperformed all previous specialized systems on real-world benchmarks (ClearPose, DREDS). They created a "lightweight" 1.3B parameter version that runs fast enough (0.17s per frame) to be used on actual robot hardware. Two reasons I find this project important: 1. It further proves that synthetic data will be essential for training the next generation vision models. 2. In real-world robotic tests, using DKT's depth maps nearly doubled the success rate of robot arms trying to pick up objects on tricky reflective or translucent surfaces. At home robots will need to interact with these types of objects on a daily basis. Check out the project page here: Code is LIVE! #Computervision #Robotics #AIshow more

Jonathan Stephens
17,712 次观看 • 8 个月前
Internet Computer contracts can hold and spend Bitcoin without... a bridge Smart contracts on DFINITY Foundation's Internet Computer, called canisters, can control Bitcoin, Ethereum, and Solana addresses outright, signing and submitting transactions with no bridge, wrapped asset, or oracle in between. ICP calls the design Chain Fusion. The trick is a private key that is never assembled. It exists only as secret shares spread across a subnet's nodes, and more than a third of which must cooperate to produce a signature. Contracts can request signatures but never see the key. Spending $BTC still works the Bitcoin way, with one signature required for every coin fragment used as an input, and $ICP's derived keys follow the same standards Bitcoin and Ethereum wallets already use.show more

BSCN
19,777 次观看 • 9 天前