🔥Spatial intelligence needs fast, *interactive* 3D world generation 🎮... — introducing WonderWorld: generating 3D scenes interactively following your movement and content requests, and see them in <10 seconds! 🧵1/6 Web: arXiv:show more

Hong-Xing (Koven) Yu
169,179 görüntüleme • 2 yıl önce
🔥Spatial intelligence requires world generation, and now we have... the first comprehensive evaluation benchmark📏 for it! Introducing WorldScore: Unifying evaluation for 3D, 4D, and video models on world generation! 🧵1/7 Web: arxiv:show more

Hong-Xing (Koven) Yu
66,876 görüntüleme • 1 yıl önce
Tripo P1.0 is live. Introducing Smart Mesh in Tripo... Studio — generate structured 3D meshes in ~2 seconds. ⚡ ~2s generation 🧠 Smart topology 🛠 Production-ready assets 🌍 Built for game pipelines, real-time rendering, web 3D tools, and scalable content production From prompt to usable mesh — almost instantly. Try it: #3DAI #3DCG #3dart #Triposhow more

Tripo
44,954 görüntüleme • 6 ay önce
DimensionX: Create Any 3D and 4D Scenes from a... Single Image with Controllable Video Diffusion TL;DR: Create 3/4DGS from Video Diffusion Note: Some first inference code released (not all yet). Contributions (cited): • We present DimensionX, a novel framework for generating photorealistic 3D and 4D scenes from only a single image using controllable video diffusion. • We propose ST-Director, which decouples the spatial and temporal priors in video diffusion models by learning (spatial and temporal) dimension-aware modules with our curated datasets. We further enhance the hybriddimension control with a training-free composition approach according to the essence of video diffusion denoising process. • To bridge the gap between video diffusion and real-world scenes, we design a trajectory-aware mechanism for 3D generation and an identity-preserving denoising approach for 4D generation, enabling more realistic and controllable scene synthesis. • Extensive experiments manifest that our DimensionX delivers superior performance in video, 3D, and 4D generation compared with baseline methods.show more

MrNeRF
17,062 görüntüleme • 1 yıl önce
Start your 3D engines 🏎️ Learn how to build... a 3D racing experience in less than an hour in Unity Studio, our web-based editor designed for industries like manufacturing, automotive, healthcare, architecture, and more. Join us live on June 3, 2026, at 9:00 AM PDT / 12:00 PM EDT to follow along step-by-step with Unity experts as we take pre-built assets and turn them into an interactive racing environment. You’ll learn how to: 🏎️ Build and customize interactive 3D scenes 💡 Add lighting, materials, and interactions quickly 🛠️ Assemble and iterate using Studio workflows 🎉 Publish and share your experience Register here:show more

Unity
34,682 görüntüleme • 3 ay önce
World Model is trending— let's revisit our HunyuanWorld journey.... We’ve been pioneering open-source 3D world generation in the past two months, and this ride’s only getting started. 🌍 📅 July: HunyuanWorld 1.0 📌 First open-source 3D world model compatible with CG pipelines (Unity/Unreal/Blender) 📌 Hit 2K+ GitHub stars in just two months ⭐—thank you for the love! 📅 August: 1.0-Lite 📌Same top-tier quality, running on consumer GPUs! 📅 September: 1.0-Voyager 📌 Direct 3D output + world memory—taking exploration further! Seamlessly integrated into CG pipelines with layered 3D modeling (assets, terrain, skybox) and fully open-sourced.. we’re fully committed to building open-source spatial intelligence for all! 🚀 💡 Why it matters? ✅ Seamless CG Pipeline Integration: Export generated 3D scenes as standard mesh formats, effortlessly integrating into industry-standard tools like Blender, Unity, and Unreal Engine for direct editing, animation, and physical simulation. ✅ Hierarchical Scene Editing: Deconstruct scenes into semantic layers (sky, background, foreground objects) via instance recognition and layer decomposition, allowing for atomic-level control—independently modify, relocate, or replace objects without rebuilding the entire world. Project page: Github: Amazing creations by Stijn Spanhove camenduru GENEL | AIを用いた動画制作 apolinario 🌐 とりにく Directive Creator 🪥 👇 #AI #3DGeneration #OpenSource #WorldModels #Hunyuan3D #HunyuanWorldshow more

Tencent HY
20,178 görüntüleme • 11 ay önce
How a 22-year-old developer built a full 3D Jet... Ski racing game in just 40 minutes with zero manual coding He used Claude Opus 5 to generate physics, WebGL 3D graphics, HUD, and audio in a single prompt and turned single-prompt gamedev into a high-margin income stream. Costs: $423 He launched a single-prompt generation workflow that built the entire HTML5 project from scratch: Top layer: A Three.js and WebGL rendering pipeline dynamically creates 3D water physics, real-time wave dynamics, dynamic lighting, and jet ski fluid mechanics, all written autonomously inside one output file without external frameworks. Bottom layer: The Claude Opus 5 engine processed a massive 690-million-token context window to generate the complete gameplay logic, collision handling, dynamic sound generation, controls, and UI layout directly from a detailed initial system prompt. The trend of single-prompt 3D game creation is rapidly exploding across media and indie development. The author monetizes this tech stack through three main channels: 1. Viral Content & Media Systems: Short-form breakdown videos driving massive reach, monetized via promo placements, prompt-pack access, and private developer communities. 2. Rapid Hypercasual Prototyping: Testing 10+ WebGL mechanics per day, flipping fully functional browser games on itch io or CodeCanyon, and licensing prototypes directly to casual game portals. 3. Interactive WebGL Client Solutions: Delivering custom 3D promotional browser games and interactive brand experiences for clients in 48 hours instead of weeks. First month results: > WebGL games generated: 24 > Viral impressions generated: 3.8M+ > Total revenue across licensing & content: $21,400 The AI completely automated the core development lifecycle: Claude Opus 5 built the physics engine, rendered 3D graphics in WebGL, hooked up audio controllers, and generated interactive browser logic with zero manual line-by-line coding. Bookmark it and check article 👇show more

Ridark
11,592 görüntüleme • 28 gün önce
Introducing Kaleido💮 from AI at Meta — a universal... generative neural rendering engine for photorealistic, unified object and scene view synthesis. Kaleido is built on a simple but powerful design philosophy: 3D perception is a form of visual common sense. Following this idea, we formulate rendering purely as a sequence-to-sequence generation problem, successfully unifying neural rendering with the architecture principles behind modern language and video models. Unlike traditional neural rendering methods, Kaleido learns 3D purely in a data-driven way, without explicit 3D representations or structures. It acquires spatial understanding directly through large-scale video pretraining, then multi-view 3D data finetuning, inspired by how LLMs acquire textual common sense from large corpora before specialising in domains like coding. Through extensive ablations, we progressively modernised the architecture design and training strategies and tackled key scaling challenges in sequence-to-sequence generative rendering, arriving at a design that’s simple, versatile, and scalable. Kaleido significantly outperforms prior generative models in few-view settings, and remarkably is the first zero-shot generative method matches InstantNGP-level rendering quality in multi-view settings. We view Kaleido also as an alternative step towards world modeling that flexibly spans a spectrum of “realities": with many views, it faithfully reconstructs grounded reality; with fewer views, it imagines plausible unseen details. 🔗 Explore more results and paper:show more

Shikun Liu
22,442 görüntüleme • 11 ay önce
WeatherEdit: Controllable Weather Editing with 4D Gaussian Field Contributions:... 1. Based on our analysis of weather editing characteristics, we introduce WeatherEdit, a comprehensive and efficient framework for realistic and controllable weather generation. Compared with existing methods that focus on either background editing or static weather effects, a progressive 2D-to-4D transformation process in WeatherEdit enhances adaptability across a wider range of scenarios. 2. We introduce an all-in-one adapter to enable a diffusion model for multi-weather (snowy, rainy, and fog) synthesis, along with a Temporal-View attention to ensure consistent editing across multi-frame and multi-view. 3. We design a 4D Gaussian field for weather particle modeling, enabling plausible simulation of raindrops, snowflakes, and fog with controllable severity. 4. We demonstrate WeatherEdit’s effectiveness in generating realistic, consistent, and controllable weather effects in 3D driving scenes, showcasing its applicability to real-world scenarios.show more

MrNeRF
10,691 görüntüleme • 1 yıl önce
Robots can now reconstruct 3D scenes in real time... from a single RGB camera. [📍 Projects page + paper] No depth sensor. No retraining. 30 FPS. Researchers at the Imperial College London introduced KV-Tracker, a training-free method that makes heavy models like π³ and Depth Anything 3 fast enough for real-time tracking. The idea is simple. These models use global self-attention, which is powerful but computationally expensive. KV-Tracker caches the key and value pairs from selected keyframes and reuses them for new frames. That cache becomes an implicit scene representation. Result: • Up to 30 FPS • 10 to 15x speedup • Accurate 6-DoF tracking on benchmarks like TUM RGB-D and 7-Scenes • Works with monocular RGB only It also supports object-level tracking with masks and allows saving the KV-cache for later reuse. For robotics, this reduces hardware constraints and moves real-time 3D perception closer to practical deployment. Credit to Marwan Taher (Marwan Taher) at Imperial’s Dyson Robotics Lab and many others who contributed to this! 📍 Save projects page + paper for later: Video: ——- if it matters in AI or Robotics you'll read it here first:show more

Ilir Aliu
53,992 görüntüleme • 5 ay önce
We ranked a B2B SaaS brand #1 on ChatGPT... for their category in 7 days. (And it's being used by marketing teams at Webflow, Chime, and Deepgram) This platform tracks AI visibility + generates cited content automatically across ChatGPT, Perplexity, Claude, and Gemini... → No more 6-12 months waiting for Google rankings to move → No more $60K agency dashboards that only show problems → No more 10 different tools to track, create, and publish content → No more manual content gap analysis taking 20+ hours weekly → No more AI slop that ChatGPT refuses to cite Just connect your data sources → autonomous visibility tracking + content generation system. Here's how it works: → AI Citation Scanner (tracks mentions across ChatGPT, Perplexity, Claude, Gemini) → Competitive Gap Analysis (identifies where competitors get cited and you don't) → First-Party Data Integration (connects Zendesk, HubSpot, Drive, product docs) → AI Content Generator (creates authoritative content with human review checkpoints) → Direct CMS Publishing (publishes to Webflow, Contentful automatically) → Performance Measurement (tracks results across traditional + AI search) Companies using this infrastructure: • Webflow: 40% traffic lift + 5X content velocity • Chime: 3X AI citations in 30 days • Deepgram: 24X organic traffic (37K → 1.5M visitors in 60 days) Built with AI-assisted workflows. Runs on human + AI collaboration. 30-day results vs 6-month SEO cycles. Want to see how you rank in AI search? Like + comment "SEO" + repost, and I'll DM you the free scanner. (must be following)show more

Aryan Mahajan
28,510 görüntüleme • 9 ay önce
Created with Seedance 2.0 on BudgetPixel AI Prompt:Create a... highly detailed, cinematic 15-second animated video set in a lively modern city completely operated by anthropomorphic animals. Use polished feature-film-quality 3D animation, expressive but natural animal movements, realistic fur simulation, vibrant environments, smooth camera motion, and consistent character design throughout. 0–3 seconds — Bear Police Officers Open with a wide establishing shot of a busy downtown intersection during a bright morning. Two large brown bears wear neat navy police uniforms with badges and caps. One bear confidently directs traffic with clear hand signals while the other stands beside a police car with flashing lights. Animal pedestrians cross the street naturally in the background. 3–6 seconds — Penguin Ice Cream Shop Use a smooth whip-pan transition to a colorful vintage ice cream truck. Two cheerful penguins wearing small aprons and bow ties serve ice cream cones through the window. A rabbit customer receives a tall strawberry cone while other animals wait in line. Include playful expressions, realistic melting ice cream, and small natural movements. 6–10 seconds — Monkey Bus Driver Transition as the ice cream truck passes in front of the camera, revealing a city bus. A friendly monkey in a professional driver’s uniform drives confidently through the busy street. Rabbits, deer, foxes, and pandas sit inside as passengers. Show the monkey checking the mirror, turning the steering wheel, and stopping smoothly at a bus stop. 10–13 seconds — Busy Animal City The camera follows the bus through the city, revealing raccoons cleaning the sidewalks, squirrels selling fruit at a street market, giraffes working near tall buildings, and birds delivering letters between rooftops. The city should feel organized, energetic, and full of believable activity. 13–15 seconds — Grand Final Shot End with a fast cinematic crane shot rising above the central city square, showing hundreds of animals working, shopping, driving, and socializing together. A fountain sits in the center while the animal city stretches into the distance under warm golden sunlight. Style: premium cinematic 3D animation, playful family-friendly comedy, highly detailed fur and clothing, expressive faces, natural body movement, realistic lighting and shadows, colorful urban production design, smooth transitions, dynamic tracking shots, shallow depth of field, 4K, 24fps, widescreen composition. Avoid: character duplication, changing uniforms, distorted paws, extra limbs, floating objects, unreadable signs, unnatural walking, chaotic traffic, stiff animation, low-detail backgrounds, sudden scene changes, text overlays, logos, subtitles, or watermarks.show more

Sarah Parker
58,693 görüntüleme • 1 ay önce
20 days ago, I connected Claude Code to my... newly created instagram handle.. I gained 4.3M views and 6500+ followers in less than a month [ i post Ai generated animated stories ] Full workflow: i let claude study my account before i write another reel.. This is the cleanest content workflow i've built on claude. give it your IG first. 4 prompts handle the rest.. niche research, the reel script, the hook, and the daily automation.. the whole loop is basically, give claude your IG → find what's working → write retention-optimized scripts → engineer the hook → automate the daily output.. ▫️ Setup: give claude your instagram open claude code. claude code has a built-in web tool that browses any public URL. or install any agentic browser like Browser Harness or Firecrawl or Comet browser paste this with your handle filled in: "Browse and pull the last 30 reels and posts. Analyze my recurring topics, top-performing hooks, formats, and engagement patterns. Then map out my actual audience and what they consistently respond to." claude reads your profile, pulls every reel down, and now has the context to personalize every prompt below to YOUR account, not a generic niche. if you're on claude desktop, the same works with firecrawl MCP connected. ▫️ Prompt 1 find what actually goes viral in your niche: "Analyze the highest-performing Instagram Reels, TikToks, and Reddit posts in the [niche] niche from the last 30 days. Identify repeating hooks, visual styles, emotional triggers, and content formats that consistently generate high engagement. Then summarize the 5 strongest content angles optimized for AI-generated content and short-form videos." run this after the setup. you get 5 angles backed by what's already working in your niche, cross-checked against what's already working on YOUR account. ▫️ Prompt 2 write a high-retention reel script "Write a short-form Instagram Reel script about [topic] with an aggressive hook in the first 2 seconds. Create immediate curiosity, tension, or controversy to stop scrolling, then deliver a fast and satisfying payoff. Keep it under 30 seconds and optimize the structure for watch time, replays, comments, and shares. Finish with a subtle CTA." the line that matters: "optimize the structure for watch time, replays, comments, and shares." claude writes for the metrics, not just the word count. ▫️ Prompt 3 engineer better hooks "Study the top-performing Reels in [niche] and break down the hook structure, pacing, and emotional triggers used in the first 3 seconds. Then generate 5 new hook variations that are even more curiosity-driven, emotionally charged, and optimized to stop scrolling instantly. Focus on triggers like surprise, fear, ego, urgency, or desire." most reels die in the first 2 seconds. this prompt has claude reverse-engineer what already works, then give you 5 sharper versions to swap in. ▫️ Prompt 4 automate the whole workflow "Build a complete AI-powered content workflow for Instagram in the [niche] niche. The system should identify trending topics daily, generate high-retention scripts, create matching AI visuals, turn them into short-form videos, and generate optimized captions and hashtags. Structure everything as a repeatable workflow designed for consistent daily posting and growth." once the niche and script structure are validated, this turns it into a daily loop. one prompt that handles topic → script → visual → video → caption. these 4 prompts are the building blocks. the setup is what makes them yours. your real value is in the [niche] you plug in. content workflow built in one weekend, daily posting on autopilot from monday.show more

Axel Bitblaze 🪓
201,149 görüntüleme • 2 ay önce
$KNDX 🤖 Theres 3 big narratives that are sending... coins left right and centre rn. 🚀 #AI, #Gamefi, & #NFTs 🔹Theres 50% mindshare for #AI. 🤖 🔹#GameFi mcap is hitting ATH's with #OfftheGrid, $XBG and $SUPER making spectacular moves. 🎮 🔹NFTs and the #Metaverse are making a strong comeback with $APE up 100% over the weekend. 🐵 What if there's a project that touches all these trending narratives with groundbreaking technology to disrupt all 3 of them? 🔥 💡- That's where $KNDX comes in. -💡 Kondux is a cutting-edge Web3 SaaS platform, combining NVIDIA’s Omniverse, AI, Blockchain, and dynamic NFTs to revolutionize secure asset management across industries. 👏 Their flagship product, kNFTs, are 3D digital assets usable across Metaverse and Gaming platforms, AR/VR/XR environments, and manufacturing applications. Kondux’s scalable model opens new revenue streams by enabling effective digital asset monetization. 💰 Kondux is the first Web3 project to integrate VFX pipelines with NVIDIA’s Omniverse and bringing it onto the Blockchain. ⛓️ It is also the only Web3 project with a *Select Status Partnership* with NVIDIA, operating under NVIDIA NDAs and working with them directly for more than 2 years. About their NVIDIA Integrations: 🤖 🔹There are three areas of the Kondux tech stack that coincide with three divisions of NVIDIA: 📡GDN (Graphics Delivery Network, the backbone of GeForce Now) 💡Omniverse for 3D aspects such as, geospatial data, real world physics, lighting, and raytracing 🤖NVIDIA AI Foundation, which covers many aspects of #AI, including inference and deployment scaling. The convergence of all these components lie within .USD file format . 🔹 They are the first blockchain project to integrate NVIDIA’s Omniverse Cloud and Graphics Delivery Network (GDN) to provide high-quality 3D content accessible on any device without requiring high-end hardware. 🔹 This setup streamlines content management, democratises access to resource-intensive 3D content, and enables real-time interaction with 3D NFTs. Now, I haven’t seen any crypto project so deeply connected with NVIDIA and NVIDIA technology. GDN is a HUGE competitive advantage. With it, the need for #GPU’s basically goes out the window. 🤯 Now lets take a look at some of the other main features... 👀 OpenUSD (Universal Scene Description): 📽️ 🔹 Kondux is leveraging USD technology, developed by Pixar and used by Meta, Apple, Microsoft and other industry leaders to enhance 3D graphics and interoperability within its creative ecosystem. 🔹 Originally created for high-end film production, USD now supports a variety of applications, including gaming and virtual reality, making it a key asset for Kondux. kNFT's: 🎨 🔹 Kondux is pioneering a new category of NFTs known as kNFTs, which aim to redefine NFT utility through innovative features. 🔹 A standout feature is the upgradeable aspect provided by Kondux DNA, allowing kNFTs to transform and combine with other NFTs, creating limitless possibilities in art, gaming, and music. 🔹Through the Kondux AI portal it will be possible to communicate with kNFTs. They can learn and adapt. This AI technology is revolutionary because it makes human to kNFT interaction possible, turning it into a unique, personalized experience. Check out the clip of kNFTs in Unreal Engine 5 gameplay below. 👇 Kondux is a very obvious utility play with huge upside because it’s multi narrative. 📈 It's seriously groundbreaking stuff that they’re about to launch. 🚀 After speaking with the team there’s no doubt in my mind this will do crazy big numbers in the next months. 🤑show more

Altcoin Miyagi🇯🇵
17,323 görüntüleme • 1 yıl önce
A 24-year-old built two AI girls with Claude and... now clears $21,800 a month from them. The build took 15 days. He trained separate LoRAs for both girls, locked their identity seeds, and kept small imperfections on purpose: a loose strand of hair, tiny skin marks, slightly uneven framing. Perfect symmetry gets flagged. Small inconsistencies make them look real. He posts 5 times a day across TikTok, Instagram and X. Morning routines, gym sessions, mirror videos, outfit changes, pool clips, and videos of the two girls together. The content is designed so they look like two real friends who actually live in the same world. The smartest part is that the accounts interact with each other. One girl comments on the other’s posts, appears in her videos, and references things they supposedly did together. Followers stop seeing them as two AI models and start following the relationship between the characters. Within three months they crossed 312,000 followers combined and started receiving hundreds of DMs every night. The private channel sits at $25 a month, while an AI memory agent keeps track of every conversation, previous message, favorite post, and personal detail each follower has shared. Replies come back in under 30 seconds. The agent checks the user's previous conversations before answering, so the response feels consistent with the personality of the girl they are talking to instead of sounding like another generic AI chatbot. By the end of month three, the two accounts were generating $13,900 from subscriptions and private chats, another $5,700 from brand deals, and around $2,200 from digital products. The brands came after the audience started growing: clothing companies, beauty products, fitness brands, and lifestyle products wanted access to the same audience that was already following the two characters every day. The Claude stack that locked them: 1Full identity, personality, lighting, camera style and body proportions locked into separate character systems. 2Separate LoRAs trained only on each girl's approved character frames. 3Apartment, bedroom, gym and outdoor locations generated once and reused to keep the world consistent. 4Every video built around natural movement, imperfect framing and small variations instead of polished AI-perfect shots. 5Memory agent connected to the conversations so both girls remember what followers previously said. 6Upscaling, face consistency and final post-processing before everything goes live. The first girl brings people into the account. The second gives them another character to follow, another story to watch, and another reason to come back. The content gets them interested. The relationship between the two characters keeps them watching. The memory agent turns that attention into recurring revenue.show more

genuenci
671,331 görüntüleme • 22 gün önce
Season 2 of Valiants: Tap-Tap is here!🌟 It’s time... to dive into the new and exciting world of Valoria! Here’s everything you need to know: Now available here! 🔥 New Points Store Your efforts in Season 1 have paid off. Now, you can use your store points to redeem amazing rewards. What will you find? Surprise boxes containing exclusive items in Valiants: Tap-Tap and Valiants: Arena, NFTs that you can trade in the future, and a portion of the $VGN Airdrop allocation. Important: For now, the boxes can only be purchased, but soon you’ll be able to open them to discover what’s inside. Stay tuned for updates so you don’t miss anything! 🕒 🎁 Wild Spin - Spin and Win Introducing the new **Wild Spin**, where each spin could change your fate. Spins will be awarded based on the activity and performance of your friends with a good User Score. So, stay active and keep connecting with your teammates to maximize your chances! What can you win? Experience, unlocks, and even USDT prizes that you can withdraw directly to your airdrop wallet. Tip: Use your USDT winnings to purchase additional spins and keep the wheel turning. How far will your luck take you? 🎰 💥New Valiants and Items The battle in Valoria heats up with the arrival of new Valiants and accessories. These exclusive items will allow you to explore new strategies and challenge your opponents like never before. 📋 Patch Notes Not only are we introducing new features, but we’ve also made several improvements and adjustments to provide you with a smoother and more satisfying gaming experience. Here are all the details: - Daily Combo Changes: The combo is now linear, meaning the challenge has increased. Can you keep up the streak? This change is designed to reward the most skilled and dedicated players. - Daily Login Update: Starting from day 10, the rewards get even better. Don’t miss a single day to maximize your benefits! We’ve also adjusted the progression so that each login brings you closer to more significant rewards. - Complete UI Redesign: We’ve revamped the entire user interface to make it more intuitive and visually appealing. Additionally, you can now choose between two visual themes: the serenity of Kai or the energy of Mimi. We’ve even added some music to accompany your adventure! These improvements are designed to make your time in Valoria even more enjoyable. We hope you love them as much as we do! ✨Don’t miss out on this new adventure! Explore all the new features in Season 2 and discover how to dominate the new world of Valoria. Remember, every decision can bring you one step closer to victory. See you in the game!🔥show more

Valiants
81,648 görüntüleme • 2 yıl önce
this effect is all over tiktok right now and... nobody's explaining how to actually do it properly... the 3d balloon character thing. where someone turns into a shiny inflatable version of themselves that still moves and talks. looks pretty smooth in feeds. the workflow is stupid simple once you see it. step 1: take any photo. drop it into an image gen tool (nano banana pro). prompt it with something like "make the person in the photo a plastic blow up balloon character with a shiny surface. keep the face details as 3d balloon details including the person in the background. don't change background" that's it for the image. don't overcomplicate the prompt. shorter = more consistent results. (learned this after wasting like 2 hours trying to get "perfect" prompts that kept giving me garbage) step 2: take that balloon image + your original video and drop both into kling motion control. prompt: "turn the motion and detailed mouth movement of the video to the setting of the image" that's literally it. kling maps the motion from the real video onto the balloon character. mouth moves. head turns. expressions transfer. the whole thing renders in a few minutes. the result looks like a $500 custom animation and costs you maybe $0.30 in kling credits. people are getting 500k+ views with these because the scroll-stop factor is insane. nobody expects to see a shiny inflatable version of someone giving a real speech or doing a product review. the play here is obvious btw. run this for client content (mix with the hook and real body, check the results yourself) or use it on your own faceless channels as a hook pattern before the algo catches up...show more

KNOX
25,773 görüntüleme • 6 ay önce
You can't 3D reconstruct glass from images... ...WRONG! Thanks... for video diffusion, now just about anything is possible! Introducing...Diffusion Knows Transparency (DKT) Transparent and reflective objects usually break robot vision and photogrammetry pipelines because they don't follow the "solid object" rules standard cameras expect. DKT is a new AI model that repurposes the "internal physics engine" found in video generation models to solve this problem. Researchers took a massive video diffusion model (WAN) and fine-tuned it using a custom-built synthetic dataset to turn it into a high-precision depth sensor. To train the AI, they built the first massive synthetic video library of transparent objects, 1.32 million frames of perfectly labeled glass and metal objects in motion. Without ever seeing a "real" labeled video of glass during training, the model (DKT) outperformed all previous specialized systems on real-world benchmarks (ClearPose, DREDS). They created a "lightweight" 1.3B parameter version that runs fast enough (0.17s per frame) to be used on actual robot hardware. Two reasons I find this project important: 1. It further proves that synthetic data will be essential for training the next generation vision models. 2. In real-world robotic tests, using DKT's depth maps nearly doubled the success rate of robot arms trying to pick up objects on tricky reflective or translucent surfaces. At home robots will need to interact with these types of objects on a daily basis. Check out the project page here: Code is LIVE! #Computervision #Robotics #AIshow more

Jonathan Stephens
17,712 görüntüleme • 8 ay önce
Paving the path for abundant verifiable compute is our... mission. We're partnering with Kaito AI 🌊 to allocate 0.25% of the Boundless token supply to creators who help shape the narrative, educate with clarity, and elevate what Boundless makes possible. Boundless is building the foundation for something bigger: scaling and interoperability for every chain through a shared network for verifiable compute. Mainnet is approaching, and we’re moving forward with focus. The next few years will be remembered. Growth should be shared every step of the way. This is not a day 1 startup. This is not an overcrowded market. This is a new category of infrastructure - built on years of research and development, merging cryptography, compute, and the open web. The Signal reminded us how deeply the industry needs this kind of ZK infrastructure. We’re grateful for your support. Kaito has been listening since the earliest days of Boundless, and continues to refine how it identifies the highest-quality creators and content driving our vision forward. Track your progress: Disclaimer: Boundless will be factoring out yappers who create AI slop, disingenuous content or use engagement farming tactics to be on the leaderboard. Not sure where to start? We have prepared plenty of content to help: > Boundless study pack pt1: > Boundless study pack pt2: > How to Yap stream: Intention, effort, and care always stand out. We want to see more of that in the world.show more

Boundless
200,744 görüntüleme • 1 yıl önce