
Superman
@thesupermanmx • 9,717 subscribers
AI & Neuroscience
Shorts
Videos

someone literally vibe-coded "twitch" for running 🤯 you stream your run live and people watch your dot move across a 3d map of the city in real time. → live 3d map tracking your exact route → pace, distance, heart rate updating live on screen → live graphs for elevation and heart rate → viewers chat with you mid-run and yell at you to push up the hill 100% free..
Superman1,596,803 次观看 • 2 天前

China open-sourced a model that reconstructs any scene in 3D from a regular video, in real-time. one camera. no LiDAR. 10,000+ frames without falling apart. just walk around with your camera and watch the entire world get rebuilt in 3D at 20 fps. → runs at ~20 FPS on a single GPU → Stable over 10,000+ frames → Beats optimization-based methods on benchmarks → Works on drone footage, driving videos, indoor walkthroughs 100% open source.
Superman1,453,932 次观看 • 1 个月前

someone open-sourced a tool that turns any phone video into a full 3D world in real time. you record a walk around your house, drop the mp4 in, and the entire scene comes back as a walkable 996k-point world at 60fps on your local GPU. no cloud. no subscription. runs 100% locally. 100% open source.
Superman817,989 次观看 • 1 个月前

Microsoft open-sourced a 4B model that turns any image into a production-ready 3D asset in 3 seconds. It’s called TRELLIS.2, a fully textured, physically accurate 3D models with PBR textures out of the box. → Full PBR (base color, roughness, metallic, opacity) → Handles hair, cloth, glass, non-manifold geometry → Exports .glb ready for Unity/Unreal/Blender → Runs locally, ships in 3 seconds It's not a demo or a research preview. The full training codebase is public. You can fine-tune it on your own asset library and get a model that generates in your studio's exact style. 100% Open Source
Superman282,217 次观看 • 1 个月前

Someone open-sourced a 66M parameter model that beats ElevenLabs, OpenAI, and Gemini and runs entirely offline on a Raspberry Pi. it's called Supertonic, a text-to-speech model so fast it turns an entire webpage into audio in under 1 second. locally. offline. for free. text-to-speech has lived in the cloud for years. every spoken character was an API call and a fraction of a cent. Supertonic 3 kills that entirely. the model is 99M parameters. ships as a single ONNX file. hits 167× faster than real-time on a laptop CPU. that's ~1,263 characters of speech per second. other open systems sit at 55–287. no cloud. no API. no GPU. → 31 languages, no separate adapters → works inside a browser tab (WebGPU/WASM) → handles phone numbers, currency, dates — no preprocessing → inline tags for laugh and breath → Turns an entire webpage into audio in under 1 second → 44.1kHz studio-grade WAV, no upsampler needed` and on real-world text, "$5.2M", "(212) 555-0142 ext. 402", "30kph”, it's the only one that reads them correctly. ElevenLabs Flash, OpenAI TTS-1, Gemini 2.5 Flash all fail. Supertonic passes. 11.2k stars. 100% Open Source.
Superman30,416 次观看 • 1 个月前
没有更多内容可加载