
Superman
@thesupermanmx • 9,588 subscribers
AI & Neuroscience
Shorts
Videos

someone literally vibe-coded "twitch" for running 🤯 you stream your run live and people watch your dot move across a 3d map of the city in real time. → live 3d map tracking your exact route → pace, distance, heart rate updating live on screen → live graphs for elevation and heart rate → viewers chat with you mid-run and yell at you to push up the hill 100% free..
Superman1,507,468 просмотров • 2 дней назад

China open-sourced a model that reconstructs any scene in 3D from a regular video, in real-time. one camera. no LiDAR. 10,000+ frames without falling apart. just walk around with your camera and watch the entire world get rebuilt in 3D at 20 fps. → runs at ~20 FPS on a single GPU → Stable over 10,000+ frames → Beats optimization-based methods on benchmarks → Works on drone footage, driving videos, indoor walkthroughs 100% open source.
Superman1,453,932 просмотров • 1 месяц назад

someone open-sourced a tool that turns any phone video into a full 3D world in real time. you record a walk around your house, drop the mp4 in, and the entire scene comes back as a walkable 996k-point world at 60fps on your local GPU. no cloud. no subscription. runs 100% locally. 100% open source.
Superman817,989 просмотров • 1 месяц назад

Microsoft open-sourced a 4B model that turns any image into a production-ready 3D asset in 3 seconds. It’s called TRELLIS.2, a fully textured, physically accurate 3D models with PBR textures out of the box. → Full PBR (base color, roughness, metallic, opacity) → Handles hair, cloth, glass, non-manifold geometry → Exports .glb ready for Unity/Unreal/Blender → Runs locally, ships in 3 seconds It's not a demo or a research preview. The full training codebase is public. You can fine-tune it on your own asset library and get a model that generates in your studio's exact style. 100% Open Source
Superman282,217 просмотров • 1 месяц назад

ssomeone vibe-coded a video stream that is secretly 100% text so it can't be blocked. It’s called ASCIline. It renders streams 360p video at 30fps with zero video element on the page. every frame is colored text characters painted on a canvas. 100% Open Source.
Superman53,731 просмотров • 1 месяц назад

Someone open-sourced a 66M parameter model that beats ElevenLabs, OpenAI, and Gemini and runs entirely offline on a Raspberry Pi. it's called Supertonic, a text-to-speech model so fast it turns an entire webpage into audio in under 1 second. locally. offline. for free. text-to-speech has lived in the cloud for years. every spoken character was an API call and a fraction of a cent. Supertonic 3 kills that entirely. the model is 99M parameters. ships as a single ONNX file. hits 167× faster than real-time on a laptop CPU. that's ~1,263 characters of speech per second. other open systems sit at 55–287. no cloud. no API. no GPU. → 31 languages, no separate adapters → works inside a browser tab (WebGPU/WASM) → handles phone numbers, currency, dates — no preprocessing → inline tags for laugh and breath → Turns an entire webpage into audio in under 1 second → 44.1kHz studio-grade WAV, no upsampler needed` and on real-world text, "$5.2M", "(212) 555-0142 ext. 402", "30kph”, it's the only one that reads them correctly. ElevenLabs Flash, OpenAI TTS-1, Gemini 2.5 Flash all fail. Supertonic passes. 11.2k stars. 100% Open Source.
Superman30,416 просмотров • 1 месяц назад
Больше нет контента для загрузки