
AshutoshShrivastava
@ai_for_success • 81,222 subscribers
Post about latest AI news, tools, tutorials and memes.
Shorts
Videos

OpenAI how it started vs how it’s going. Can we really trust Sam Altman or OpenAI with AGI? 🤔 📹 Credit: u/katxwoods
AshutoshShrivastava12,895,574 görüntüleme • 1 yıl önce

Fable 5.1 + Three.js + Cursor 🔥 I built this game, and I can’t even believe how good it turned out. It feels so good to play with a PS5 DualSense controller with haptic enabled. I’m also working on a new stage, The only downside is that the prompt rejection rate is high
AshutoshShrivastava21,351 görüntüleme • 1 gün önce

Grok-3 image editing feature is really underrated. It's incredibly good, you guys should try it out! It's free for everyone. if you combine it with Pika new feature, Pikaframes, it awesome. I created this video in less than 3 minutes. 1/2 Here’s how 👇
AshutoshShrivastava1,947,726 görüntüleme • 1 yıl önce

Grok-3 DeepSearch + Gamma : productivity hack 🔥 - Grok-3: DeepSearch is really powerful use it to pull relevant data and people's opinions from X and the internet on any topic. - Gamma : Turn this data into sleek, AI-generated slides in seconds. step-by-step guide 👇
AshutoshShrivastava1,586,491 görüntüleme • 1 yıl önce

China is on 🔥 ByteDance drops another banger AI paper! OmniHuman-1 can generate realistic human videos at any aspect ratio and body proportion using just a single image and audio. This is the best i have seen so far. 10 incredible examples and the research paper Link👇
AshutoshShrivastava1,549,513 görüntüleme • 1 yıl önce

The improvement in AI-generated videos over the past 15-16 months is remarkable. The first video was released in March or April 2023, while the second video came out in August 2024. The progress is immense, and it's expected to continue in the coming year. Some of the best AI video tools right now include: - Gen-3 Alpha - Kling AI - Luma Labs - Sora (OpenAI, not yet available for everyone)
AshutoshShrivastava1,865,350 görüntüleme • 2 yıl önce

I vibe coded a game called Aether Breach for my son today. Built the whole thing in about 2 hours using Cursor, Grok 4.5, and GPT Image 2.0. He gave it a try and really liked it. I just need to make Easy mode a bit more relaxed so he can make it to the later stages. The full game has DualSense support too. I published my WebHID DualSense npm package a few days ago, so adding controller support to new games is pretty straightforward now. Really happy with how it turned out.
AshutoshShrivastava116,280 görüntüleme • 1 ay önce

xAI is cooking 🔥 They just launched Grok Voice, a platform for building AI voice agents. - You can build voice agents in under 2 minutes - No code required - Human-like conversations with sub-second latency - Supports 25+ languages - 80+ built-in voices or clone your own - Works with Gmail, Google Calendar, Outlook, Notion, MCPs, and more - Starts at $0.05/min and includes a free phone number
AshutoshShrivastava144,824 görüntüleme • 2 ay önce

I love Hermes Agent and Gemini Live is also one of my for voice conversation, so over the weekend I built a small vibe-coding project that brings the two together called IRIS. The idea is simple: Gemini Live becomes the conversational interface, while Hermes Agent handles execution, automation, and long-running tasks behind the scenes. I can talk to Gemini naturally, and whenever something needs real work, Gemini formulates the request and hands it off to Hermes. The best part is that the conversation never stops. Gemini keeps talking with me naturally while Hermes works in the background. When Hermes finishes, Gemini proactively tells me the results are ready, switches context to the Hermes output, helps me review it, and then we return to our normal conversation. That manager/worker flow feels really good. Because I wanted the app to be as hands-free as possible, I also added a camera-based gesture layer using MediaPipe Gesture Recognizer. I can point at Hermes task cards, open details, scroll through results with my hand, and close them using gestures. Small weekend project, but I had a lot of fun building it. Also, the anti-flicker macOS app I vibe-coded yesterday ended up being incredibly useful for this project. It fixed my webcam flickering issue, so I can finally record videos properly. 🙂
AshutoshShrivastava149,739 görüntüleme • 2 ay önce

LMAO 😂 Sam Altman is probably not sleeping well. Kai-Fu Lee (Taiwanese businessman and computer scientist)
AshutoshShrivastava976,214 görüntüleme • 1 yıl önce

Google just launched Gemini 3.5 Transcribe, its most precise speech-to-text model yet. I built this app using Gemini 3.5 Transcribe and have been playing with the model for quite some time. Thanks to the Google DeepMind team for the early access. It’s already available in the Gemini API, AI Studio, and Antigravity, and is now in public preview.
AshutoshShrivastava20,167 görüntüleme • 8 gün önce

🚨You can now use the new upcoming OpenAI model GPT 5.2 inside Cursor. Here is the full walkthrough. - Open the editor, go to settings and then the model tab. Add a custom model and enter the text "gpt-5.2-high" and "gpt-5.2". - After that you can select the model and ask questions. To verify, I started my test on the usage page which had zero gpt-5.2-high requests and consumption. After the test I could see the details in usage and the cost incurred while using it. Enjoy
AshutoshShrivastava424,035 görüntüleme • 8 ay önce

Vibe coded a new project using Gemini 3.7 Flash inside Antigravity CLI over the weekend. It’s a real-time global flight radar and route intelligence platform. Some information is slightly delayed because it’s running entirely on free-tier APIs. - Tracks 4,800+ airborne aircraft live using OpenSky telemetry - 60 FPS 2D/3D canvas rendering with dead-reckoning extrapolation and geodesic route arcs - Local SQLite route caching with strict airframe hex binding to stay within free-tier API quotas - ATC callsign decomposition engine - Emergency squawk monitoring
AshutoshShrivastava22,838 görüntüleme • 11 gün önce

You can run Google new Gemma 4 on mobile easily. I am using Gemma 4 version E2B on my Pixel 10 Pro. Here is all you need to do: - Go to the App Store and install Google AI Edge Gallery. If you already have it, just update it. - From there, you can install the model directly and start using it. Still exploring what this small but powerful model can do :)
AshutoshShrivastava226,599 görüntüleme • 5 ay önce

