
Vaibhav (VB) Srivastav
@reach_vb • 56,737 subscribers
founder mode @OpenAI | ex @huggingface | F1 fan | *opinions my own
Shorts
Videos

Starting today you can use Codex in Claude Code 👀 /plugin marketplace add openai/codex-plugin-cc Try it out today with: /codex:review for a normal read-only Codex review /codex:adversarial-review for a steerable challenge review /codex:rescue to let codex rescue your code Enjoy Codex-ing!
Vaibhav (VB) Srivastav956,233 views • 5 months ago

Excited to announce OpenAI Build Week! A week full of excitement, joy and dedication to pushing what's possible with GPT 5.6 and Codex We're introducing a Global challenge starting July 13 and over 40+ events across the globe to bring this experience closer to you Kudos to our Codex Ambassadors for putting together such tremendous experiences across the planet from SF to Tokyo in the coming week!! Sign up now and excited to see what you build!
Vaibhav (VB) Srivastav54,140 views • 1 month ago

GPT 5.6 Sol is pretty GOATED at making games, it made an entire rollercoaster simulator with textures and assets included! It all started with a simple prompt in a /goal and with few artistic directions from my side, it's still not complete and there are some UI inconsistencies But you can go pretty far with a broad idea and feedback along the way! Goes to say, you're truly bounded by your own ambition Try it out yourself 🤗
Vaibhav (VB) Srivastav49,923 views • 1 month ago

With ChatGPT Work, you can kick off cloud tasks directly from the desktop or mobile app. It can use your configured plugins to work across your repos, docs, Slack, and the other tools you already use. I use it constantly from my phone: hand off a task, let it run asynchronously in the cloud, and come back when the work is done.
Vaibhav (VB) Srivastav37,684 views • 1 month ago

HOLY FUCK! Zyphra just dropped Zonos - Apache 2.0 licensed, Multilingual, Text to Speech model with INSTANT voice cloning! 🔥 > Zero-shot TTS with Voice Cloning: Input text and a 10-30 second speaker sample to generate high-quality text-to-speech output > Audio Prefix Inputs: Enhance speaker matching by adding an audio prefix to the text, enabling behaviors like whispering that are hard to achieve with voice cloning alone > Multilingual Support: Supports English, Japanese, Chinese, French, and German > Audio Quality & Emotion Control: Fine-tune speaking rate, pitch, frequency, audio quality, and emotions (e.g., happiness, anger, sadness, fear) > Fast Performance: Runs at ~2x real-time speed on an RTX 4090 > Available on the Hugging Face Hub 🤗
Vaibhav (VB) Srivastav299,066 views • 1 year ago

Fuck it! You can now run *any* GGUF on the Hugging Face Hub directly with ollama 🔥 This has been a constant ask from the community, starting today you can point to any of the 45,000 GGUF repos on the Hub* *Without any changes whatsoever! ⚡ All you need to do is: ollama run hf. co/{username}/{reponame}:latest For example, to run the Llama 3.2 1B, you can run: ollama run hf. co/bartowski/Llama-3.2-1B-Instruct-GGUF:latest If you want to run a specific quant, all you need to do is specify the Quant type: ollama run hf. co/bartowski/Llama-3.2-1B-Instruct-GGUF:Q8_0 That's it! We'll work closely with Ollama to continue developing this further! ⚡
Vaibhav (VB) Srivastav317,790 views • 1 year ago

Super excited to host our first OpenAI Developer Office Hours tomorrow! We’ll cover everything new across Codex and the OpenAI platform - /goal, mobile, plugins, Amazon Bedrock and more! Followed by a live AMA. Come with questions. Come with ideas. Come one, come all!
Vaibhav (VB) Srivastav48,003 views • 2 months ago

Fuck it, 685B parameter, DeepSeek V3 0324 running locally on M3 Ultra, fully private 🔥 Powered by llama.cpp & dynamic quants from Unsloth AI ⚡ Step 1: brew install llama.cpp Step 2: llama-cli -hf unsloth/DeepSeek-V3-0324-GGUF:Q2_K_XL That's it! 🤗 Honestly a bit surreal to be able to chat with such a chunky model at the touch of the keyboard - future is going to be wild!!
Vaibhav (VB) Srivastav168,225 views • 1 year ago

Excited to announce the Codex App: run multiple projects and threads in one focused app! 🔥 The app natively packs a lot of features making it easier to maximise your productivity: > Worktree mode keeps changes isolated - parallel tasks without touching your checkout > Automations run in background worktrees and drop findings into your inbox > Built‑in Git review: diff, stage/revert hunks, inline comments > Integrated terminal for test, lint, git - no need to switch apps > Voice dictation: hold Ctrl+M and speak your prompt > Skills + slash commands for faster workflows. > IDE sync with auto context - ask about files you’re viewing > Local / Worktree / Cloud modes - choose where tasks run > Shared MCP config across app/CLI/IDE Bonus: For a limited time, we've doubled the rate limits across the tiers from Free all the way to Enterprise! Enjoy! 🤗
Vaibhav (VB) Srivastav74,106 views • 7 months ago

MARS5 TTS: Open Source Text to Speech with insane prosodic control! 🔥 > Voice cloning with less than 5 seconds of audio > Two stage Auto-Regressive (750M) + Non-Auto Regressive (450M) model architecture > Used BPE tokenizer to enable control over punctuations, pauses, stops etc. > AR model predicts L0 coarse tokens, refined further by the NAR DDPM model followed by the vocoder Great job Camb AI team! Kudos for open sourcing the artifacts - looking forward to what comes next ;)
Vaibhav (VB) Srivastav162,281 views • 2 years ago

Kyutai released their Streaming Text to Speech model, ~2B param model, ultra low latency (220ms), CC-BY-4.0 license 🔥 Trained on 2.5 Million Hours of audio, it can serve up to 32 users w/ less than 350ms latency on a SINGLE L40 🤯 Incredible release by kyutai folks, go check out their hugging face page now!
Vaibhav (VB) Srivastav93,586 views • 1 year ago