Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

I've built an async Voice Agent for WhatsApp 💬 It's simple, yet makes for a powerful experience! → Scribe transcribes voice message → AI SDK generates text response → ElevenLabs generates speech in Opus format → Hono & Cloudflare write file to R2 → twilio serves files to WhatsApp...

26,288 Aufrufe • vor 1 Jahr •via X (Twitter)

11 Kommentare

Profilbild von Thor 雷神 ⚡️
Thor 雷神 ⚡️vor 1 Jahr

You can find the code here:

Profilbild von Alex ZAP
Alex ZAPvor 1 Jahr

🚀 Revolutionize your QA testing with ZAPTEST AI! ZAPTEST automates testing, slashes costs & boosts ROI up to 10X. No coding needed—just faster, smarter automation. 📅 Book your live demo now:

Profilbild von angelo
angelovor 1 Jahr

@elevenlabsio @aisdk @honojs @Cloudflare now i can answer my mum's voice notes

Profilbild von Luke Harries
Luke Harriesvor 1 Jahr

@elevenlabsio @aisdk @honojs @Cloudflare Love this! Native feature for Conversational AI soon too?!

Profilbild von Thor 雷神 ⚡️
Thor 雷神 ⚡️vor 1 Jahr

@elevenlabsio @aisdk @honojs @Cloudflare Sounds about right 🫡

Profilbild von Ray Fernando
Ray Fernandovor 1 Jahr

@elevenlabsio @aisdk @honojs @Cloudflare Yooooooo 👀👀👀 cooking up a code with AirPods workflow. Exactly what I needed RN.

Profilbild von Thor 雷神 ⚡️
Thor 雷神 ⚡️vor 1 Jahr

@elevenlabsio @aisdk @honojs @Cloudflare Yay, awesome, please help yourself to the code 🙌

Profilbild von Rexan Wong
Rexan Wongvor 1 Jahr

@elevenlabsio @aisdk @honojs @Cloudflare bro is boom boom bang different sexy tech tgt and boom he has a crazy product

Profilbild von Thor 雷神 ⚡️
Thor 雷神 ⚡️vor 1 Jahr

@elevenlabsio @aisdk @honojs @Cloudflare That's why they call me boomer! 💥

Profilbild von Ash
Ashvor 1 Jahr

@elevenlabsio @aisdk @honojs @Cloudflare Waiting till you eventually build out your own version of Jarvis now 😂

Profilbild von Thor 雷神 ⚡️
Thor 雷神 ⚡️vor 1 Jahr

@elevenlabsio @aisdk @honojs @Cloudflare Imagine simply ordering a pizza from a voice note! 🤯

Ähnliche Videos

Learn to build conversational AI voice agents in "Building AI Voice Agents for Production", created in collaboration with LiveKit and RealAvatar, and taught by dsa (Co-founder & CEO of LiveKit), Shayne (Developer Advocate, LiveKit), and Nedelina Teneva (Head of AI at RealAvatar, an AI Fund portfolio company). Voice agents combine speech and reasoning capabilities to enable real-time conversations. They're already being used to support customer service, to improve accessibility in healthcare, for entertainment applications, and for talk therapy. In this course, you’ll learn to build voice agents that listen, reason, and respond naturally. You’ll follow the architecture used to create the "AI Andrew" Avatar, a collaborative project between and RealAvatar that responds to users in what sounds like my voice. You’ll build a voice agent from scratch and deploy it to the cloud, enabling support for many simultaneous users. What you’ll learn: - Understand the fundamentals of voice agents, including key components like speech-to-text (STT), text-to-speech (TTS), and LLMs, and how latency is introduced at each layer. - Explore voice agent architectures and the trade-offs between modular pipelines and speech-to-speech APIs. - Explore how platforms like LiveKit mitigate latency issues with optimized networking infrastructure and low-latency communication protocols. - Learn how to connect client devices to voice agents using WebRTC—and why it outperforms HTTP and WebSocket for low-latency audio streaming. - Incorporate voice activity detection (VAD), end-of-turn detection, and context management to detect turns, handle interruptions, and manage conversational flow. - Understand the trade-offs between latency, quality, and cost in an example in which you build a voice agent and change its voice. - Equip your agent with metrics to measure latency at each stage of the voice pipeline and learn the key levers you can pull to make your agent faster and more responsive. The voice agents built in this course also incorporate voice technology from , a supporting contributor to the project. By the end of this course, you'll have learned the components of an AI voice agent pipeline, combined them into a system with low-latency communication, and deployed them on cloud infrastructure so it scales to many users. I’m looking forward to seeing what voice agents you build from this course! Please sign up here:

Andrew Ng

87,711 Aufrufe • vor 1 Jahr