正在加载视频...

视频加载失败

Meet the new Soniox voice library. Soniox Text-to-Speech now includes more than 200 built-in voices, carefully curated for different applications, styles, accents, ages, and voice characteristics. Every voice works across 60+ languages, so you can choose the voice that fits your product and use that same voice globally while...

126,206 次观看 • 13 天前 •via X (Twitter)

3 条评论

Yonatan Gross 的头像
Yonatan Gross13 天前

I love it, I use it for a personal assistant, now I'll have the ability to fire one every couple of days ^^

RAZA | AI EXPLORER 的头像
RAZA | AI EXPLORER12 天前

200+ voices across 60+ languages is a huge upgrade for global voice experiences.

Kumar Deepanshu 的头像
Kumar Deepanshu13 天前

nice 👍

相关视频

Typing just became… Typeless. Meet Typeless 1.0.2 for Mac — a tool that transforms your voice into clear, accurate writing across your Mac. Speak naturally and let Typeless handle the typing, corrections, and structure for you. With Typeless, you can dictate, translate, or ask for quick edits in any language or accent. It converts your speech into polished text up to 10× faster than traditional typing, while automatically fixing mistakes along the way. --- 1️⃣ Dictation Typeless acts as a powerful voice keyboard that works across all applications on your Mac. When you speak, it understands your intent, organizes your ideas, and converts your natural speech into well-structured writing. Whether you're drafting emails, notes, documents, or messages, Typeless helps you turn spoken thoughts into clean text instantly. Controls - Press Fn to start or stop dictation - Hold Fn for quick, short dictation --- 2️⃣ Translation Typeless makes writing in other languages effortless. You can speak in your native language and have Typeless translate your words instantly into the language you want. This allows you to communicate, write, and respond in foreign languages smoothly and naturally. Controls - Press Fn + Space to start translation - Press Fn to stop translation --- 3️⃣ Ask Anything Typeless Typeless also lets you interact with your text using voice commands. You can select any text and simply say how you want it changed. Typeless can edit, rewrite, answer questions, or perform quick actions based on your request, making editing and improving text much faster. Controls - Press Fn + Space to start Ask Anything - Press Fn to stop Ask Anything --- With Typeless, your voice becomes the fastest and easiest way to write, edit, and communicate on your Mac. Your voice is now your keyboard. Get Typeless → Available now on Mac, Windows, iOS, and Android. #Typeless

Kuria Chronicles

43,623 次观看 • 6 个月前

Learn to build conversational AI voice agents in "Building AI Voice Agents for Production", created in collaboration with LiveKit and RealAvatar, and taught by dsa (Co-founder & CEO of LiveKit), Shayne (Developer Advocate, LiveKit), and Nedelina Teneva (Head of AI at RealAvatar, an AI Fund portfolio company). Voice agents combine speech and reasoning capabilities to enable real-time conversations. They're already being used to support customer service, to improve accessibility in healthcare, for entertainment applications, and for talk therapy. In this course, you’ll learn to build voice agents that listen, reason, and respond naturally. You’ll follow the architecture used to create the "AI Andrew" Avatar, a collaborative project between and RealAvatar that responds to users in what sounds like my voice. You’ll build a voice agent from scratch and deploy it to the cloud, enabling support for many simultaneous users. What you’ll learn: - Understand the fundamentals of voice agents, including key components like speech-to-text (STT), text-to-speech (TTS), and LLMs, and how latency is introduced at each layer. - Explore voice agent architectures and the trade-offs between modular pipelines and speech-to-speech APIs. - Explore how platforms like LiveKit mitigate latency issues with optimized networking infrastructure and low-latency communication protocols. - Learn how to connect client devices to voice agents using WebRTC—and why it outperforms HTTP and WebSocket for low-latency audio streaming. - Incorporate voice activity detection (VAD), end-of-turn detection, and context management to detect turns, handle interruptions, and manage conversational flow. - Understand the trade-offs between latency, quality, and cost in an example in which you build a voice agent and change its voice. - Equip your agent with metrics to measure latency at each stage of the voice pipeline and learn the key levers you can pull to make your agent faster and more responsive. The voice agents built in this course also incorporate voice technology from , a supporting contributor to the project. By the end of this course, you'll have learned the components of an AI voice agent pipeline, combined them into a system with low-latency communication, and deployed them on cloud infrastructure so it scales to many users. I’m looking forward to seeing what voice agents you build from this course! Please sign up here:

Andrew Ng

87,868 次观看 • 1 年前