
LiveKit
@livekit • 10,597 subscribers
Open source framework and cloud platform for building voice, video, and physical AI agents. https://t.co/OWLvFH82oN
Videos

Today we're launching Expressive mode in LiveKit Agents. Most voice agents deliver every sentence in that same, upbeat register we all know and hate (including bad news). Expressive mode makes your agent sound like it registers what the caller is actually feeling.
LiveKit378,230 görüntüleme • 25 gün önce

Protect sensitive data without giving up voice agent observability. LiveKit PII Redaction automatically removes personal information from transcripts and recordings before it is stored. It is included with LiveKit Agent Observability at no additional cost. Read more:
LiveKit69,544 görüntüleme • 19 gün önce

Introducing Agents UI, an open-source shadcn component library for building polished React frontends for your voice agents. Audio visualizers. Media controls. Session management tools. Chat transcripts. All wired to LiveKit Agents. Install via the shadcn CLI and own the code.
LiveKit183,463 görüntüleme • 6 ay önce

Today we’re launching our first homegrown AI model: an open source turn detection model for building voice agents. Instead of relying solely on voice activity detection (VAD), which only considers when a user is speaking, our model also considers what has and is being said in the context of a conversation and predicts when a user is finished expressing their thoughts before the agent responds. Conversations with AI voice agents using this new model flow much more naturally without constant interruptions from the AI— check it out (more videos, details, and code in the thread):
LiveKit126,961 görüntüleme • 1 yıl önce

Voice agents do not sound robotic because they are slow. They sound robotic because the model writes like an essay and then reads it out loud. We just shared a post on making STT to LLM to TTS sound human. Make the model more human by including ums, sos, real pauses, and even laughter tags. Tiny rhythm changes can make a huge difference.
LiveKit46,010 görüntüleme • 6 ay önce

Gemini 3.1 Flash Live just dropped and it's available with LiveKit today. This is the first Gemini 3 native audio model on the Live API. Better instruction following, improved tool calling, reduced speaker drift, and support for 70+ languages. Audio in, audio out. No text conversion in between.
LiveKit40,277 görüntüleme • 5 ay önce

Introducing LiveKit Inference — a new cloud service that gives you access to the most popular voice AI models with just your LiveKit API key. We manage rate limits for you, report on usage, and consolidate billing. All LiveKit Cloud plans now include free monthly inference credits. A single string update allows you to call models from: AssemblyAI Deepgram Google DeepMind Inworld AI OpenAI Rime
LiveKit37,237 görüntüleme • 11 ay önce

We shipped LiveKit Turn Detector v1. Instead of reading transcripts, it listens to speech directly, combining semantic and acoustic cues into one end-of-turn prediction. The result: high accuracy, low latency—the best model we tested across 14 languages. Available on LiveKit Cloud.
LiveKit12,053 görüntüleme • 2 ay önce

Voice cloning is now available on LiveKit Inference. We’re launching with Inworld AI and Cartesia. Clone a voice once and use it across multiple TTS providers, with automatic fallback to the same voice if a provider fails mid-call. Free to create and available on all paid plans today.
LiveKit11,426 görüntüleme • 4 ay önce

Add a face to your voice agent. LiveAvatar by HeyGen is now supported in LiveKit Agents. Add a realtime human avatar to your agent without rebuilding the conversation loop. Your LiveKit agent still owns the room, turn-taking, model orchestration, and voice pipeline. LiveAvatar renders the synchronized face and video stream. Useful for product demos, onboarding, tutoring, and support agents that need a visual layer.
LiveKit10,885 görüntüleme • 4 ay önce
Daha fazla içerik yok.
