正在加载视频...

视频加载失败

Introducing Speech Engine. Developers can now turn their existing chat agent into a full voice agent with one prompt. Speech Engine combines our leading speech, transcription, and voice orchestration models into a single pipeline - all custom built to work best together.

135,012 次观看 • 4 个月前 •via X (Twitter)

39 条评论

ElevenLabs 的头像
ElevenLabs4 个月前

Connect to your existing chat agent. Your text-based agent remains untouched - Speech Engine integrates on top of your existing stack so nothing is rearchitected.

ElevenLabs 的头像
ElevenLabs4 个月前

Install with one command using our skill. npx skills add elevenlabs/skills --skill speech-engine The skill sets up everything you need so you can go from chat to voice in a single prompt.

ElevenLabs 的头像
ElevenLabs4 个月前

Add expressive, human-like voices in 70+ languages. Voice is the fastest and richest way to exchange information, making your product and services more accessible.

ElevenLabs 的头像
ElevenLabs4 个月前

Industry-leading transcription. Our transcription models are optimized for conversational use cases, delivering ultra-low latency and built to handle messy, real-world environments.

ElevenLabs 的头像
ElevenLabs4 个月前

Enterprise-grade security. Our platform is designed for deployments at scale with enterprise-level data protections, including support for SOC 2, HIPAA, and GDPR compliance. EU Data Residency and Zero Retention Mode are available for stricter data control.

ElevenLabs 的头像
ElevenLabs4 个月前

Watch the full Speech Engine walkthrough from the @aiDotEngineer conference in London.

ElevenLabs 的头像
ElevenLabs4 个月前

Migrate to ElevenAgents at any time. Get additional deployment channels, monitoring, analytics, and the full suite of agent tools.

ElevenLabs 的头像
ElevenLabs4 个月前

Speech Engine is available now in ElevenAPI. Starting at 8¢ per minute, decreasing with scale.

nickster 的头像
nickster4 个月前

guys you are covered by ai radio!! ai host just picked your release

Rayan A Cader 的头像
Rayan A Cader4 个月前

Whoever did the motion graphics, needs a raise. This thing is crazy good 🔥

Carlos 的头像
Carlos4 个月前

Too bad we can't use any of this on YouTube without getting demonetized! THANKS ALOT! SYNTH ID BS.

Anjali Thacker 的头像
Anjali Thacker4 个月前

🔥🔥

Ken 的头像
Ken3 个月前

Voice agents get interesting when they inherit an existing text workflow instead of replacing it. The shift is chat agent -> callable phone/voice worker.

kiyosaki 的头像
kiyosaki3 个月前

すごい進化ですね!音声エージェントに変えることができるなんて!

Oliver Finn 的头像
Oliver Finn4 个月前

Jarvis? Is that you?

@skully 的头像
@skully4 个月前

camping your page

blcfyp123 的头像
blcfyp1233 个月前

so good :3

Hasan 的头像
Hasan4 个月前

this is straight up ridiculous in the best way elevenlabs made full voice ai stupidly easy with one prompt low latency magic across 70 languages and zero rearchitecting devs are eating good tonight well played team 🔥

Paolo Perazzo 的头像
Paolo Perazzo3 个月前

This is great, but what’s the cost to run it?

Everhett Grimes 的头像
Everhett Grimes4 个月前

Cool

Hermes Agent Tips 的头像
Hermes Agent Tips4 个月前

evolutionary... awesome

Junk 🛠️ 的头像
Junk 🛠️4 个月前

one prompt is doing a lot here

Nischay 的头像
Nischay3 个月前

Voice to agent action is interesting. We took a different angle: speak your intent during a coding session, run /act, and the agent executes it with full screen context already in view.

Aaliya 的头像
Aaliya4 个月前

So now one prompt can also make my chatbot talk more than me in real life 😄

8Zamania 的头像
8Zamania4 个月前

@ElevenLabs es lo más horrible y caro que he visto en la vida, no pierdan su tiempo. ¿Un simple video de una mujer bailando? Pfff

dawood10x 的头像
dawood10x4 个月前

voice was the last thing keeping agents feeling like agents. one prompt to flip the switch is wild. the gap between chat and human-like just collapsed.

Soma 的头像
Soma4 个月前

Voice may be the next real interface shift for agents. Turning a chat agent into a voice agent with one prompt lowers the barrier from “build a product” to “test a new behavior.”

ingrid souza 的头像
ingrid souza3 个月前

新式の会話仲介システム、素晴らしいですね。

Naveen 的头像
Naveen3 个月前

Does Speech Engine support multi-modal interfaces, allowing developers to combine voice, text, and visual inputs for a more seamless user experience?

Nacho Papi 的头像
Nacho Papi3 个月前

check out what I built using @ElevenLabs gimmicky or useful?honest feedback only.

Ramit Koul 的头像
Ramit Koul4 个月前

voice agents need this kind of boring pipeline work. the magic only works if the whole stack is smooth

toni 的头像
toni3 个月前

More models = more complexity, right? Not here. One prompt replaces three separate integrations. Simpler stack, same power.

Hùng Phan 的头像
Hùng Phan3 个月前

Evaluating this for our voice stack. Is $0.08/min flat across all voice models? Agents page suggests bundled, but ElevenAPI shows per-model rates.

Tadas Petra 的头像
Tadas Petra4 个月前

Happy to finally see this live!

Anastasios-Antonios Toulkeridis 的头像
Anastasios-Antonios Toulkeridis4 个月前

so i use OpenAI for the LLM and ElevenLabs for TTS. Are you saying that ElevenLabs can now handle the LLM part also? I'm confused

kayvon! (building murray ai) 的头像
kayvon! (building murray ai)4 个月前

Nice engineering move — lowering the friction to add voice to existing agents is smart. The next layer that matters for production use is when that voice agent has to work reliably in noisy environments, with hands full, and with strong guardrails before anything actually executes. Software layer on top of chat is necessary. Dedicated hardware + controlled handoff is what makes it usable all day.

Vladimir Gusev 的头像
Vladimir Gusev4 个月前

One prompt to voice is a big unlock.

あいり|海外AIニュースを毎日届ける人 的头像
あいり|海外AIニュースを毎日届ける人4 个月前

この内容、日本語で詳しく書きました Wrote a detailed take in Japanese:

Adel Bucetta 的头像
Adel Bucetta4 个月前

the real unlock is turning existing agents into voice ones with a simple prompt. that's huge for devs who already have chat setup, now they can just flip the switch.

相关视频