Загрузка видео...

Не удалось загрузить видео

На главную

Kyutai Speech-To-Text is now open-source! It’s streaming, supports batched inference, and runs blazingly fast: perfect for interactive applications. Check out the details here:

66,503 просмотров • 1 год назад •via X (Twitter)

Комментарии: 9

Фото профиля kyutai
kyutai1 год назад

Today we are releasing two models. The first one is a 2.6B English-only model that beats Whisper Large v3 on benchmarks even though it’s a streaming model that doesn’t process all the audio at once. It can process 400 sequences in parallel on a single H100.

Фото профиля kyutai
kyutai1 год назад

The other model is a lightweight English/French 1B model optimized for real-time voice chat apps like It comes with a semantic voice activity detector that predicts if you’re done talking or just pausing mid-sentence. The open-source releases of Kyutai Text-To-Speech and will follow soon!

Фото профиля clem 🤗
clem 🤗1 год назад

Magnifique !

Фото профиля Alex Volkov (Thursd/AI)
Alex Volkov (Thursd/AI)1 год назад

This is great!! Well cover on @thursdai_pod on an hour

Фото профиля @gerry
@gerry1 год назад

That is really good. Well done :)

Фото профиля Dan Western
Dan Western1 год назад

Interesting... Great conversation with this ai. Wondering about potential opportunities to embed this functionality into apps...

Фото профиля karai
karai1 год назад

It needs mooore languages

Фото профиля ratwell
ratwell1 год назад

@dankvr finally

Фото профиля Simon Icard 
Simon Icard 1 год назад

👏

Похожие видео