正在加载视频...

视频加载失败

Further tinkering with my little French tutor app. This version is using the Gemini Multimodal Live API. The speech understanding in Gemini is quite something. In this video you can see Gemini correcting my pronunciation. (Very patiently.) The language tutor use case really highlights the strengths of a next-generation...

13,005 次观看 • 1 年前 •via X (Twitter)

12 条评论

kwindla 的头像
kwindla1 年前

Code is here: Here are the docs for Pipecat's open source client SDKs for javascript, React, iOS, Android, React Native, and C++: The Pipecat clients all support WebRTC, WebSocket, and HTTP network transports.

Nexus 👾 的头像
Nexus 👾1 年前

Nexus combines CMC’s market insights with Zerion and Debank’s wallet tracking—all in one easy-to-use dashboard. Simplify your crypto journey today, completely free 🚀 Try it now!

The Canaanite 的头像
The Canaanite1 年前

bravo, excellente prononciation monsieur 🤣

عای‌ خان 的头像
عای‌ خان1 年前

Amazing. I need to code up a Maths tutor for my kids for those inevitable times where my patience sadly failed. 😅

Emanuel Perez 的头像
Emanuel Perez1 年前

Whoa

Samuel Ekpe 的头像
Samuel Ekpe1 年前

interesting

Mitja Martini 的头像
Mitja Martini1 年前

Thanks for sharing this very nice use case! I've been using ChatGPT with my daughter to help her learning Latin, not in voice mode, though :)

Nicolas MICHEL 的头像
Nicolas MICHEL1 年前

it is funny, the machine has an African accent. This will lead to some funny situations if you start speaking in public with this accent 😂

Jake 的头像
Jake1 年前

cool any other interesting uncommon use cases? I feel like the voice use cases are going to blow up this year

kwindla 的头像
kwindla1 年前

> any other interesting uncommon use cases? Check out the voice interaction experiments that @trudypainter has been posting!

Dan Goodman 🍊 的头像
Dan Goodman 🍊1 年前

Can you have it take in audio but output text in pipecat?

kwindla 的头像
kwindla1 年前

> Can you have it take in audio but output text in pipecat? Yes!

相关视频