Artwork

Changelog Media에서 제공하는 콘텐츠입니다. 에피소드, 그래픽, 팟캐스트 설명을 포함한 모든 팟캐스트 콘텐츠는 Changelog Media 또는 해당 팟캐스트 플랫폼 파트너가 직접 업로드하고 제공합니다. 누군가가 귀하의 허락 없이 귀하의 저작물을 사용하고 있다고 생각되는 경우 여기에 설명된 절차를 따르실 수 있습니다 https://ko.player.fm/legal.
Player FM -팟 캐스트 앱
Player FM 앱으로 오프라인으로 전환하세요!

Full-duplex, real-time dialogue with Kyutai

50:05
 
공유
 

Manage episode 453647338 series 2385063
Changelog Media에서 제공하는 콘텐츠입니다. 에피소드, 그래픽, 팟캐스트 설명을 포함한 모든 팟캐스트 콘텐츠는 Changelog Media 또는 해당 팟캐스트 플랫폼 파트너가 직접 업로드하고 제공합니다. 누군가가 귀하의 허락 없이 귀하의 저작물을 사용하고 있다고 생각되는 경우 여기에 설명된 절차를 따르실 수 있습니다 https://ko.player.fm/legal.

Kyutai, an open science research lab, made headlines over the summer when they released their real-time speech-to-speech AI assistant (beating OpenAI to market with their teased GPT-driven speech-to-speech functionality). Alex from Kyutai joins us in this episode to discuss the research lab, their recent Moshi models, and what might be coming next from the lab. Along the way we discuss small models and the AI ecosystem in France.

Join the discussion

Changelog++ members save 10 minutes on this episode because they made the ads disappear. Join today!

Sponsors:

  • Fly.ioThe home of Changelog.com — Deploy your apps close to your users — global Anycast load-balancing, zero-configuration private networking, hardware isolation, and instant WireGuard VPN connections. Push-button deployments that scale to thousands of instances. Check out the speedrun to get started in minutes.
  • TimescalePurpose-built performance for AI Build RAG, search, and AI agents on the cloud and with PostgreSQL and purpose-built extensions for AI: pgvector, pgvectorscale, and pgai.
  • WorkOSAuthKit offers 1,000,000 monthly active users (MAU) free — The world’s best login box, powered by WorkOS + Radix. Learn more and get started at WorkOS.com and AuthKit.com

Featuring:

Show Notes:

Something missing or broken? PRs welcome!

  continue reading

챕터

1. Welcome to Practical AI (00:00:00)

2. Sponsor: Fly (00:00:35)

3. What is Kyutai? (00:03:16)

4. French AI ecosystem (00:05:59)

5. Formin a non-profit (00:08:41)

6. Connecting to open science (00:10:31)

7. What makes Kyutai stand out? (00:12:28)

8. Sponsor: Timescale (00:16:26)

9. Moshi's capabilities (00:19:04)

10. History of speech-to-speech models (00:22:58)

11. Cool things to try (00:30:53)

12. Sponsor: WorkOS (00:34:13)

13. Fine tuning data sets (00:37:13)

14. Model sizes (00:42:28)

15. Things to come (00:45:10)

16. Thanks for joining us! (00:48:34)

17. Outro (00:49:16)

302 에피소드

Artwork
icon공유
 
Manage episode 453647338 series 2385063
Changelog Media에서 제공하는 콘텐츠입니다. 에피소드, 그래픽, 팟캐스트 설명을 포함한 모든 팟캐스트 콘텐츠는 Changelog Media 또는 해당 팟캐스트 플랫폼 파트너가 직접 업로드하고 제공합니다. 누군가가 귀하의 허락 없이 귀하의 저작물을 사용하고 있다고 생각되는 경우 여기에 설명된 절차를 따르실 수 있습니다 https://ko.player.fm/legal.

Kyutai, an open science research lab, made headlines over the summer when they released their real-time speech-to-speech AI assistant (beating OpenAI to market with their teased GPT-driven speech-to-speech functionality). Alex from Kyutai joins us in this episode to discuss the research lab, their recent Moshi models, and what might be coming next from the lab. Along the way we discuss small models and the AI ecosystem in France.

Join the discussion

Changelog++ members save 10 minutes on this episode because they made the ads disappear. Join today!

Sponsors:

  • Fly.ioThe home of Changelog.com — Deploy your apps close to your users — global Anycast load-balancing, zero-configuration private networking, hardware isolation, and instant WireGuard VPN connections. Push-button deployments that scale to thousands of instances. Check out the speedrun to get started in minutes.
  • TimescalePurpose-built performance for AI Build RAG, search, and AI agents on the cloud and with PostgreSQL and purpose-built extensions for AI: pgvector, pgvectorscale, and pgai.
  • WorkOSAuthKit offers 1,000,000 monthly active users (MAU) free — The world’s best login box, powered by WorkOS + Radix. Learn more and get started at WorkOS.com and AuthKit.com

Featuring:

Show Notes:

Something missing or broken? PRs welcome!

  continue reading

챕터

1. Welcome to Practical AI (00:00:00)

2. Sponsor: Fly (00:00:35)

3. What is Kyutai? (00:03:16)

4. French AI ecosystem (00:05:59)

5. Formin a non-profit (00:08:41)

6. Connecting to open science (00:10:31)

7. What makes Kyutai stand out? (00:12:28)

8. Sponsor: Timescale (00:16:26)

9. Moshi's capabilities (00:19:04)

10. History of speech-to-speech models (00:22:58)

11. Cool things to try (00:30:53)

12. Sponsor: WorkOS (00:34:13)

13. Fine tuning data sets (00:37:13)

14. Model sizes (00:42:28)

15. Things to come (00:45:10)

16. Thanks for joining us! (00:48:34)

17. Outro (00:49:16)

302 에피소드

모든 에피소드

×
 
Loading …

플레이어 FM에 오신것을 환영합니다!

플레이어 FM은 웹에서 고품질 팟캐스트를 검색하여 지금 바로 즐길 수 있도록 합니다. 최고의 팟캐스트 앱이며 Android, iPhone 및 웹에서도 작동합니다. 장치 간 구독 동기화를 위해 가입하세요.

 

빠른 참조 가이드