Artwork

内容由Changelog Media提供。所有播客内容(包括剧集、图形和播客描述)均由 Changelog Media 或其播客平台合作伙伴直接上传和提供。如果您认为有人在未经您许可的情况下使用您的受版权保护的作品,您可以按照此处概述的流程进行操作https://zh.player.fm/legal
Player FM -播客应用
使用Player FM应用程序离线!

Full-duplex, real-time dialogue with Kyutai

50:05
 
分享
 

Manage episode 453647338 series 2385063
内容由Changelog Media提供。所有播客内容(包括剧集、图形和播客描述)均由 Changelog Media 或其播客平台合作伙伴直接上传和提供。如果您认为有人在未经您许可的情况下使用您的受版权保护的作品,您可以按照此处概述的流程进行操作https://zh.player.fm/legal

Kyutai, an open science research lab, made headlines over the summer when they released their real-time speech-to-speech AI assistant (beating OpenAI to market with their teased GPT-driven speech-to-speech functionality). Alex from Kyutai joins us in this episode to discuss the research lab, their recent Moshi models, and what might be coming next from the lab. Along the way we discuss small models and the AI ecosystem in France.

Join the discussion

Changelog++ members save 10 minutes on this episode because they made the ads disappear. Join today!

Sponsors:

  • Fly.ioThe home of Changelog.com — Deploy your apps close to your users — global Anycast load-balancing, zero-configuration private networking, hardware isolation, and instant WireGuard VPN connections. Push-button deployments that scale to thousands of instances. Check out the speedrun to get started in minutes.
  • TimescalePurpose-built performance for AI Build RAG, search, and AI agents on the cloud and with PostgreSQL and purpose-built extensions for AI: pgvector, pgvectorscale, and pgai.
  • WorkOSAuthKit offers 1,000,000 monthly active users (MAU) free — The world’s best login box, powered by WorkOS + Radix. Learn more and get started at WorkOS.com and AuthKit.com

Featuring:

Show Notes:

Something missing or broken? PRs welcome!

  continue reading

章节

1. Welcome to Practical AI (00:00:00)

2. Sponsor: Fly (00:00:35)

3. What is Kyutai? (00:03:16)

4. French AI ecosystem (00:05:59)

5. Formin a non-profit (00:08:41)

6. Connecting to open science (00:10:31)

7. What makes Kyutai stand out? (00:12:28)

8. Sponsor: Timescale (00:16:26)

9. Moshi's capabilities (00:19:04)

10. History of speech-to-speech models (00:22:58)

11. Cool things to try (00:30:53)

12. Sponsor: WorkOS (00:34:13)

13. Fine tuning data sets (00:37:13)

14. Model sizes (00:42:28)

15. Things to come (00:45:10)

16. Thanks for joining us! (00:48:34)

17. Outro (00:49:16)

302集单集

Artwork
icon分享
 
Manage episode 453647338 series 2385063
内容由Changelog Media提供。所有播客内容(包括剧集、图形和播客描述)均由 Changelog Media 或其播客平台合作伙伴直接上传和提供。如果您认为有人在未经您许可的情况下使用您的受版权保护的作品,您可以按照此处概述的流程进行操作https://zh.player.fm/legal

Kyutai, an open science research lab, made headlines over the summer when they released their real-time speech-to-speech AI assistant (beating OpenAI to market with their teased GPT-driven speech-to-speech functionality). Alex from Kyutai joins us in this episode to discuss the research lab, their recent Moshi models, and what might be coming next from the lab. Along the way we discuss small models and the AI ecosystem in France.

Join the discussion

Changelog++ members save 10 minutes on this episode because they made the ads disappear. Join today!

Sponsors:

  • Fly.ioThe home of Changelog.com — Deploy your apps close to your users — global Anycast load-balancing, zero-configuration private networking, hardware isolation, and instant WireGuard VPN connections. Push-button deployments that scale to thousands of instances. Check out the speedrun to get started in minutes.
  • TimescalePurpose-built performance for AI Build RAG, search, and AI agents on the cloud and with PostgreSQL and purpose-built extensions for AI: pgvector, pgvectorscale, and pgai.
  • WorkOSAuthKit offers 1,000,000 monthly active users (MAU) free — The world’s best login box, powered by WorkOS + Radix. Learn more and get started at WorkOS.com and AuthKit.com

Featuring:

Show Notes:

Something missing or broken? PRs welcome!

  continue reading

章节

1. Welcome to Practical AI (00:00:00)

2. Sponsor: Fly (00:00:35)

3. What is Kyutai? (00:03:16)

4. French AI ecosystem (00:05:59)

5. Formin a non-profit (00:08:41)

6. Connecting to open science (00:10:31)

7. What makes Kyutai stand out? (00:12:28)

8. Sponsor: Timescale (00:16:26)

9. Moshi's capabilities (00:19:04)

10. History of speech-to-speech models (00:22:58)

11. Cool things to try (00:30:53)

12. Sponsor: WorkOS (00:34:13)

13. Fine tuning data sets (00:37:13)

14. Model sizes (00:42:28)

15. Things to come (00:45:10)

16. Thanks for joining us! (00:48:34)

17. Outro (00:49:16)

302集单集

Alle episoder

×
 
Loading …

欢迎使用Player FM

Player FM正在网上搜索高质量的播客,以便您现在享受。它是最好的播客应用程序,适用于安卓、iPhone和网络。注册以跨设备同步订阅。

 

快速参考指南

边探索边听这个节目
播放