Top stories
Top stories
Top stories
Kyutai — forck
The
last 8 weeks
3
posts
·
First-party
·
peak 2
0
outlets reported
·
On 0 posts
@kyutai_labs
·
RSS
Lead story
Kyutai
· 11 days ago
Diffusable Latents from Structure-Agnostic Distillation
|
|
@kyutai_labs
RSS
Announcements
Chatter
Kyutai
Announcements
Chatter
Models
The week
Connect your AI
Announcements
Chatter
Companies
Perplexity
Anthropic
Luma
Runway
Apple
Google Research
GitHub
Kyutai
The
last 8 weeks
3
posts
·
First-party
·
peak 2
0
outlets reported
·
On 0 posts
Pool, then distill: one vector per image is enough to make a latent diffusable, whatever its shape.
Everything else, by date
Building a Pocket TTS with a drifting objective
28 Sep
Pocket TTS training code released
25 Aug
MuScriptor: Automatic Multi-instrument Transcription
10 Jul
Surflo: Consistent 3D surfaces from a global state
1 Jul
The FID Lottery
18 Jun
Post-training speech models for better interactivity
10 Jun
Kairos: Understanding Data Temporality Impact on LLM pre-training
26 May
Introducing KE:SAI
20 May
Pocket TTS now supports six languages
4 May
MoshiRAG: Asynchronous Knowledge Retrieval for Full-Duplex Speech Language Models
30 Apr
ARC-Encoder: learning compressed text representations for LLMs
28 Apr
OVIE: One View Is Enough
14 Apr
+16 more posts
Show fewer
Invincible Voice online demo released
24 Feb
Hibiki-Zero: Simultaneous Speech-to-Speech Translation Without Aligned Data
12 Feb
Pocket TTS: a high-quality TTS with voice cloning that runs on CPU
13 Jan
CASA: The return of the cross-attention
23 Dec
Neural audio codecs: how to get audio into LLMs
21 Oct
Kyutai TTS 1.6B
3 Jul
Kyutai TTS and Unmute now open-source
3 Jul
Kyutai Speech-To-Text released as open-source
19 Jun
Unmute: Make LLMs listen and speak
22 May
Helium 1: a modular and multilingual LLM
30 Apr
MoshiVis: Teaching Moshi to Converse about Images
21 Mar
Simultaneous, on-device, high fidelity speech-to-speech translation with Hibiki
10 Feb
Announcing Helium-1 Preview
13 Jan
Moshi open-source release: run Moshi locally!
18 Sep
Meet Moshi, the first real-time voice AI
3 Jul
Hello Kyutai!
17 Nov