kyutai

Tag

Cards List
#kyutai

@kyutai_labs: Our audio-to-MIDI model, MuScriptor, now also detects tempo! You can directly drag-and-drop the MIDI into a DAW and it …

X AI KOLs Timeline · 2026-08-06 Cached

Kyutai Labs announces that their audio-to-MIDI model MuScriptor now detects tempo, allowing direct drag-and-drop of MIDI into a DAW without manual tempo matching.

0 favorites 0 likes
#kyutai

Kyutai's Pocket TTS clones a voice from 5 seconds of audio, on CPU, under MIT. Benchmarked against Kokoro, Supertonic, and Inflect-Nano for Eng. TTS

Reddit r/LocalLLaMA · 2026-07-06

Kyutai released Pocket TTS, a text-to-speech model capable of cloning a voice from just 5 seconds of audio, running on CPU and released under the MIT license. It was benchmarked against Kokoro, Supertonic, and Inflect-Nano for English TTS.

0 favorites 0 likes
#kyutai

@t0m1ab: Heading to ICML 2026 in Seoul next week with @romfbr31 to present Hibiki-Zero[https://kyutai.org/blog/2026-02-12-hibiki…

X AI KOLs Following · 2026-06-30 Cached

Kyutai presents Hibiki-Zero, a real-time speech-to-speech translation model, at ICML 2026 in Seoul, with an oral presentation scheduled for July 8.

0 favorites 0 likes
#kyutai

Continuous Audio Language Models

Papers with Code Trending · 2025-09-08 Cached

This paper introduces Continuous Audio Language Models (CALM), which generate audio using continuous frames instead of discrete tokens to improve fidelity and reduce computational cost in speech and music generation.

0 favorites 0 likes
← Back to home

Submit Feedback