@honualx: Music is missing its "openai-whisper". It's time that we can turn any music into notes, the same way transcribing speec…
Summary
Kyutai releases MuScriptor, an open model for multi-instrument transcription that converts any music recording into MIDI notes, similar to OpenAI Whisper for speech.
View Cached Full Text
Cached at: 07/12/26, 11:03 PM
Music is missing its “openai-whisper”. It’s time that we can turn any music into notes, the same way transcribing speech is now a given. Are we there yet? Let us know after testing MuScriptor 🎶 Work by @simonrouard (@kyutai_labs, @Ircam) and Michael Krause (@MireloAI).
kyutai (@kyutai_labs): We’re releasing MuScriptor, the best open model for multi-instrument transcription to date, created in collaboration with @MireloAI. Give it a recording in any genre: pop, classical, metal, jazz, whatever, and it transcribes the individual instruments into MIDI. Link in 🧵
Similar Articles
@kyutai_labs: We're releasing MuScriptor, the best open model for multi-instrument transcription to date, created in collaboration wi…
Kyutai Labs and MireloAI release MuScriptor, the best open model for multi-instrument transcription to date, which transcribes recordings from any genre into individual instrument MIDI tracks.
MuScriptor: An Open Model for Multi-Instrument Music Transcription
MuScriptor is an open-source model for multi-instrument music transcription, capable of transcribing recordings into MIDI notes for each instrument without prior knowledge of the instruments present.
@MireloAI: Today, together with @kyutai_labs, we’re introducing our new Audio-to-MIDI model. It takes a finished recording, identi…
Mirelo AI, in collaboration with Kyutai Labs, introduces an open-source Audio-to-MIDI model that transcribes full music mixes into separate MIDI tracks per instrument, detecting chords, key, and tempo directly from the mix without requiring isolated stems.
@kyutai_labs: Our audio-to-MIDI model, MuScriptor, now also detects tempo! You can directly drag-and-drop the MIDI into a DAW and it …
Kyutai Labs announces that their audio-to-MIDI model MuScriptor now detects tempo, allowing direct drag-and-drop of MIDI into a DAW without manual tempo matching.
@simplifyinAI: This might be the Suno killer for AI music generation. HeartMuLa is an open source family of models that turns lyrics a…
HeartMuLa is an open-source family of AI models for music generation that converts lyrics and tags into full songs, including a language model, music codec, lyrics transcriber, and audio-text alignment model, with demos available on Hugging Face Spaces and ModelScope.