@Open_MOSS: MOSS-Transcribe-Diarize is among the top trending models on @huggingface . A few weeks after release, the model has bee…
Summary
MOSS-Transcribe-Diarize, an open-source ASR model with multi-speaker diarization and hotword biasing, trends on Hugging Face after release.
View Cached Full Text
Cached at: 07/16/26, 06:21 PM
MOSS-Transcribe-Diarize is among the top trending models on @huggingface .
A few weeks after release, the model has been explored by developers and researchers worldwide.@MosiAI_Official
Built with an end-to-end audio-to-structured-transcript paradigm: • 0.9B open-source ASR model • Apache license 2.0 • 128k long-context transcription • Up to ~90-min audio input • Speaker labels + timestamps in one generation • Multi-speaker diarization for meetings, interruptions, and overlapping voices • Hotword biasing for names, terms, and domain-specific vocabulary • ~100 token/s on NVIDIA RTX 4090, RTF ~0.017
If you haven’t
Similar Articles
@MosiAI_Official: MOSS-Transcribe-Diarize-0.9B is now open source on @huggingface. Built with an end-to-end audio-to-structured-transcrip…
MOSS-Transcribe-Diarize-0.9B is an open-source end-to-end audio understanding model for long-form multi-speaker transcription, diarization, and timestamp generation, released by Mosi AI under Apache 2.0.
@MosiAI_Official: MOSS-TTS-v1.5 just reached #1 on Hugging Face Trending for Text-to-Speech, with 20.6K downloads. A multilingual, contro…
MOSS-TTS-v1.5, a multilingual controllable TTS model with voice cloning and long-form generation, reached #1 on Hugging Face Trending with 20.6K downloads.
OpenMOSS-Team/MOSS-TTS-v1.5 · Hugging Face
MOSS-TTS v1.5 is an updated open-source text-to-speech model with improved multilingual synthesis (supporting 31 languages), more stable zero-shot voice cloning, and explicit inline pause control.
@cohere: Cohere Transcribe, our open-source speech recognition model, is #1 on the new @huggingface Far-Field ASR benchmark.
Cohere Transcribe, an open-source speech recognition model, achieved first place on Hugging Face's new Far-Field ASR benchmark.
I fine-tuned Cohere Transcribe to support diarization and timestamps
Fine-tuned Cohere Transcribe, the best open-source speech-to-text model, to support diarization and timestamps. The new model is available on Hugging Face.