@Thom_Wolf: Most people should probably update their priors on the state of open-source speech-to-speech. It's honestly kind of min…

X AI KOLs Following Tools

Summary

Thom Wolf and Cerebras released a fully open-source realtime voice demo with models and code, showcasing state-of-the-art speech-to-speech capabilities.

Most people should probably update their priors on the state of open-source speech-to-speech. It's honestly kind of mind-blowing. We teamed up with @cerebras to build a fully open-source realtime voice demo (models + code) to show what's possible today. Demo : https://huggingface.co/spaces/smolagents/hf-realtime-voice… Blog: https://huggingface.co/blog/cerebras-gemma4-voice-ai… Go test it, fork it, tweak it, and impress your friends. video is raw, no cut, no speed-up, first take
Original Article
View Cached Full Text

Cached at: 07/03/26, 08:33 AM

Most people should probably update their priors on the state of open-source speech-to-speech.

It’s honestly kind of mind-blowing.

We teamed up with @cerebras to build a fully open-source realtime voice demo (models + code) to show what’s possible today.

Demo : https://huggingface.co/spaces/smolagents/hf-realtime-voice…

Blog: https://huggingface.co/blog/cerebras-gemma4-voice-ai…

Go test it, fork it, tweak it, and impress your friends.

video is raw, no cut, no speed-up, first take


HF Realtime Voice - a Hugging Face Space by smolagents

Source: https://huggingface.co/spaces/smolagents/hf-realtime-voice Fetching metadata from the HF Docker repository...

Similar Articles

@kwindla: https://x.com/kwindla/status/2062544580105359686

X AI KOLs Timeline

NVIDIA released Nemotron 3.5 ASR, an open-source multilingual speech-to-text model with the lowest latency tested, available in multilingual and English-only variants, ideal for voice agents and self-hosted deployments.