Skip to content
Hugging Face Blog·· 2026-07-01

Hugging Face and Cerebras bring Gemma 4 to real-time voice AI

Hugging Face and Cerebras bring Gemma 4 to real-time voice AI

AI summary

Hugging Face and Cerebras jointly demonstrated a real-time speech-to-speech pipeline, accelerating Gemma 4 31B inference with Cerebras and combining Nvidia Parakeet speech recognition with Alibaba Qwen3TTS synthesis.

Selection record

AdmittedSum of both 124 ≥ twice the threshold 120

Source tier
Official, first-hand; this tier's threshold is 60
Pre-filter
passed:Gemma4语音AI架构与推理演示
Why it was chosen
The reusable open-source speech-to-speech architecture and components at each layer help assess how to build real-time voice interaction.

A model scores each item twice, independently, against one written standard, out of 100. An item is admitted only when the two scores add up to twice the threshold. The threshold is set per source tier.

Source: Hugging Face Blog · huggingface.co