← Back to the wire

Reachy Mini goes fully local

AnnouncementProductMay 27, 2026

Hugging Face's speech-to-speech pipeline now enables Reachy Mini to run conversations fully locally without cloud dependency. Amir Mahla and Andres Marafioti detail a cascaded VAD, STT, LLM, and TTS architecture that serves a Realtime API-compatible WebSocket endpoint. Users can serve LLMs via llama.cpp or vLLM, with opinionated defaults including Parakeet-TDT for speech recognition and Qwen3TTS for synthesis. The setup allows component swapping for customization across latency and quality trade-offs.

Receipt № 4891 source · awaiting confirmation ◐

Evidence

1source· awaiting independent confirmation

No score is assigned. Sources and their independence are shown in the citation chain below.

Citation chain · 1 source

Hugging FaceCompanyAmir MahlaPersonAndres MarafiotiPerson
Canonical: https://huggingface.co/blog/local-reachy-mini-conversation