speech-to-speech
Build voice agents with open-source models
About this project
Speech To Speech: Build voice agents with open-source models A low-latency, fully modular voice-agent pipeline: VAD - STT - LLM - TTS, exposed through the core OpenAI Realtime GA event set over WebSocket and WebRTC. Every component is swappable. The LLM slot speaks OpenAI-compatible protocols, so you can point it at a hosted provider, at HF Inference Providers, or at a vLLM or llama.cpp server on your own hardware for a fully local, fully open stack. This pipeline runs in production as the conversation backend for thousands of Reachy Mini robots. Quickstart Choose where the language model should run. All three configurations use local Parakeet…
Technologies
Project health
GitHub
Reviews
Built by
Maintain huggingface/speech-to-speech? Claiming verifies admin access through your GitHub account and gives you control of this listing.

