speech-to-speech

by huggingface · Machine Learning

Build voice agents with open-source models

New0 ratings13,082 starsActive
Machine LearningAI ApplicationPythonDockerfile
↓ Download v1.0.0GitHub⚑ Report
speech-to-speech — image 1
speech-to-speech — image 2
speech-to-speech — image 3
1 / 3

About this project

  Speech To Speech: Build voice agents with open-source models A low-latency, fully modular voice-agent pipeline: VAD - STT - LLM - TTS, exposed through the core OpenAI Realtime GA event set over WebSocket and WebRTC. Every component is swappable. The LLM slot speaks OpenAI-compatible protocols, so you can point it at a hosted provider, at HF Inference Providers, or at a vLLM or llama.cpp server on your own hardware for a fully local, fully open stack. This pipeline runs in production as the conversation backend for thousands of Reachy Mini robots. Quickstart Choose where the language model should run. All three configurations use local Parakeet…

Technologies

aiassistantPythonDockerfilelanguage-model

Project health

Actively maintained
Last update2 days ago
Contributors42
Latest releasev1.0.0
Open issues & PRs98
LicenseApache-2.0
On GitHubsince 2024

GitHub

13,082
stars
1,650
forks
42
contributors
98
open issues & PRs
Python
language
2 days ago
last commit
View on GitHub ↗All releases ↗

Reviews

out of 5 · 0 ratings
★★★★★
0%
★★★★
0%
★★★
0%
★★
0%
0%
Sign in to write a review
No reviews yet
Be the first to review speech-to-speech.

Built by

huggingface
Imported from GitHub · not yet claimed on GitPalace
View developer pageSign in with GitHub to claim

Maintain huggingface/speech-to-speech? Claiming verifies admin access through your GitHub account and gives you control of this listing.

You might also like