speech-to-speech
Build local voice agents with open-source models
Master Hugging Face speech-to-speech CLI commands. Explore run, demo, benchmark, and realtime options for seamless voice interaction. Get started now.
How to Run the Hugging Face Speech-to-Speech Pipeline in Offline ModeRun the Hugging Face Speech-to-Speech pipeline offline. Cache models locally or use explicit paths to operate without a network connection and maintain your workflow.
Supported TTS Backends in Hugging Face Speech-to-Speech and Installation GuideDiscover supported TTS backends in Hugging Face Speech-to-Speech: chatTTS, MMS, pocket, kokoro, and qwen3. Get installation guides for seamless setup.
How to Select Different TTS Backends for Speech-to-Speech: A Complete Configuration GuideConfigure your Speech-to-Speech project easily. Learn how to select TTS backends like MLX GGML or Torch within the huggingface repository for optimal performance.
How to Skip STT and Send Audio Directly to an Audio-Capable LLM in Hugging Face Speech-to-SpeechLearn how to skip STT and send audio directly to audio-capable LLMs in Hugging Face Speech-to-Speech by setting --stt none. Stream raw audio for efficient processing.
How to Use the Chat Completions API for LLM Backend in Speech-to-SpeechLearn how to use the Chat Completions API for your LLM backend with huggingface speech-to-speech. Manage your LLM lifecycle easily and efficiently.
How to Use vLLM with the Responses API Backend for Speech-to-Speech PipelinesLearn how to use vLLM with the Responses API backend for speech-to-speech pipelines. Point the speech-to-speech library to your vLLM HTTP endpoint using responses_api_base_url and responses_api_model_name.
How to Configure the Responses API for LLM Backend in huggingface/speech-to-speechEasily configure the Responses API for your LLM backend in huggingface/speech-to-speech. Learn to set API keys, model names, and endpoint URLs for seamless integration.
How to Use an OpenAI-Compatible API for the LLM Backend in speech-to-speechIntegrate an OpenAI-compatible API with the huggingface speech-to-speech pipeline. Connect to Ollama, Azure OpenAI, or self-hosted LLMs using the responses_api backend.
How to Select Different LLM Backends for Speech-to-Speech: A Complete GuideMaster selecting LLM backends for speech-to-speech with this guide. Easily switch between transformers, mlx-lm, and more using CLI flags or Python functions.
How to Use Faster Whisper as an STT Backend in Speech-to-SpeechIntegrate Faster Whisper as your STT backend for enhanced speech-to-speech performance. Learn how to set up this powerful tool easily for your projects. Improve your audio processing today.
How to Use Whisper as an STT Backend in the Speech-to-Speech PipelineLearn how to use Whisper as an STT backend for your speech-to-speech pipeline. Configure model parameters and generation settings easily via the CLI for efficient transcription.
Have a question about this repo?
These articles cover the highlights, but your codebase questions are specific. Give your agent direct access to the source. Share this with your agent to get started:
curl -s "https://instagit.com/install.md" Maintain an open-source project? Get it listed too →