speech-to-speech

Build local voice agents with open-source models

351 articles 5.6k View on GitHub ↗
351 articles
Speech-to-Speech CLI Commands: Complete Guide to the Hugging Face Toolkit

Master Hugging Face speech-to-speech CLI commands. Explore run, demo, benchmark, and realtime options for seamless voice interaction. Get started now.

how-to-guide
Aug 11, 2026
How to Run the Hugging Face Speech-to-Speech Pipeline in Offline Mode

Run the Hugging Face Speech-to-Speech pipeline offline. Cache models locally or use explicit paths to operate without a network connection and maintain your workflow.

how-to-guide
Aug 11, 2026
Supported TTS Backends in Hugging Face Speech-to-Speech and Installation Guide

Discover supported TTS backends in Hugging Face Speech-to-Speech: chatTTS, MMS, pocket, kokoro, and qwen3. Get installation guides for seamless setup.

installation-guide
Aug 11, 2026
How to Select Different TTS Backends for Speech-to-Speech: A Complete Configuration Guide

Configure your Speech-to-Speech project easily. Learn how to select TTS backends like MLX GGML or Torch within the huggingface repository for optimal performance.

how-to-guide
Aug 11, 2026
How to Skip STT and Send Audio Directly to an Audio-Capable LLM in Hugging Face Speech-to-Speech

Learn how to skip STT and send audio directly to audio-capable LLMs in Hugging Face Speech-to-Speech by setting --stt none. Stream raw audio for efficient processing.

how-to-guide
Aug 11, 2026
How to Use the Chat Completions API for LLM Backend in Speech-to-Speech

Learn how to use the Chat Completions API for your LLM backend with huggingface speech-to-speech. Manage your LLM lifecycle easily and efficiently.

how-to-guide
Aug 11, 2026
How to Use vLLM with the Responses API Backend for Speech-to-Speech Pipelines

Learn how to use vLLM with the Responses API backend for speech-to-speech pipelines. Point the speech-to-speech library to your vLLM HTTP endpoint using responses_api_base_url and responses_api_model_name.

how-to-guide
Aug 11, 2026
How to Configure the Responses API for LLM Backend in huggingface/speech-to-speech

Easily configure the Responses API for your LLM backend in huggingface/speech-to-speech. Learn to set API keys, model names, and endpoint URLs for seamless integration.

how-to-guide
Aug 11, 2026
How to Use an OpenAI-Compatible API for the LLM Backend in speech-to-speech

Integrate an OpenAI-compatible API with the huggingface speech-to-speech pipeline. Connect to Ollama, Azure OpenAI, or self-hosted LLMs using the responses_api backend.

how-to-guide
Aug 11, 2026
How to Select Different LLM Backends for Speech-to-Speech: A Complete Guide

Master selecting LLM backends for speech-to-speech with this guide. Easily switch between transformers, mlx-lm, and more using CLI flags or Python functions.

how-to-guide
Aug 11, 2026
How to Use Faster Whisper as an STT Backend in Speech-to-Speech

Integrate Faster Whisper as your STT backend for enhanced speech-to-speech performance. Learn how to set up this powerful tool easily for your projects. Improve your audio processing today.

how-to-guide
Aug 11, 2026
How to Use Whisper as an STT Backend in the Speech-to-Speech Pipeline

Learn how to use Whisper as an STT backend for your speech-to-speech pipeline. Configure model parameters and generation settings easily via the CLI for efficient transcription.

how-to-guide
Aug 11, 2026

Have a question about this repo?

These articles cover the highlights, but your codebase questions are specific. Give your agent direct access to the source. Share this with your agent to get started:

Share the following with your agent to get started:
curl -s "https://instagit.com/install.md"

Works with
Claude Codex Cursor VS Code OpenClaw Any MCP Client

Maintain an open-source project? Get it listed too →