Fish Speech
State-of-the-art open-source multilingual text-to-speech and voice cloning system
Overview
Requirements
Recommended VPS
Hostinger · KVM 4
4 vCPU · 16384 MB · 200 GB
Hostinger · KVM 4
4 vCPU · 16384 MB · 200 GB
Hostinger · KVM 8
8 vCPU · 32256 MB · 400 GB
Affiliate disclosure
Docker Compose
# Generated by Run This Ai — docker-compose.yml
services:
fish-speech:
image: fishaudio/fish-speech:latest
restart: unless-stopped
ports:
- 8080:8080
volumes:
- ./data/fish-speech:/data
Related tools
Stable Diffusion WebUI
Generative image models in your browser
ComfyUI
The most powerful and modular diffusion model GUI with a graph/nodes interface for Stable Diffusion
whisper.cpp
High-performance C++ port of OpenAI Whisper for fast local speech recognition
Fooocus
AI image generator focusing on prompts and generating — a Midjourney-like experience offline
Coqui TTS
Open-source deep learning toolkit for text-to-speech, battle-tested in research and production
ChatTTS
High-quality conversational text-to-speech model optimized for natural daily dialogue
Guides & articles
Getting Started with Fish Speech: Docker Deployment Guide
Step-by-step guide to deploying Fish Speech with Docker. Covers CPU, GPU, WebUI, and server modes for self-hosted TTS.
Fish Speech: The Ultimate Open-Source TTS System for Self-Hosting
Discover Fish Speech — the state-of-the-art open-source TTS system with 31k+ stars. Multilingual, emotionally expressive voice cloning with Docker support.