Run This Ai
EN DE

F5-TTS

High-quality non-autoregressive text-to-speech with flow matching

★ 14,848 GitHub MIT text-to-speechttsvoice-cloningaudioflow-matchingspeech-synthesisai-voice Bild & Video

Überblick

F5-TTS is a state-of-the-art text-to-speech synthesis system that uses flow matching for natural, expressive speech generation. With over 14,800 GitHub stars, it offers fairytale-like voice quality. Key features include zero-shot voice cloning from just a short audio reference, adjustable speech speed and emotion, bilingual support (Chinese and English), a Gradio web UI for easy inference, and both CLI and Docker deployment options. Built on a flow-matching architecture, F5-TTS produces remarkably fluent and faithful speech output.

Anforderungen

Min vCPU
2
Min RAM
4096 MB
Min Disk
10 GB
Rec vCPU
4
Rec RAM
8192 MB
Rec Disk
20 GB

Empfohlener VPS

Hostinger · KVM 4

4 vCPU · 16384 MB · 200 GB

$11.99
Zum Anbieter

Hostinger · KVM 4

4 vCPU · 16384 MB · 200 GB

$11.99
Zum Anbieter

Hostinger · KVM 8

8 vCPU · 32256 MB · 400 GB

$20.00
Zum Anbieter

Affiliate-Hinweis

Docker Compose

# Generated by Run This Ai — docker-compose.yml
services:
  f5-tts:
    image: ghcr.io/swivid/f5-tts:main
    restart: unless-stopped
    ports:
      - 8080:8080
    volumes:
      - ./data/f5-tts:/data

Bester VPS für F5-TTS →

Verwandte Tools

Anleitungen & Artikel