Local TTS model
Fish Speech
Catalogue summary: Fast, high-quality TTS with voice cloning and multilingual support. Optimized for real-time applications.
Repository editorial metadata; verify comparative claims in the linked upstream material.
Apple Silicon readytext-to-speech generation10 languagesApache 2.0
Choose an app
Compare speech modelsStart with Recommended. No terminal commands are shown.
Open model filesRecommended verified starting pointOpen in UnslothLaunches Unsloth Desktop on this modelOpen Python packageVerified package page without exposing a commandOpen on Hugging FaceFiles, licence and model cardCheck my machineOptional: verify the fit with your saved hardware
Desktop app links require the app to be installed. If nothing opens, LocalClaw will show app-download and model-file fallbacks.
Catalogue quality
9/10
Catalogue speed
8.5/10
Model size
3 GB
Voices
Unlimited cloning
Can Fish Speech run locally?
Fish Speech can generate speech locally for private voice workflows. Use the verified setup options on this page; no terminal command is required to choose the right path.
Apache 2.0 license. Still verify upstream usage notes before shipping.
streamingrealtimecloningmultilingual
Audio profile
Best fit
Fish Speech is best for local voice cloning and expressive speech generation.
Hardware: gpuapple
Model details
Type
Local TTS model
Family
fish
Latency
ultra-low
Formats
pytorchonnx
Languages
en, zh, ja, ko, fr, de, es, it, pt, ru
Context
Instant cloning
Install locally
01
Check runtimeConfirm the backend supports pytorch, onnx on your machine.02
Open recommended setupUse the app and model links above. LocalClaw does not expose a terminal command.03
Test locallyRun a short private audio prompt before moving into production workflows.Good for
- text-to-speech generation
- Apple Silicon ready local workflows
- streaming, realtime, cloning
Watch before shipping
- Validate pronunciation, latency and artifacts with your own voice samples.
- Review the upstream license and acceptable-use notes.
- Benchmark on your target CPU, Apple Silicon or GPU setup.
Related TTS and speech models
Phạm Nguyễn Ngọc Bảo
VieNeu-TTS v3 Turbo
Local TTS model · Q 9.2 · Speed 9.4
Samuel Vitorino / HALO Research
Sopro V2 Turbo
Local TTS model · Q 9.2 · Speed 9.4
Zyphra
ZONOS2
Local TTS model · Q 9.6 · Speed 8.5
Alibaba FunAudioLLM
CosyVoice 2
Local TTS model · Q 9.3 · Speed 8.8
OpenBMB
VoxCPM2
Local TTS model · Q 9.4 · Speed 8.3
Nineninesix
Gepard 1.0
Local TTS model · Q 8.9 · Speed 9.2
OpenMOSS / MOSI.AI
MOSS-TTS-Nano
Local TTS model · Q 8.5 · Speed 9.7
hexgrad
Kokoro TTS
Local TTS model · Q 9.2 · Speed 9.8