Local TTS model
F5-TTS
Catalogue summary: Flow-matching based TTS with SOTA quality and extremely fast inference. Simple and efficient architecture.
Repository editorial metadata; verify comparative claims in the linked upstream material.
Apple Silicon readytext-to-speech generation2 languagesCC-BY-NC-4.0 weights; MIT code
Choose an app
Compare speech modelsStart with Recommended. No terminal commands are shown.
Open model filesRecommended verified starting pointOpen in UnslothLaunches Unsloth Desktop on this modelOpen Python packageVerified package page without exposing a commandOpen on Hugging FaceFiles, licence and model cardCheck my machineOptional: verify the fit with your saved hardware
Desktop app links require the app to be installed. If nothing opens, LocalClaw will show app-download and model-file fallbacks.
Catalogue quality
9.4/10
Catalogue speed
9/10
Model size
1.5 GB
Voices
Reference cloning
Can F5-TTS run locally?
F5-TTS can generate speech locally for private voice workflows. Use the verified setup options on this page; no terminal command is required to choose the right path.
CC-BY-NC-4.0 weights; MIT code license. Review upstream restrictions before commercial use.
realtimecloningstreaming
Audio profile
Best fit
F5-TTS is best for local voice cloning and expressive speech generation.
Hardware: gpuapple
Model details
Type
Local TTS model
Family
f5
Latency
ultra-low
Formats
pytorchsafetensors
Languages
en, zh
Context
Flow matching
Install locally
01
Check runtimeConfirm the backend supports pytorch, safetensors on your machine.02
Open recommended setupUse the app and model links above. LocalClaw does not expose a terminal command.03
Test locallyRun a short private audio prompt before moving into production workflows.Good for
- text-to-speech generation
- Apple Silicon ready local workflows
- realtime, cloning, streaming
Watch before shipping
- Validate pronunciation, latency and artifacts with your own voice samples.
- Review the upstream license and acceptable-use notes.
- Benchmark on your target CPU, Apple Silicon or GPU setup.
Related TTS and speech models
Speech Research (SWivid)
F5-TTS v1 Base
Local TTS model · Q 9.5 · Speed 9.2
Amphion Team
MaskGCT
Local TTS model · Q 9.4 · Speed 9
Phạm Nguyễn Ngọc Bảo
VieNeu-TTS v3 Turbo
Local TTS model · Q 9.2 · Speed 9.4
Samuel Vitorino / HALO Research
Sopro V2 Turbo
Local TTS model · Q 9.2 · Speed 9.4
Zyphra
ZONOS2
Local TTS model · Q 9.6 · Speed 8.5
BreezeBlue
Breeze TTS 2
Local TTS model · Q 9.6 · Speed 8.4
Zyphra
Zonos v0.1
Local TTS model · Q 9.5 · Speed 8.5
Neuphonic
NeuTTS Air
Local TTS model · Q 9 · Speed 9.5