F5-TTS
F5-TTS. Open-source TTS with voice cloning.
Task Fit
Not marked for code agent in the current library.
Not marked for code in the current library.
Not marked for chat in the current library.
Not marked for rag in the current library.
Not marked for vision in the current library.
Not marked for image generation in the current library.
Not marked for video generation in the current library.
Speech recognition, TTS, or audio workflows.
Source Confidence
Variants and Quant Artifacts
Choose the artifact first; hardware fit follows from RAM, VRAM, format, and runtime.
| Quant | Format | Quality | Min RAM | Reco RAM | Runtime | Action |
|---|---|---|---|---|---|---|
| FP16 | gguf | high | 4GB | 8GB | ollama, llama.cpp, lm-studio | Plan with this |
| Q8 | gguf | high | 4GB | 8GB | ollama, llama.cpp, lm-studio | Plan with this |
Recommended Hardware
Lowest estimated 5-year cost that can run this model.
Enough unified/system memory with a balanced 5-year cost.
Highest local performance signal among compatible hardware.
Benchmarks
No benchmark data is available for this model yet.
Source and Review
Execution evidence
Run F5-TTS with a documented recipe
Recipes connect hardware, a model artifact, tools, settings, verification, and a reportable result.
No verified recipe is linked to this record yet.
Compatibility estimates remain available in the planner. A recipe appears here only after its exact stack and verification protocol are documented.