lenML/Speech-AI-Forge
A Text-to-Speech generation framework supporting multiple TTS/STT models with API server and Gradio WebUI.

Not currently ranked — collecting fresh signals.
star history
Speech-AI-Forge is a project built around TTS generation models, providing an API server and a Gradio-based WebUI for deployment. It integrates popular speech models including ChatTTS, CosyVoice, Fish-Speech, and supports speech-to-text via Whisper. Users can deploy locally, via Docker containers, use a Windows integration package, or run in Google Colab.
Frequently asked
- What is lenML/Speech-AI-Forge?
- A Text-to-Speech generation framework supporting multiple TTS/STT models with API server and Gradio WebUI.
- Is Speech-AI-Forge open source?
- Yes — lenML/Speech-AI-Forge is open source, released under the AGPL-3.0 license.
- What language is Speech-AI-Forge written in?
- lenML/Speech-AI-Forge is primarily written in Python.
- How popular is Speech-AI-Forge?
- lenML/Speech-AI-Forge has 1.4k stars on GitHub.
- Where can I find Speech-AI-Forge?
- lenML/Speech-AI-Forge is on GitHub at https://github.com/lenML/Speech-AI-Forge.