TTS WebUI: Releases and Overview

September 18, 2026

Web UI that brings many text-to-speech and music generation models under one roof.

Category: Image, video and audio. Part of our engine and tool tracker.

At a Glance

Item Value
Repository rsxdalv/tts-webui
License MIT
Main language TypeScript
GitHub stars 3,264 (as of 2026-09-17)
Latest release v1.5.2 (2026-09-01)
Install / docs official documentation

The project describes itself as: “A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro, OpenVoice, ParlerTTS, Stable Audio, MMS, StyleTTS2, MAGNet, AudioGen, MusicGen, Tortoise, RVC, Vocos, Demucs, SeamlessM4T, and Bark!”

Release History

Compiled by Local Model Watch from the project’s GitHub releases. Pre-releases are omitted. Where we wrote an article about a release, it is linked in the last column; smaller releases are tracked here only.

Version Released Release notes Our article
v1.5.2 2026-09-01 GitHub
v1.5.1 2026-05-15 GitHub
v1.5.0 2026-05-01 GitHub
v1.4.1 2026-04-27 GitHub
v1.4.0 2026-04-23 GitHub
v1.2.0 2026-04-23 GitHub
v1.1.0 2026-04-06 GitHub
v1.0.0 2026-04-06 GitHub
v0.5.1 2026-04-04 GitHub
v0.4.0 2025-11-24 GitHub
v0.3.0 2025-11-14 GitHub
v0.0.1 2025-09-29 GitHub
v0.0.0 (Latest release) 2025-08-17 GitHub

Articles on Local Model Watch

No articles yet.

Last updated 2026-09-18 (JST). Facts above come from the GitHub API; the one-paragraph summary is written by the site’s editors.