Qwen3-TTS WebUI is an advanced text-to-speech application based on the Qwen3-TTS model, offering robust features like custom voice creation, voice design from natural language descriptions, and voice cloning. With support for multiple languages and an intuitive interface, it caters to diverse voices and applications.
Qwen3-TTS WebUI is a comprehensive text-to-speech web application built on the Qwen3-TTS model, featuring 1.7 billion parameters. This application not only allows for customizable voice options but also offers innovative features like voice design and voice cloning, enhancing the versatility of text-to-speech solutions.
The application interface provides a modern user experience:
Desktop Light Mode

Desktop Dark Mode

Desktop Voice Design List

Desktop Save Voice Design Dialog

Desktop Voice Cloning

Mobile Light & Dark Mode
![]() | ![]() |
![]() | ![]() |
The application provides a RESTful API with endpoints for:
POST /auth/register - Register new users
POST /auth/token - User login
POST /tts/custom-voice - Create a custom voice
POST /tts/voice-design - Design a new voice
POST /tts/voice-clone - Clone a specified voice
GET /jobs - List all jobs
GET /jobs/{id}/download - Download results of a specific job
The Qwen3-TTS WebUI stands out by integrating advanced features, robust backend support, and multi-language capabilities, making it an ideal choice for developers and businesses looking for effective text-to-speech solutions.
No comments yet.
Sign in to be the first to comment.