-
This repository offers a high-quality multilingual text-to-speech (TTS) integration for ComfyUI, featuring advanced capabilities such as voice cloning and emotional control. It supports various Qwen3-TTS models, allowing for customized voice generation and nuanced emotional expression.
-
Supports multiple languages including Russian, English, and more.
-
Features a voice cloning capability that allows users to create unique voice outputs from reference audio.
-
Provides batch processing functionality for generating multiple audio files simultaneously.
Context
This tool, known as DVA Qwen TTS, is designed to enhance ComfyUI by integrating advanced text-to-speech functionalities. Its main purpose is to allow users to synthesize speech with emotional nuances and clone voices, making it particularly useful for creators looking to generate personalized audio content.
Key Features & Benefits
DVA Qwen TTS offers several practical features that enhance the text-to-speech experience. The ability to synthesize speech with emotional tones allows for more engaging audio outputs, while voice cloning enables users to replicate specific voices for various applications. Additionally, the tool supports multiple languages, broadening its usability for diverse audiences.
Advanced Functionalities
This tool includes advanced features such as emotion mixing, which allows users to blend different emotional tones in their speech synthesis. Users can adjust emotional weights in real-time, facilitating the creation of complex emotional transitions. Furthermore, the batch generation capability enables the processing of multiple text inputs at once, streamlining workflows for users who need to produce large volumes of audio.
Practical Benefits
DVA Qwen TTS significantly improves workflow efficiency in ComfyUI by allowing for easy integration of TTS capabilities. Its voice cloning and emotional synthesis features provide users with greater control over the audio output, enhancing the quality of the generated content. The ability to batch process audio files also saves time, making it easier to manage larger projects.
Credits/Acknowledgments
This project is maintained by contributors on GitHub and is distributed under the Apache 2.0 license. The community is encouraged to participate by reporting issues, suggesting improvements, and contributing to the codebase.




