ComfyUI-SparkTTS is a custom ComfyUI node implementation of SparkTTS, an advanced text-to-speech system that harnesses the power of large language models (LLMs) to generate highly accurate and natural-sounding speech.
ComfyUI-SparkTTS is a custom ComfyUI node implementation of SparkTTS, an advanced text-to-speech system that harnesses the power of large language models (LLMs) to generate highly accurate and natural-sounding speech.
ComfyUI_SparkTTS is a custom ComfyUI node implementation of SparkTTS, an advanced text-to-speech system that harnesses the power of large language models (LLMs) to generate highly accurate and natural-sounding speech.
Ensure pip install comfy-cli is installed.
Installing ComfyUI comfy install (if you don't have ComfyUI Installed)
install the ComfyUI-SparkTTS, use the following command:
comfy node registry-install Comfyui-Spark-TTS
install requirment.txt in the ComfyUI-SparkTTS folder
This node allows you to clone a voice from a reference audio sample. Inputs: - text: Text to synthesize with the cloned voice. - referenceaudio: The audio sample to clone the voice from. - referencetext: Transcript of the reference audio to improve cloning quality. - maxtokens: Controls the maximum length of generated speech. - batchtexts (optional): Additional texts for better control over pacing and intonation. Outputs: - audio: Generated audio with the cloned voice.
SparkTTS_VoiceCreator
SparkTTS_VoiceCreator
This node allows you to create a customized voice by adjusting parameters. Inputs: - text: Text to synthesize. - gender: Gender of the voice (female or male). - pitch: Pitch level of the voice (verylow, low, moderate, high, veryhigh). - speed: Speed level of the voice (verylow, low, moderate, high, veryhigh). - batchtexts (optional): Additional texts for better control over pacing and intonation. Outputs: - audio: Generated audio with the customized voice.