
This ComfyUI workflow leverages the power of ElevenLabs to transform written text into ultra-realistic speech. Utilizing nodes like ElevenLabsTextToSpeech and ElevenLabsVoiceSelector, users can either select from a range of preset voices or upload a sample to clone a specific voice for synthesis. The workflow is versatile, allowing for both text-to-speech conversion and voice cloning, making it a powerful tool for content creators, educators, and developers who need high-quality audio outputs. By integrating nodes such as LoadAudio and SaveAudioMP3, users can manage audio inputs and outputs efficiently, ensuring a seamless experience from text input to audio file generation.
API
Use this workflow from code
Every Comfy workflow is a JSON graph. The payload below is this workflow, exactly as ComfyUI runs it — fetch it from the URL, keep it in version control, or load it in ComfyUI and run it node by node.
GET https://comfy.org/workflows/download/a8749454536e.jsonFetching workflow JSON…
Run it from Python or TypeScript with the Comfy SDK. The same code targets Comfy Cloud or a ComfyUI you host yourself — only the base URL changes.
# Install (beta)
pip install comfy-sdk # Python
npm i @comfyorg/sdk # TypeScript
# Run "ElevenLabs: Text to Speech" (Python)
from comfy_sdk import Comfy
client = Comfy(api_key="comfyui-...")
# This workflow, exported in API format (see note below)
wf = client.workflows.from_file("api_elevenlabs_text_to_speech_api.json")
job = client.run(wf)
for output in job.get_outputs("215"): # SaveAudioMP3
output.to_file(output.name)The SDK takes a workflow in API format: open this workflow in ComfyUI and use File → Export Workflow (API). API access requires a Comfy Cloud plan with an API key. SDK docs · Get an API key
FAQ














