ElevenLabs: Texto para Voz

Este fluxo de trabalho do ComfyUI utiliza o ElevenLabs para transformar texto escrito em fala ultra-realista. Utilizando nós como ElevenLabsTextToSpeech e ElevenLabsVoiceSelector, é possível selecionar entre vozes pré-definidas ou enviar uma amostra para clonar uma voz específica para síntese. O fluxo é versátil, permitindo tanto a conversão texto-para-fala quanto a clonagem de voz, sendo uma ferramenta útil para criadores de conteúdo, educadores e desenvolvedores que precisam de saídas de áudio de alta qualidade. Com a integração dos nós LoadAudio e SaveAudioMP3, é possível gerenciar entradas e saídas de áudio de forma eficiente, garantindo uma experiência fluida do texto até o arquivo de áudio gerado.

API

Use this workflow from code

Every Comfy workflow is a JSON graph. The payload below is this workflow, exactly as ComfyUI runs it — fetch it from the URL, keep it in version control, or load it in ComfyUI and run it node by node.

api_elevenlabs_text_to_speech.json
Fetching workflow JSON…

Run it from TypeScript or Python with the Comfy SDK. The same code targets Comfy Cloud or a ComfyUI you host yourself — only the base URL changes.

// Install (beta)
npm i @comfyorg/sdk

// Run "ElevenLabs: Texto para Voz" (TypeScript)
import { Comfy } from "@comfyorg/sdk";

const client = new Comfy({ apiKey: "comfyui-..." });

// This workflow, exported in API format (see note below)
const wf = await client.workflows.fromFile("api_elevenlabs_text_to_speech_api.json");

const job = await client.run(wf);
await job.getOutputs("215")[0].toFile("output.png"); // SaveAudioMP3

The SDK takes a workflow in API format: open this workflow in ComfyUI and use File → Export Workflow (API).

Comfy Cloud API access requires a plan with an API key.

FAQ

Perguntas Frequentes

Ver todos os workflows
Character
Cinematic
Image to Video
Lip Sync
Multiple Angles
Portrait
Style Reference
Style Transfer
Text to Video
Video Generation
Video
Showing 30 of 30 templates