
This ComfyUI workflow performs voice conversion using the ElevenLabs Speech-to-Speech API. You provide a source audio clip (via LoadAudio or RecordAudio), choose a target voice (via ElevenLabsVoiceSelector for a preset voice or ElevenLabsInstantVoiceClone to create a new clone), and the ElevenLabsSpeechToSpeech node generates a new recording that preserves the timing and content of the original while changing the vocal timbre to match the selected voice. The output can be saved with SaveAudioMP3 or SaveAudioOpus.
Technically, the pipeline routes your source waveform into ElevenLabsSpeechToSpeech along with a target voice ID. In Option 1, ElevenLabsVoiceSelector queries your ElevenLabs account and outputs a selected voice for direct use. In Option 2, ElevenLabsInstantVoiceClone uploads a clean reference sample and returns a new voice resource that is then fed into the Speech-to-Speech node. The Note node contains guidance, and several nodes are initially bypassed so you can enable only what you need. This design makes it easy to A/B between preset voices and instant clones while keeping file management simple with explicit save nodes.
FAQ
Frequently Asked Questions
Seedance 2.0: Reference to Video


Nano Banana 2 Lite: Text to Image

Seedream 5.0 Pro: Image Edit

Krea-2: Text to Image
4K Seedance 2.0 - Reference to Video


Z-Image-Turbo Text to Image
Grok: Image Edit
Nano Banana 2 Lite: Image Edit
Grok: Video generation

Seedance2.0 4K: Reference to Video

Grok Imagine Image Quality: Generation
LTX 2.3 - Lipdub LoRA + Voice Clone
Seedance 2.0 Mini: Reference to Video
SCAIL-2: Character Replacement

Ideogram v4: Text to Image
Googly Eyes

Seedance 2.0 - Viral Videos Character Swap
Gemini Omni Flash: Image to Video

Nano Banana 2: Image Edit


cinematic_annotate_video
HappyHorse 1.1: Text to Video
HappyHorse 1.1: Image to Video
Beeble SwitchX: Video Edit

3x3 Contact Sheet
Restore Archival Footage - LTX 2.3 Dearchive LoRA













