Fish Audio: Text to Speech

This ComfyUI workflow turns written scripts into natural, expressive speech using the Fish Audio S2 Pro text‑to‑speech model. It supports over 80 languages and lets you direct performance with inline style tags (for example, adding [whisper] or [excited] directly in your text). The core path runs your text through the FishAudioTextToSpeech node and saves a single high‑quality audio file via SaveAudioAdvanced.

You can choose between two voice sources. Option 1 (Preset Audio) uses FishAudioVoiceSelector to supply a curated, ready‑to‑use voice to the TTS node. Option 2 (Voice Clone) uses LoadAudio plus FishAudioInstantVoiceClone to create a temporary voice profile from a short reference clip, then feeds that cloned voice to FishAudioTextToSpeech. Toggle between these options by enabling the corresponding workflow group (Ctrl‑B). A MarkdownNote node on the canvas provides quick reminders, while SaveAudioAdvanced writes the generated speech to disk in your chosen format. Because this uses API nodes, processing happens on the Fish Audio service and is pay‑per‑use once you’re signed in.

API

從程式碼使用此工作流

每個 Comfy 工作流都是一個 JSON 圖。下方的 payload 就是此工作流,與 ComfyUI 執行時完全相同 — 你可以從此 URL 取得、放入版本控制,或在 ComfyUI 載入並逐節點執行。

8b24ee5765ae.json
正在取得工作流 JSON…

使用 Comfy SDK 以 TypeScript 或 Python 執行。相同程式碼可用於 Comfy Cloud 或你自架的 ComfyUI — 只需更改 base URL。

// Install (beta)
npm i @comfyorg/sdk

// Run this workflow (TypeScript)
import { Comfy } from "@comfyorg/sdk";

const client = new Comfy({ apiKey: "comfyui-..." });
const wf = await client.workflows.fromFile("workflow_api.json");
const job = await client.run(wf);
await job.getOutputs("<output-node-id>")[0].toFile("output.png");

SDK 需使用 API 格式的工作流:請在 ComfyUI 開啟此工作流,並使用 檔案 → 匯出工作流(API)。

Comfy Cloud API 存取需有 API 金鑰的方案。

查看所有工作流
Showing 30 of 30 templates