ComfyUI Node: FL FishSpeech Transcribe
Run ComfyUI workflows without the setup
No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.
Category
πFL FishSpeech
Inputs
audio AUDIO
model
- openai/whisper-large-v3-turbo
- openai/whisper-large-v3
- openai/whisper-medium
- openai/whisper-small
- openai/whisper-base
- openai/whisper-tiny
language
- auto
- en
- zh
- ja
- ko
- de
- fr
- es
- pt
- ru
- it
device
- auto
- cuda
- cpu
Outputs
STRING
Extension: FL FishSpeech
FL FishSpeech - AI Text-to-Speech & Voice Cloning for ComfyUI. High-quality 44.1kHz speech synthesis with zero-shot voice cloning using OpenAudio S1-mini (Fish Audio). Features DualAR Transformer, DAC codec, emotion control tags, and built-in Whisper transcription.
Authored by filliptm
Looking for a different node?
Run ComfyUI workflows without the setup
No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.