Comfyui_URN_AudioTools
A ComfyUI extension with 12 custom nodes.
Nodes (12)
Your audio is glued to the centre — URN Audio Channel unsticks it
New lyrics that actually fit the vocal line, because they were written to the ABC score
Fetch them online, transcribe them as a last resort
A multitrack timeline inside ComfyUI? URN Audio Mixer actually pulls it off
Pick music tags by clicking instead of retyping a comma-separated mess
The audio loader that reads your lyrics back out of the file
Trim, fade and normalise ComfyUI audio without a detour through Audacity
Save audio with the words welded in — a file that knows its own lyrics
Chop a song into 9-second pieces without cutting mid-word
Stretch a 30-second loop into three minutes without an obvious seam
Documentation that doesn't eat your canvas
The node that stops the queue and waits for you to fix the text
ComfyUI - URN Audio Nodes (Yue2)
A collection of custom ComfyUI nodes for audio editing, splitting, mixing, lyric handling, style selection, YuE2 workflows, and workflow documentation.
The package is designed to keep common audio tasks inside ComfyUI while providing interactive visual controls where useful.
Install Intructions at the bottom of the page.
Tutorial video after you have download the nodes. https://youtu.be/VmiJJPEz8Mg
Included Nodes
URN Audio Lyrics (gets online Lyrics for Yue2 etc)
Retrieves lyrics for songs or transcribes them when online lyrics cannot be found.
<img width="610" height="678" alt="image" src="https://github.com/user-attachments/assets/661945fe-6f2d-4d2b-8955-a78ac1bc2690" />Supports automatic title detection, interactive song selection, manual artist/title search, and Whisper as the final fallback.
It can also output duration, filtered MusicBrainz style information, BPM analysis, and the original audio.
URN Audio Style Selector (gets online Styles and customer sytles maker for Yue2 etc)
<img width="890" height="657" alt="image" src="https://github.com/user-attachments/assets/1a7d809e-1abf-41e6-924c-d7eead7c18e9" />Visual style/tag selector designed for music-generation workflows such as YuE2.
Incoming generated styles can be reviewed, removed, or supplemented with user-selected tags from configurable categories.
Tabs and available styles are loaded from JSON, and custom user tags can be added and saved directly from the node.
URN Audio Load and Save with Lyrics metadata
<img width="850" height="267" alt="image" src="https://github.com/user-attachments/assets/f39074d3-51ea-41ec-b481-e99d33e69324" />URN Audio Smart Splitter
Splits longer audio into manageable chunks while attempting to avoid cutting through vocals or words.
Can use Whisper analysis and optional Mel-RoFormer vocal/music separation to improve cut placement.
Supports vocal/music stem export, vocal chunks, transcript sidecars, and FLAC/MP3 chunk output.
URN Audio Mixer
<img width="1199" height="636" alt="image" src="https://github.com/user-attachments/assets/3c321473-884d-4599-b45f-6cb936ecd863" />A visual multi-track audio mixer for arranging and combining multiple clips inside ComfyUI.
Clips can be positioned and adjusted independently with trim, fades, and track-level controls.
Useful for assembling generated music, vocals, effects, or other audio elements into a final mix.
URN Audio Channel
Provides straightforward stereo-channel processing for connected audio.
Includes left/right level control, channel swapping, static panning, and whole-clip auto-pan.
Designed as a lightweight utility node for quick stereo adjustments.
URN Audio Trim Fade
Visual trim, fade, gain, normalization, and silence-padding utility for connected audio.
Includes an interactive waveform editor with start/end and fade controls, plus local preview support.
Can work by output length or end position and passes the processed audio back into the workflow.
URN Text Edit
Interactive text-review node that pauses the workflow until the user accepts the current text.
The text can be edited directly, saved to a .txt file, or replaced by loading an existing .txt file.
Only Accept Changes resumes downstream workflow execution.
Optional Models and Downloads
Some features require external models:
- Faster-Whisper models are downloaded and cached locally when transcription is first used.
- Mel-RoFormer is used by the Smart Audio Splitter for optional vocal/music separation and downloads its required model when needed.
Model files are not bundled with this repository.
Editable Configuration Files
urn_audio_style_selector_styles.json
Controls the tabs and tags displayed by URN Audio Style Selector.
Each top-level JSON section automatically becomes a tab, so categories and tags can be expanded without editing the node code.
urn_musicbrainz_tag_blocklist.json
Contains tags that should be excluded from MusicBrainz-derived music descriptions.
The file can be edited directly to customise the filtering behaviour.
Dependencies
The package uses several non-core Python dependencies, including:
faster-whisperaudio-separatorlibrosascipysoundfile
Use the included install.bat or install the versions listed in requirements.txt with the Python environment used by ComfyUI.
Installation
-
Download or clone the repository.
-
Place the
URN Audio Nodesfolder inside:ComfyUI/custom_nodes/ -
Run
install.batonce to install the required Python dependencies. -
Restart ComfyUI.
-
After updating frontend files, a browser hard refresh (
Ctrl+F5) may be required.
If you already manage dependencies manually, you can install them from
requirements.txtusing ComfyUI's Python environment.
Manual installation with pip
If you do not want to use install.bat, open a Command Prompt in your ComfyUI installation folder and run:
python_embeded\python.exe -m pip install -r "ComfyUI\custom_nodes\URN Audio Nodes\requirements.txt"
If your ComfyUI installation does not use the Windows python_embeded layout, run the same requirements file using the Python environment that launches ComfyUI:
python -m pip install -r "ComfyUI/custom_nodes/URN Audio Nodes/requirements.txt"
First run: some features may take a little longer the first time they are used because required models such as Faster-Whisper and Mel-RoFormer may need to be downloaded and cached locally. Later runs will reuse the downloaded model files.
Workflow Compatibility
Where visible node names have changed, internal node IDs have been preserved where possible so existing workflows continue to load correctly.
Documentation
More detailed per-node documentation is included in:
URN Audio Nodes/Node Readmes/
These files can also be copied into ComfyUI Markdown/Note nodes and placed beside the matching node in a workflow.