This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHubThe TextEncodeAceStepAudio node processes text inputs for audio conditioning by combining tags and lyrics into tokens, then encoding them with adjustable lyrics strength. It takes a CLIP model along with text descriptions and lyrics, tokenizes them together, and generates conditioning data suitable for audio generation tasks. The node allows fine-tuning the influence of lyrics through a strength parameter that controls their impact on the final output.
Conditioning
TextEncodeAceStepAudio - ComfyUI Built-in Node Documentation
Complete documentation for the TextEncodeAceStepAudio node in ComfyUI. Learn its inputs, outputs, parameters and usage.