ACE-Step
/v1/runAce Step Audio To Audio by ace - advanced AI model for audio-to-audio. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
Request body
Submit an async generation request. The model field selects the model; other fields are model-specific input parameters.
Model identifier. Set to ace/ace-step/audio-to-audio.
Default: ace/ace-step/audio-to-audio
URL of the audio file to be outpainted.
Lyrics to be sung in the audio. If not provided or if [inst] or [instrumental] is the content of this field, no lyrics will be sung. Use control structures like [verse], [chorus] and [bridge] to control the structure of the song.
Default:
Guidance scale for the generation.
Range: 0 to 200
Default: 15
Random seed for reproducibility. If not provided, a random seed will be used.
Minimum guidance scale for the generation after the decay.
Range: 0 to 200
Default: 3
Tag guidance scale for the generation.
Range: 0 to 10
Default: 5
Lyric guidance scale for the generation.
Range: 0 to 10
Default: 1.5
Whether to edit the lyrics only or remix the audio.
Allowed values: lyrics, remix
Default: remix
Guidance interval for the generation. 0.5 means only apply guidance in the middle steps (0.25 * infer_steps to 0.75 * infer_steps)
Range: 0 to 1
Default: 0.5
Original seed of the audio file.
Scheduler to use for the generation process.
Allowed values: euler, heun
Default: euler
Granularity scale for the generation process. Higher values can reduce artifacts.
Range: -100 to 100
Default: 10
Type of CFG to use for the generation process.
Allowed values: cfg, apg, cfg_star
Default: apg
Original lyrics of the audio file.
Default:
Original tags of the audio file.
Guidance interval decay for the generation. Guidance scale will decay from guidance_scale to min_guidance_scale in the interval. 0.0 means no decay.
Range: 0 to 1
Default: 0
Number of steps to generate the audio.
Range: 3 to 60
Default: 27
Comma-separated list of genre tags to control the style of the generated audio.
Response Schema
The submit endpoint returns an accepted generation task. Poll the result endpoint with the returned id for terminal outputs or errors.
Error message if the task failed. Empty on success.
Unique identifier for the generation task.
Model ID used for the prediction.
Array of generated content. Empty when status is not completed.
Status of the task: pending, running, completed, failed, or timeout.
Allowed values: pending, running, completed, failed, timeout
Model capabilities
Capabilities declared by the model registry.
Default: audio-to-audio
Execution mode declared by the model registry.
Default: async

