← All workflows

specialty / Audio / Intermediate

ACE Step 1.5 — Text to Song

Generate audio from musical tags and lyrics with the ACE Step 1.5 checkpoint graph.

Recorded audio output

Generated locally with the settings in the test record.

Actual output from the recorded RTX 5080 run. See test settings below.

What it produces

Generate audio from musical tags and lyrics with the ACE Step 1.5 checkpoint graph.

HARDWARE & SETTINGS

TESTED on NVIDIA GeForce RTX 5080 - UNCLASSIFIED

This is a catalog estimate, not a tested minimum. GPU architecture, precision, resolution and offloading affect whether a run fits.

Output settings
Audio · 120 seconds requested
Workflow version
v1.5
System RAM
62 GB (test system)
Test hardware
RTX 5080 · 16 GB
Generation time
214.70 seconds in this run
Peak memory
Not measured
Read the compatibility policy →

HARDWARE INTELLIGENCE LAYER

Will this run on my GPU?

Evaluate evidence-backed execution feasibility for your specific hardware.

TESTED ON RTX 5080 - UNCLASSIFIEDConfidence: HIGH

Directly Tested on NVIDIA GeForce RTX 5080

Current workflow hash matches the successful unclassified execution. Success applies to its recorded settings.

CLOSEST EVIDENCE

NVIDIA GeForce RTX 5080 (16GB VRAM)

Exact GPU model and current workflow hash; applies only to the recorded settings and software.

Measured: 214.70s runtime

MEASURED RUNTIME SPECS

Resolution:
Audio · 120 seconds requested
Steps / Sampler:
8
Runtime:
214.70s
Test Date:
2026-09-08

Hardware & Runtime Considerations:

  • Historical success: STANDARD versus OPTIMIZED was not recorded.

Execution history (1)

Success applies to the recorded settings. Failure is an outcome; STANDARD and OPTIMIZED describe execution settings. Historical runs without that distinction remain unclassified.

Hardware / dateProfile / outcomeWorkflow evidenceDetails
NVIDIA GeForce RTX 5080
2026-09-08T02:33:37.220Z
UNCLASSIFIED
SUCCESS · 214.70s
CURRENT
d364f30dab2d
Exact record · Output
Settings and notes
{
  "model": [
    "78",
    0
  ],
  "seed": 31,
  "steps": 8,
  "cfg": 1,
  "sampler_name": "euler",
  "scheduler": "simple",
  "positive": [
    "94",
    0
  ],
  "negative": [
    "47",
    0
  ],
  "latent_image": [
    "98",
    0
  ],
  "denoise": 1,
  "resolution": "Audio · 120 seconds requested",
  "batchSize": 1
}

Historical run: STANDARD versus OPTIMIZED was not recorded. Original evidence preserved.

No comparable cross-GPU executions. Timing comparisons require matching workflow, inputs, model hashes, settings, software, precision and timing protocol.

Models & custom nodes

These model filenames are extracted from the downloadable graph, including subgraphs. Check each source’s license and access requirements.

ace_step_1.5_turbo_aio.safetensorsComfyUI/models/checkpoints/

Custom nodes

The graph uses ComfyUI core nodes. Use a recent ComfyUI version that includes these node types.

Standard ComfyUI graph.

Custom node setup guide →

Installation & first run

  1. Install ComfyUI using the official requirements and installation documentation for your operating system and GPU.
  2. Download this workflow JSON. Use the listed model sources above and place the required files in the indicated model folders. Model folder guide →
  3. Load the JSON in ComfyUI. Resolve missing nodes and select installed model files in the loader nodes. Supply any required input image, video or audio.
  4. Review the graph’s settings before running. See the settings extracted below; a recorded test only establishes the specific configuration used in that run.
  5. Run the workflow and check the ComfyUI console if it fails. Keep the exact error text, software versions and your hardware details for troubleshooting.

Example prompt from the catalog

Neo-Soul: A warm, organic neo-soul track dripping with live instrumentation and effortless groove. A live drummer plays a loose, hip-hop influenced pocket—soft kick drum with lazy swing, snare hits that sit just behind the beat, and brushed hi-hats that breathe and shuffle with human imperfection.

This prompt is a starting point, not evidence that the preview was produced by this workflow.

Explore the workflow

When a run fails

Missing model or a model name is not available

Compare the graph’s loader file names with your installed files and the dependency list. Check model folders →

Missing or unrecognized node type

Record the exact node type and check its source and version before installing extensions. Custom node guide →

GPU memory or execution error

The VRAM estimate is not a guarantee. Record the error, input dimensions, model precision and batch size. GPU error guide → · Workflow error guide →

Graph settings & required inputs

No Load Image input file is required by this graph.

View the graph’s saved controls
EmptyAceStep1.5LatentAudio · node 98
[120,1]
KSampler · node 3
[31,"fixed",8,1,"euler","simple",1]

Testing & source information

This exact download completed a local run on an RTX 5080 with 16 GB VRAM. This does not establish a minimum VRAM requirement.

Test date
2026-09-08
ComfyUI
0.33.0
PyTorch
2.11.0.dev20260211+cu128
Resolution / batch
Audio · 120 seconds requested / 1
Steps / CFG
8 / 1
Sampler / scheduler
euler / simple
Seed
31
Generation time
214.70 seconds, including model loading
Peak memory
Not measured
Download the test record →

Catalog version: v1.5. Original publication date is not recorded.

Graph source: audio_ace_step_1_5_checkpoint.json. Upstream license

What a complete test record includes →

Related workflows