Forward Future Tools Library

DiffRhythm logo

DiffRhythm

DiffRhythm generates full songs with vocals and accompaniment from lyrics and a style prompt, making it useful for musicians and creators producing fast song demos.

Try DiffRhythm

aslp-lab.github.io·From $6.99/month·Checked 2026-08-29

DiffRhythm screenshotOSUM: Advancing Open Speech Understanding Models with Limited Resources in  AcademiaASLP-lab (ASLP-lab)ASLP-lab (@AslpLab_NWPU) / Posts / X
DiffRhythmaslp-lab.github.io
DiffRhythm screenshot
DiffRhythmaslp-lab.github.io

What is DiffRhythm?

DiffRhythm is a latent diffusion model for generating complete songs with vocals and accompaniment. The model page describes songs up to 4 minutes 45 seconds generated in about ten seconds. Inference uses lyrics and a style prompt, with English and Chinese song examples.

What are the pros and cons of DiffRhythm?

Strengths

Generates vocals and accompaniment together rather than requiring separate tracks
Produces full-length songs instead of only short musical segments
Requires a relatively simple input of lyrics and a style prompt
Designed for fast generation through a non-autoregressive architecture
Supports English and Chinese song generation

Trade-offs

The hosted Starter plan is credit-based and allows three concurrent generations
Unlimited hosted generation requires the Unlimited subscription
The model page and hosted pricing page describe different maximum song durations

What are DiffRhythm’s key features?

Generates complete songs with vocal and accompaniment tracks
Creates songs up to 4 minutes 45 seconds according to the model page
Uses lyrics and a style prompt as the main inference inputs
Uses a non-autoregressive architecture for fast inference
Includes English and Chinese song modes
Provides pure music, pure vocal, and audio reference modes

What are the best use cases for DiffRhythm?

Turn drafted lyrics and a style description into a complete song demo
Create vocal-and-instrumental material for songwriting and composition
Generate English or Chinese tracks for content projects
Produce pure music or pure vocal versions for further creative work

What is the pricing for DiffRhythm?

PlanPriceDetails
Starter monthly$6.99/monthIncludes 18,000 credits, all four models, WAV and MP3 downloads, stem extraction, vocal removal, commercial licensing, and three concurrent generations.
Starter yearly$83.9/year billed yearlyIncludes the Starter plan's 18,000 credits, four models, downloads, stem and vocal tools, commercial licensing, and three concurrent generations.
Unlimited monthly$59/monthIncludes unlimited generation with all four models plus 18,000 bonus credits for advanced tools such as vocal separation and stem extraction.
Unlimited yearly$712.9/year billed yearlyIncludes the Unlimited plan's unlimited generation, four models, and 18,000 bonus credits for advanced tools.

The pricing page says unused credits roll over. It also advertises 8-minute songs, which differs from the model page's stated maximum of 4 minutes 45 seconds.

Checked 2026-08-29 · source

Who is DiffRhythm best for?

musiciansA fit for musicians who want to turn lyrics and style ideas into complete vocal demos quickly.
content creatorsUseful for producing original vocal-and-accompaniment tracks for content projects.
small teamThe hosted plans suit teams that need repeated generation, with Unlimited removing the song-generation cap.

What are the best DiffRhythm alternatives?

Where can I try DiffRhythm?

Open aslp-lab.github.io