Forward Future Tools Library

DiffRhythm
DiffRhythm generates full songs with vocals and accompaniment from lyrics and a style prompt, making it useful for musicians and creators producing fast song demos.
aslp-lab.github.io·From $6.99/month·Checked 2026-08-29




›What is DiffRhythm?
DiffRhythm is a latent diffusion model for generating complete songs with vocals and accompaniment. The model page describes songs up to 4 minutes 45 seconds generated in about ten seconds. Inference uses lyrics and a style prompt, with English and Chinese song examples.
›What are the pros and cons of DiffRhythm?
Strengths
Trade-offs
›What are DiffRhythm’s key features?
›What are the best use cases for DiffRhythm?
›What is the pricing for DiffRhythm?
| Plan | Price | Details |
|---|---|---|
| Starter monthly | $6.99/month | Includes 18,000 credits, all four models, WAV and MP3 downloads, stem extraction, vocal removal, commercial licensing, and three concurrent generations. |
| Starter yearly | $83.9/year billed yearly | Includes the Starter plan's 18,000 credits, four models, downloads, stem and vocal tools, commercial licensing, and three concurrent generations. |
| Unlimited monthly | $59/month | Includes unlimited generation with all four models plus 18,000 bonus credits for advanced tools such as vocal separation and stem extraction. |
| Unlimited yearly | $712.9/year billed yearly | Includes the Unlimited plan's unlimited generation, four models, and 18,000 bonus credits for advanced tools. |
The pricing page says unused credits roll over. It also advertises 8-minute songs, which differs from the model page's stated maximum of 4 minutes 45 seconds.
Checked 2026-08-29 · source