WO2026022528 - GENERATING MULTI-TRACK MUSIC FROM TEXT PROMPTS WITH DIFFUSION MODELS
National phase entry is expected:
Publication Number
WO/2026/022528
Publication Date
29.01.2026
International Application No.
PCT/IB2025/054129
International Filing Date
18.04.2025
Title **
[English]
GENERATING MULTI-TRACK MUSIC FROM TEXT PROMPTS WITH DIFFUSION MODELS
[French]
GÉNÉRATION DE MUSIQUE À PISTES MULTIPLES À PARTIR D'INVITES DE TEXTE AVEC DES MODÈLES DE DIFFUSION
Applicants **
FUTUREVERSE CORPORATION LIMITED
Inventors
WANG, Yijun
YAO, Yao
CHEN, Boyu
Priority Data
63/674,952
24.07.2024
US
19/182,474
17.04.2025
US
Application details
| Total Number of Claims/PCT | * |
| Number of Independent Claims | * |
| Number of Priorities | * |
| Number of Multi-Dependent Claims | * |
| Number of Drawings | * |
| Pages for Publication | * |
| Number of Pages with Drawings | * |
| Pages of Specification | * |
| * | |
| Number of Office Actions | * |
| * | |
International Searching Authority |
USPTO
* |
| * | |
| Recordal of a Change of the Applicant's Name/Address |
Change of Applicant's Name and Address
* |
| Type of Assignment |
The Standard Agent's Assignment
* |
| Applicant's Legal Status |
Legal Entity
* |
| * | |
| * | |
| * | |
| * | |
| * | |
| Entry into National Phase under |
Chapter I
* |
| Patent Delivery |
Send the Letters Patent by Courier
* |
| Translation |
|
* The data is based on automatic recognition. Please verify and amend if necessary.
** IP-Coster compiles data from publicly available sources. If this data includes your personal information, you can contact us to request its removal.
Quotation for National Phase entry
| Country | Stages | Total | |
|---|---|---|---|
| China | Filing, Examination, Granting | 2544 | |
| EPO | Filing, Examination, Granting | 16662 | |
| Japan | Filing, Examination, Granting | 2450 | |
| South Korea | Filing, Examination, Granting | 2736 | |
| USA | Filing, Examination, Granting | 5110 |

Total:
29,502
Contact Us
Abstract[English]
Methods, systems, and devices for multi-track music generation are described. In some examples, a method includes receiving a text prompt describing desired musical attributes and generating, using a diffusion model, multiple audio tracks based on the text prompt, wherein each audio track corresponds to a different musical component. The method can further include assigning individual timestep vectors respectively to each of multiple audio tracks and generating, using a diffusion model, one or more enhanced audio tracks based on the individual timestep vectors and corresponding audio track. Finally, the method can include combining the generated audio tracks to produce a multi-track musical composition.[French]
L'invention concerne des procédés, des systèmes et des dispositifs de génération de musique à pistes multiples. Dans certains exemples, un procédé consiste à recevoir une invite de texte décrivant des attributs musicaux souhaités et à générer, à l'aide d'un modèle de diffusion, de multiples pistes audio sur la base de l'invite de texte, chaque piste audio correspondant à un composant musical différent. Le procédé peut en outre consister à attribuer des vecteurs de pas de temps individuels respectivement à chaque piste parmi de multiples pistes audio et à générer, à l'aide d'un modèle de diffusion, une ou plusieurs pistes audio améliorées sur la base des vecteurs de pas de temps individuels et de la piste audio correspondante. Enfin, le procédé peut comprendre la combinaison des pistes audio générées pour produire une composition musicale à pistes multiples.