WO2026009132 - GENERATION OF ACOUSTICS-MATCHED AUDIO FOR THREE-DIMENSIONAL (3D) SCENE ASSOCIATED WITH METAVERSE APPLICATION
National phase entry is expected:
Publication Number
WO/2026/009132
Publication Date
08.01.2026
International Application No.
PCT/IB2025/056635
International Filing Date
30.06.2025
Title **
[English]
GENERATION OF ACOUSTICS-MATCHED AUDIO FOR THREE-DIMENSIONAL (3D) SCENE ASSOCIATED WITH METAVERSE APPLICATION
[French]
GÉNÉRATION D'AUDIO ADAPTÉ À L'ACOUSTIQUE POUR UNE SCÈNE TRIDIMENSIONNELLE (3D) ASSOCIÉE À UNE APPLICATION DE MÉTAVERS
Applicants **
SONY GROUP CORPORATION
Inventors
MOHAMED, Shehnaz
VENKATESWARAN SABARITA, Kalpana
Priority Data
202411051082
03.07.2024
IN
Application details
| Total Number of Claims/PCT | * |
| Number of Independent Claims | * |
| Number of Priorities | * |
| Number of Multi-Dependent Claims | * |
| Number of Drawings | * |
| Pages for Publication | * |
| Number of Pages with Drawings | * |
| Pages of Specification | * |
| * | |
| Number of Office Actions | * |
| * | |
International Searching Authority |
EPO
* |
| Recordal of a Change of the Applicant's Name/Address |
Change of Applicant's Name and Address
* |
| Type of Assignment |
The Standard Agent's Assignment
* |
| Applicant's Legal Status |
Legal Entity
* |
| * | |
| * | |
| * | |
| * | |
| * | |
| Entry into National Phase under |
Chapter I
* |
| Patent Delivery |
Send the Letters Patent by Courier
* |
| Translation |
|
* The data is based on automatic recognition. Please verify and amend if necessary.
** IP-Coster compiles data from publicly available sources. If this data includes your personal information, you can contact us to request its removal.
Quotation for National Phase entry
| Country | Stages | Total | |
|---|---|---|---|
| China | Filing, Examination, Granting | 2640 | |
| EPO | Filing, Examination, Granting | 12581 | |
| Japan | Filing, Examination, Granting | 2328 | |
| South Korea | Filing, Examination, Granting | 2345 | |
| USA | Filing, Examination, Granting | 4740 |

Total:
24,634
Contact Us
Abstract[English]
An electronic device and method for generation of immersive metaverse audio experiences is disclosed. The electronic device receives multi-view images corresponding to a three-dimensional (3D) scene associated with a first user of a metaverse application. The electronic device applies a deep learning model to the multi-view images and determines volume aggregation information associated with the 3D scene. The electronic device receives source audio associated with the first user and applies cross-modal encoder model to the volume aggregation information and the source audio. The electronic device generates acoustics-matched audio associated with the 3D scene based on the applied cross-modal encoder model. The electronic device controls the display device to render the 3D scene for a second user of the metaverse application based on acoustics-matched audio.[French]
Il est divulgué un dispositif électronique et un procédé de génération d'expériences audio de métavers immersives. Le dispositif électronique reçoit des images multivues correspondant à une scène tridimensionnelle (3D) associée à un premier utilisateur d'une application de métavers. Le dispositif électronique applique un modèle d'apprentissage profond aux images multivues et détermine des informations d'agrégation de volumes associées à la scène 3D. Le dispositif électronique reçoit un audio source associé au premier utilisateur et applique un modèle de codeur intermodal aux informations d'agrégation de volumes et à l'audio source. Le dispositif électronique génère un audio adapté à l'acoustique associé à la scène 3D sur la base du modèle de codeur intermodal appliqué. Le dispositif électronique commande au dispositif d'affichage de réaliser un rendu de la scène 3D pour un second utilisateur de l'application de métavers sur la base de l'audio adapté à l'acoustique.