WO2025248487 - EVENT-BASED REINFORCEMENT LEARNING FOR RRM PARAMETER OPTIMIZATION
National phase entry is expected:
Publication Number
WO/2025/248487
Publication Date
04.12.2025
International Application No.
PCT/IB2025/055572
International Filing Date
29.05.2025
Title **
[English]
EVENT-BASED REINFORCEMENT LEARNING FOR RRM PARAMETER OPTIMIZATION
[French]
APPRENTISSAGE PAR RENFORCEMENT BASÉ SUR DES ÉVÉNEMENTS POUR L'OPTIMISATION DE PARAMÈTRES RRM
Applicants **
NOKIA TECHNOLOGIES OY
Inventors
SONG, Jian
FEKI, Afef
HÖHNE, Hans Thomas
ALI-TOLPPA, Janne
VEIJALAINEN, Teemu Mikael
DOSTI, Endrit
ALI, Samad
KHATIBI, Sina
Priority Data
20245702
31.05.2024
FI
Application details
| Total Number of Claims/PCT | * |
| Number of Independent Claims | * |
| Number of Priorities | * |
| Number of Multi-Dependent Claims | * |
| Number of Drawings | * |
| Pages for Publication | * |
| Number of Pages with Drawings | * |
| Pages of Specification | * |
| * | |
| Number of Office Actions | * |
| * | |
International Searching Authority |
EPO
* |
| Recordal of a Change of the Applicant's Name/Address |
Change of Applicant's Name and Address
* |
| Type of Assignment |
The Standard Agent's Assignment
* |
| Applicant's Legal Status |
Legal Entity
* |
| * | |
| * | |
| * | |
| * | |
| * | |
| Entry into National Phase under |
Chapter I
* |
| Patent Delivery |
Send the Letters Patent by Courier
* |
| Translation |
|
* The data is based on automatic recognition. Please verify and amend if necessary.
** IP-Coster compiles data from publicly available sources. If this data includes your personal information, you can contact us to request its removal.
Quotation for National Phase entry
| Country | Stages | Total | |
|---|---|---|---|
| China | Filing, Examination, Granting | 2367 | |
| EPO | Filing, Examination, Granting | 9444 | |
| Japan | Filing, Examination, Granting | 2155 | |
| South Korea | Filing, Examination, Granting | 2071 | |
| USA | Filing, Examination, Granting | 4740 |

Total:
20,777
Contact Us
Abstract[English]
According to an aspect, there is provided an apparatus for performing the following. The apparatus transmits, to a network entity, a configuration request requesting configuration of one or more reinforcement learning, RL, strategies for an RL model. The apparatus receives, from the network entity, at least one configuration message comprising the one or more RL strategies which comprise one or more RL event conditions for entering and/or exiting exploration and/or exploitation events. The apparatus performs exploration and/or exploitation using the RL model for determining one or more radio resource management parameters of the apparatus based on one or more radio measurement metrics. The apparatus evaluates the one or more RL event conditions. Based on the results of the evaluating, the apparatus performs triggering the exploration or exploitation event and/or exiting the exploration or the exploitation event.[French]
Selon un aspect, l'invention concerne un appareil pour mettre en œuvre les étapes suivantes : l'appareil transmet à une entité de réseau une demande de configuration demandant la configuration d'une ou de plusieurs stratégies d'apprentissage par renforcement (RL) pour un modèle RL ; l'appareil reçoit de l'entité de réseau au moins un message de configuration comprenant la ou les stratégies RL qui comprennent une ou plusieurs conditions d'événement RL liées à l'entrée et/ou la sortie d'événements d'exploration et/ou d'exploitation ; l'appareil réalise une exploration et/ou une exploitation à l'aide du modèle RL pour déterminer un ou plusieurs paramètres de gestion de ressources radio de l'appareil sur la base d'une ou de plusieurs mesures radio ; l'appareil évalue la ou les conditions d'événement RL ; sur la base des résultats de l'évaluation, l'appareil déclenche l'événement d'exploration ou d'exploitation et/ou la sortie de l'événement d'exploration ou d'exploitation.