WO2026009085 - TECHNIQUES FOR CLASSIFYING DATA USING LARGE LANGUAGE MODELS
National phase entry is expected:
Publication Number
WO/2026/009085
Publication Date
08.01.2026
International Application No.
PCT/IB2025/056411
International Filing Date
24.06.2025
Title **
[English]
TECHNIQUES FOR CLASSIFYING DATA USING LARGE LANGUAGE MODELS
[French]
TECHNIQUES DE CLASSIFICATION DE DONNÉES À L'AIDE DE GRANDS MODÈLES DE LANGAGE
Applicants **
CYERA, LTD.
Inventors
SEGEV, Yotam
BAR-ILAN, Itamar
ITAI, Yonatan
BARELI, Shiran
NIKITIN, Andrey
VERED, Guye
SHAKED, Michal
HOROVITZ, Dvir
Priority Data
18/763,531
03.07.2024
US
Application details
| Total Number of Claims/PCT | * |
| Number of Independent Claims | * |
| Number of Priorities | * |
| Number of Multi-Dependent Claims | * |
| Number of Drawings | * |
| Pages for Publication | * |
| Number of Pages with Drawings | * |
| Pages of Specification | * |
| * | |
| Number of Office Actions | * |
| * | |
International Searching Authority |
ILPO
* |
| Recordal of a Change of the Applicant's Name/Address |
Change of Applicant's Name and Address
* |
| Type of Assignment |
The Standard Agent's Assignment
* |
| Applicant's Legal Status |
Legal Entity
* |
| * | |
| * | |
| * | |
| * | |
| * | |
| Entry into National Phase under |
Chapter I
* |
| Patent Delivery |
Send the Letters Patent by Courier
* |
| Translation |
|
* The data is based on automatic recognition. Please verify and amend if necessary.
** IP-Coster compiles data from publicly available sources. If this data includes your personal information, you can contact us to request its removal.
Quotation for National Phase entry
| Country | Stages | Total | |
|---|---|---|---|
| China | Filing, Examination, Granting | 2425 | |
| EPO | Filing, Examination, Granting | 16070 | |
| Japan | Filing, Examination, Granting | 2425 | |
| South Korea | Filing, Examination, Granting | 2638 | |
| USA | Filing, Examination, Granting | 5340 |

Total:
28,898
Contact Us
Abstract[English]
A system and method for classification. A method includes identifying candidate entities among text data by applying at least one entity identification rule to the text data. Inputs are constructed based on the identified candidate entities, where each input includes a first portion of text indicating a candidate entity and at least one second portion of text and where the at least one second portion of text of each input is adjacent to the first portion of text of the input. Multiple language models are applied to the inputs, where each language model is trained to identify a respective set of entities and where outputs of the language models include at least one portion of entity-indicating text for each input. Based on the outputs of the language models, at least one named entity in the text data is determined.[French]
L'invention concerne un système et un procédé de classification. Un procédé consiste à identifier des entités candidates parmi des données de texte par application d'au moins une règle d'identification d'entité aux données de texte. Des entrées sont construites sur la base des entités candidates identifiées, chaque entrée comprenant une première partie de texte indiquant une entité candidate et au moins une seconde partie de texte et la ou les secondes parties de texte de chaque entrée étant adjacentes à la première partie de texte de l'entrée. De multiples modèles de langage sont appliqués aux entrées, chaque modèle de langage étant entraîné pour identifier un ensemble respectif d'entités et les sorties des modèles de langage comprenant au moins une partie du texte indiquant l'entité pour chaque entrée. Sur la base des sorties des modèles de langage, au moins une entité nommée dans les données de texte est déterminée.