-
公开(公告)号:US11789953B2
公开(公告)日:2023-10-17
申请号:US16979197
申请日:2019-03-13
Applicant: Semiconductor Energy Laboratory Co., Ltd.
Inventor: Kazuki Higashi , Junpei Momo
IPC: G06F16/2457 , G06F16/93 , G06N20/00 , G06F40/279 , G06F40/268 , G06N3/08 , G06Q10/10 , G06Q50/18
CPC classification number: G06F16/24578 , G06F16/93 , G06F40/268 , G06F40/279 , G06N3/08 , G06N20/00 , G06F2216/11 , G06Q10/10 , G06Q50/184
Abstract: A highly accurate document search, particularly a search for a document relating to intellectual property, is achieved with an easy input method. A document search system includes a processing portion. The processing portion has a function of extracting a keyword included in text data, a function of extracting a related term of the keyword from words included in a plurality of pieces of first reference text analysis data, a function of giving a weight to each of the keyword and the related term, a function of giving a score to each of a plurality of pieces of second reference text analysis data on the basis of the weight, a function of ranking the plurality of pieces of second reference text analysis data on the basis of the score to generate ranking data, and a function of outputting the ranking data.
-
公开(公告)号:US12019636B2
公开(公告)日:2024-06-25
申请号:US17064871
申请日:2020-10-07
Applicant: Semiconductor Energy Laboratory Co., Ltd.
Inventor: Kazuki Higashi , Junpei Momo
IPC: G06F16/2457 , G06F16/93 , G06F40/268 , G06F40/279 , G06N3/08 , G06N20/00 , G06Q10/10 , G06Q50/18
CPC classification number: G06F16/24578 , G06F16/93 , G06F40/268 , G06F40/279 , G06N3/08 , G06N20/00 , G06F2216/11 , G06Q10/10 , G06Q50/184
Abstract: A highly accurate document search, particularly a search for a document relating to intellectual property, is achieved with an easy input method. A document search system includes a processing portion. The processing portion has a function of extracting a keyword included in text data, a function of extracting a related term of the keyword from words included in a plurality of pieces of first reference text analysis data, a function of giving a weight to each of the keyword and the related term, a function of giving a score to each of a plurality of pieces of second reference text analysis data on the basis of the weight, a function of ranking the plurality of pieces of second reference text analysis data on the basis of the score to generate ranking data, and a function of outputting the ranking data.
-
公开(公告)号:US12169515B2
公开(公告)日:2024-12-17
申请号:US17612248
申请日:2020-05-11
Applicant: SEMICONDUCTOR ENERGY LABORATORY CO., LTD.
Inventor: Kunitaka Yamamoto , Junpei Momo , Kazuki Higashi
IPC: G06F16/35 , G06F16/93 , G06F40/268
Abstract: A document search system that enables efficient document search regardless of the ability of a user is achieved. Document search is performed using a document search system in which database document data is stored. After first document data and second document data are input to the document search system, the document search system extracts a plurality of terms from the first document data. The extraction of the terms is performed using morphological analysis, for example. Next, the extracted terms are weighted on the basis of the second document data. For example, texts included in a document represented by the second document data are classified into first and second texts. Among the terms extracted from the first document data, the weight of the term included in the first text is set larger than the weights of the other terms. The classification of the texts can be performed in accordance with a rule basis or using machine learning. After that, the similarity of the database document data to the first document data is calculated on the basis of the weighted term.
-
公开(公告)号:US12099543B2
公开(公告)日:2024-09-24
申请号:US17600280
申请日:2020-04-16
Applicant: Semiconductor Energy Laboratory Co., Ltd.
Inventor: Kazuki Higashi , Junpei Momo
IPC: G06F16/36 , G06F40/242 , G06F40/247 , G06F40/279
CPC classification number: G06F16/374 , G06F40/242 , G06F40/247 , G06F40/279
Abstract: Highly accurate document search, especially intellectual property-related document search, is achieved with a simple input method. A processing portion has a function of generating text analysis data from text data input to an input portion; a function of extracting a search word from words included in the text analysis data; and a function of generating first search data from the search word on the basis of weight dictionary data and thesaurus data. A memory portion stores second search data generated when the first search data is modified by a user. The processing portion updates the thesaurus data in accordance with the second search data.
-
公开(公告)号:US12086181B2
公开(公告)日:2024-09-10
申请号:US17791316
申请日:2020-12-28
Applicant: Semiconductor Energy Laboratory Co., Ltd.
Inventor: Junpei Momo , Kazuki Higashi , Motoki Nakashima
IPC: G06F16/00 , G06F16/33 , G06F16/901 , G06F16/93 , G06F40/211 , G06F40/284
CPC classification number: G06F16/9024 , G06F16/3344 , G06F16/93 , G06F40/211 , G06F40/284
Abstract: A document retrieval system retrieving a document with the concept of the document taken into account is provided. The system includes a processing portion and the processing portion creates a retrieval graph from a retrieval composition. The retrieval graph includes first to m-th retrieval local graphs (m is an integer of greater than or equal to 1), and the retrieval local graphs are each constituted by two nodes and one edge. The processing portion performs retrieval of first to m-th sentences on a reference document. The i-th sentence (i is an integer of greater than or equal to 1 and less than or equal to m) includes one of the two nodes in the i-th retrieval local graph or a related term or a hyponym of the one of the two nodes; the other of the two nodes in the i-th retrieval local graph or a related term or a hyponym of the other of the two nodes; and the edge in the i-th retrieval local graph or a related term or a hyponym of the edge. A mark is assigned to the score of the reference document in accordance with the number of sentences included in the reference document among the first to m-th sentences.
-
-
-
-