Patent search ap:("GOOGLE LLC") AND inv:"Pu-sen Chao" Page 1

1.

发明授权
Hot-word free pre-emption of automated assistant response presentation 有权

公开(公告)号：US12125477B2

公开(公告)日：2024-10-22

申请号：US18235726

申请日：2023-08-18

Applicant: GOOGLE LLC

Inventor： Pu-sen Chao , Alex Fandrianto

IPC: G10L15/16 , G06N3/08 , G10L15/08 , G10L17/00 , G10L21/0208

CPC classification number: G10L15/16 , G06N3/08 , G10L17/00 , G10L21/0208 , G10L2015/088 , G10L2021/02082

Abstract: The presentation of an automated assistant response may be selectively pre-empted in response to a hot-word free utterance that is received during the presentation and that is determined to be likely directed to the automated assistant. The determination that the utterance is likely directed to the automated assistant may be performed, for example, using an utterance classification operation that is performed on audio data received during presentation of the response, and based upon such a determination, the response may be pre-empted with another response associated with the later-received utterance. In addition, the duration that is used to determine when a session should be terminated at the conclusion of a conversation between a user and an automated assistant may be dynamically controlled based upon when the presentation of a response has completed.

2.

发明申请
TEXT INDEPENDENT SPEAKER RECOGNITION 有权

公开(公告)号：US20230113617A1

公开(公告)日：2023-04-13

申请号：US18078476

申请日：2022-12-09

Applicant: GOOGLE LLC

Inventor： Pu-sen Chao , Diego Melendo Casado , Ignacio Lopez Moreno , Quan Wang

IPC: G10L15/06 , G10L15/07 , G10L15/22 , G10L15/32 , G10L17/24

Abstract: Text independent speaker recognition models can be utilized by an automated assistant to verify a particular user spoke a spoken utterance and/or to identify the user who spoke a spoken utterance. Implementations can include automatically updating a speaker embedding for a particular user based on previous utterances by the particular user. Additionally or alternatively, implementations can include verifying a particular user spoke a spoken utterance using output generated by both a text independent speaker recognition model as well as a text dependent speaker recognition model. Furthermore, implementations can additionally or alternatively include prefetching content for several users associated with a spoken utterance prior to determining which user spoke the spoken utterance.

3.

发明授权
Text independent speaker recognition 有权

公开(公告)号：US11527235B2

公开(公告)日：2022-12-13

申请号：US17046994

申请日：2019-12-02

Applicant: Google LLC

Inventor： Pu-sen Chao , Diego Melendo Casado , Ignacio Lopez Moreno , Quan Wang

IPC: G10L15/06 , G10L15/07 , G10L15/22 , G10L15/32 , G10L17/24

Abstract: Text independent speaker recognition models can be utilized by an automated assistant to verify a particular user spoke a spoken utterance and/or to identify the user who spoke a spoken utterance. Implementations can include automatically updating a speaker embedding for a particular user based on previous utterances by the particular user. Additionally or alternatively, implementations can include verifying a particular user spoke a spoken utterance using output generated by both a text independent speaker recognition model as well as a text dependent speaker recognition model. Furthermore, implementations can additionally or alternatively include prefetching content for several users associated with a spoken utterance prior to determining which user spoke the spoken utterance.

4.

发明授权
Multi-user authentication on a device 有权

公开(公告)号：US11238848B2

公开(公告)日：2022-02-01

申请号：US16709132

申请日：2019-12-10

Applicant: Google LLC

Inventor： Meltem Oktem , Taral Pradeep Joglekar , Fnu Heryandi , Pu-sen Chao , Ignacio Lopez Moreno , Salil Rajadhyaksha , Alexander H. Gruenstein , Diego Melendo Casado

IPC: G10L15/08 , G06F21/32 , G10L17/06 , G06F16/635 , G10L15/22 , G10L17/00 , G06K9/00 , G10L15/07 , G10L15/26

Abstract: In some implementations, authentication tokens corresponding to known users of a device are stored on the device. An utterance from a speaker is received. The speaker of the utterance is classified as not a known user of the device. A query that includes the authentication tokens that correspond to known users of the device, a representation of the utterance, and an indication that the speaker was classified as not a known user of the device is provided to the server. A response to the query is received at the device and from the server based on the query.

5.

发明申请
AUTOMATICALLY DETERMINING LANGUAGE FOR SPEECH RECOGNITION OF SPOKEN UTTERANCE RECEIVED VIA AN AUTOMATED ASSISTANT INTERFACE 有权

公开(公告)号：US20210097981A1

公开(公告)日：2021-04-01

申请号：US17120906

申请日：2020-12-14

Applicant: Google LLC

Inventor： Pu-sen Chao , Diego Melendo Casado , Ignacio Lopez Moreno

IPC: G10L15/14 , G10L15/02 , G10L15/18 , G06F3/16 , G10L15/00 , G10L15/183 , G10L15/22 , G10L15/30

Abstract: Implementations relate to determining a language for speech recognition of a spoken utterance, received via an automated assistant interface, for interacting with an automated assistant. Implementations can enable multilingual interaction with the automated assistant, without necessitating a user explicitly designate a language to be utilized for each interaction. Selection of a speech recognition model for a particular language can based on one or more interaction characteristics exhibited during a dialog session between a user and an automated assistant. Such interaction characteristics can include anticipated user input types, anticipated user input durations, a duration for monitoring for a user response, and/or an actual duration of a provided user response.

6.

发明申请
ADAPTIVE INTERFACE IN A VOICE-BASED NETWORKED SYSTEM 审中-公开

公开(公告)号：US20190318729A1

公开(公告)日：2019-10-17

申请号：US15973461

申请日：2018-05-07

Applicant: Google LLC

Inventor： Pu-sen Chao , Diego Melendo Casado , Ignacio Lopez Moreno

IPC: G10L15/18 , G10L15/08

Abstract: Determining a language for speech recognition of a spoken utterance received via an automated assistant interface for interacting with an automated assistant. The system can enable multilingual interaction with the automated assistant, without necessitating a user explicitly designate a language to be utilized for each interaction. The system can determine a user profile that corresponds to audio data that captures a spoken utterance, and utilize language(s), and optionally corresponding probabilities, assigned to the user profile in determining a language for speech recognition of the spoken utterance. The system can perform speech recognition in each of multiple languages assigned to the user profile, and utilize criteria to select only one of the speech recognitions as appropriate for generating and providing content that is responsive to the spoken utterance.

7.

发明申请
MULTI-USER AUTHENTICATION ON A DEVICE 审中-公开

公开(公告)号：US20180308491A1

公开(公告)日：2018-10-25

申请号：US15956350

申请日：2018-04-18

Applicant: Google LLC

Inventor： Meltem Oktem , Taral Pradeep Joglekar , Fnu Heryandi , Pu-sen Chao , Ignacio Lopez Moreno , Salil Rajadhyaksha , Alexander H. Gruenstein , Diego Melendo Casado

IPC: G10L17/06 , G10L15/07 , G06F21/32 , G10L15/08 , G06F17/30 , G06K9/00

Abstract: In some implementations, authentication tokens corresponding to known users of a device are stored on the device. An utterance from a speaker is received. The utterance is classified as spoken by a particular known user of the known users. A query that includes a representation of the utterance and an indication of the particular known user as the speaker is provided using the authentication token of the particular known user.

8.

发明授权
Assessing speaker recognition performance 有权

公开(公告)号：US12154574B2

公开(公告)日：2024-11-26

申请号：US18506105

申请日：2023-11-09

Applicant: Google LLC

Inventor： Jason Pelecanos , Pu-sen Chao , Yiling Huang , Quan Wang

IPC: G10L17/12 , G06N3/045 , G06N3/08 , G10L17/18 , G10L25/30 , G10L25/51

Abstract: A method for evaluating a verification model includes receiving a first and a second set of verification results where each verification result indicates whether a primary model or an alternative model verifies an identity of a user as a registered user. The method further includes identifying each verification result in the first and second sets that includes a performance metric. The method also includes determining a first score of the primary model based on a number of the verification results identified in the first set that includes the performance metric and determining a second score of the alternative model based on a number of the verification results identified in the second set that includes the performance metric. The method further includes determining whether a verification capability of the alternative model is better than a verification capability of the primary model based on the first score and the second score.

9.

发明授权
Automatically determining language for speech recognition of spoken utterance received via an automated assistant interface 有权

公开(公告)号：US12046233B2

公开(公告)日：2024-07-23

申请号：US18361408

申请日：2023-07-28

Applicant: GOOGLE LLC

Inventor： Pu-sen Chao , Diego Melendo Casado , Ignacio Lopez Moreno , William Zhang

IPC: G10L15/22 , G10L13/00 , G10L15/00 , G10L15/08 , G10L15/14 , G10L15/18 , G10L15/197 , G10L15/30

CPC classification number: G10L15/197 , G10L13/00 , G10L15/005 , G10L15/08 , G10L15/14 , G10L15/1822 , G10L15/22 , G10L15/30 , G10L2015/088 , G10L2015/223 , G10L2015/228

Abstract: Determining a language for speech recognition of a spoken utterance received via an automated assistant interface for interacting with an automated assistant. Implementations can enable multilingual interaction with the automated assistant, without necessitating a user explicitly designate a language to be utilized for each interaction. Implementations determine a user profile that corresponds to audio data that captures a spoken utterance, and utilize language(s), and optionally corresponding probabilities, assigned to the user profile in determining a language for speech recognition of the spoken utterance. Some implementations select only a subset of languages, assigned to the user profile, to utilize in speech recognition of a given spoken utterance of the user. Some implementations perform speech recognition in each of multiple languages assigned to the user profile, and utilize criteria to select only one of the speech recognitions as appropriate for generating and providing content that is responsive to the spoken utterance.

10.

发明授权
Assessing speaker recognition performance 有权

公开(公告)号：US11837238B2

公开(公告)日：2023-12-05

申请号：US17076743

申请日：2020-10-21

Applicant: Google LLC

Inventor： Jason Pelecanos , Pu-sen Chao , Yiling Huang , Quan Wang

IPC: G10L17/12 , G06N3/08 , G10L17/18 , G10L25/30 , G10L25/51 , G06N3/04 , G06N3/045

CPC classification number: G10L17/12 , G06N3/045 , G06N3/08 , G10L17/18 , G10L25/30 , G10L25/51

Abstract: A method for evaluating a verification model includes receiving a first and a second set of verification results where each verification result indicates whether a primary model or an alternative model verifies an identity of a user as a registered user. The method further includes identifying each verification result in the first and second sets that includes a performance metric. The method also includes determining a first score of the primary model based on a number of the verification results identified in the first set that includes the performance metric and determining a second score of the alternative model based on a number of the verification results identified in the second set that includes the performance metric. The method further includes determining whether a verification capability of the alternative model is better than a verification capability of the primary model based on the first score and the second score.

Search Results

Country/Region

Patent validity

Application date

Publication (announcement) day

applicant

The country/region where the applicant is located

Inventor

IPC

IPC Department

IPC class

IPC subclass

IPC group

IPC team

Appearance classification