SYSTEM AND METHOD FOR ACCENT-AGNOSTIC FRAME-LEVEL WAKE WORD DETECTION
Abstract:
A method includes accessing, using at least one processor of an electronic device, a machine learning model. The machine learning model is a trained student model that is trained using audio samples in a plurality of accent types. The method also includes receiving, using the at least one processor, an audio input from an audio input device. The method further includes providing, using the at least one processor, the audio input to the trained student model. The method also includes receiving, using the at least one processor, an output from the trained student model including frame-level probabilities associated with the audio input. In addition, the method includes instructing, using the at least one processor, at least one action based on the frame-level probabilities associated with the audio input.
Public/Granted literature
Information query
Patent Agency Ranking
0/0