摘要:
A method for adjusting audio input signal gain in a speech system can include seven steps. First, an upper and a lower threshold can be predetermined in which the upper and lower threshold define an optimal range of audio data signal amplitude measurements. Second, a frame of unpredicted digital audio data samples can be received. Each sample can indicate an amplitude measurement of the audio data signal at a particular point in time. Third, a maximum signal amplitude can be calculated for a configurable measurement percentile of the unpredicted digital audio data samples. Fourth, the audio input signal gain can be incrementally adjusted downward if the maximum signal amplitude exceeds the upper threshold. Conversely, fifth, the audio input signal gain can be incrementally adjusted upward if the maximum signal amplitude falls below the lower threshold. Sixth, additional frames of unpredicted digital audio data samples can be received. Finally, seventh, each of the third through the sixth steps can be repeated with the received additional frames until the calculated maximum signal amplitude falls within the optimal range of audio signal amplitude.
摘要:
A device and method are disclosed which perform diagnostics on a microphone and display diagnostic information and instructions to a user. The invention uses a processor to create histograms of the PCM (Pulse Code Modulation) signal after removing any dc bias to determine signal and noise levels and ratios, as well as other parameters. Messages are generated and displayed by the device and method to inform a user that the microphone is working correctly or about possible malfunctions, such as low gain. The messages can advise the user on steps to take to correct the malfunctions, for example, to try a different adapter cable or plug.
摘要:
An automatic gain control method in accordance with the inventive arrangements can include the following steps. Initially, an audio signal can be provided to an audio device which has a range of permissible signal level settings and a signal level controller for establishing a particular signal level setting. In addition, an actual signal level can be measured for the audio signal at an established signal level setting. The measured actual signal level further can be stored in a volume map along with the corresponding established signal level setting. Following the storage of the measured actual signal level in the volume map, a different signal level setting can be established using the signal level controller. Subsequently, the actual signal level can be re-measured and the re-measured actual signal level and corresponding established different signal level setting can be stored in the volume map. Finally, the volume map can be used during an audio processing session to determine a signal level setting for the audio device, wherein the signal level setting corresponds to a desired actual audio signal level. In one aspect of the present invention, the method can also include detecting a hysteresis condition in the volume map.
摘要:
A method of automatically adjusting volume of speech generated by a text-to-speech application can include measuring an ambient noise level of an audio environment. A target volume for speech output generated by a text-to-speech application can be calculated based in part upon the ambient noise level. A volume of speech generated by the text-to-speech application can be automatically adjusted responsive to the performed calculation.
摘要:
A method of optimizing audio input for speech recognition applications can include identifying a source waveform and at least one optimization parameter, wherein the optimization parameter is configured to adjust audio input to a speech recognition application. The source waveform can be modified according to the optimization parameter resulting in a modified waveform. At least one optimization parameter can be synchronized with the source waveform. At least two time dependant graphs can be displayed, where the time dependant graphs can include the source waveform, the modified waveform, and/or a graph for the optimization parameter plotted against time.
摘要:
The present invention records a loudness-level-reference segment of audio when creating speech audio files and audio files including background sounds. The speech audio files can then be combined with the background sound containing audio files in any desirable combination. When combining the files, the relative audio level of the files is matched, by matching the loudness-level-reference segments with each other. Any of a variety of known digital signal processing techniques can be used to normalize the component audio files. The combined audio files containing speech and background sounds (e.g. ambient noise) having matching relative audio levels can be used to test and/or train a speech recognition engine or a speech processing system.
摘要:
A method of optimizing audio input for speech recognition applications can include identifying a source waveform and at least one optimization parameter, wherein the optimization parameter is configured to adjust audio input to a speech recognition application. The source waveform can be modified according to the optimization parameter resulting in a modified waveform. At least one optimization parameter can be synchronized with the source waveform. At least two time dependant graphs can be displayed, where the time dependant graphs can include the source waveform, the modified waveform, and/or a graph for the optimization parameter plotted against time.
摘要:
A portable paperless parcel tracking system capable of reading bar codes on packages, capturing signatures and alphanumeric data related to the parcels when entered into a touch panel display, disabling manual or finger touch entry when a stylus is near the surface of the display, storing the parcel related data electronically and transmitting the data to a host at a convenient time.
摘要:
A method of optimizing audio input for speech recognition applications can include identifying a source waveform and at least one optimization parameter, wherein the optimization parameter is configured to adjust audio input to a speech recognition application. The source waveform can be modified according to the optimization parameter resulting in a modified waveform. At least one optimization parameter can be synchronized with the source waveform. At least two time dependant graphs can be displayed, where the time dependant graphs can include the source waveform, the modified waveform, and/or a graph for the optimization parameter plotted against time.
摘要:
A stylus assembly in which a stylus is mounted in a housing for movement relative thereto. An electrical circuit is provided for sensing contact of the stylus with a writing surface and measuring the pressure exerted on the surface by the stylus.