SYSTEM AND METHOD FOR COMBINING FRAME AND SEGMENT LEVEL PROCESSING, VIA TEMPORAL POOLING, FOR PHONETIC CLASSIFICATION
    1.
    发明申请
    SYSTEM AND METHOD FOR COMBINING FRAME AND SEGMENT LEVEL PROCESSING, VIA TEMPORAL POOLING, FOR PHONETIC CLASSIFICATION 有权
    用于组合框架和分段水平处理的系统和方法,通过时间池,用于电话分类

    公开(公告)号:US20130103402A1

    公开(公告)日:2013-04-25

    申请号:US13281102

    申请日:2011-10-25

    IPC分类号: G10L15/04 G10L15/00

    CPC分类号: G10L15/02 G10L15/08 G10L15/16

    摘要: Disclosed herein are systems, methods, and non-transitory computer-readable storage media for combining frame and segment level processing, via temporal pooling, for phonetic classification. A frame processor unit receives an input and extracts the time-dependent features from the input. A plurality of pooling interface units generates a plurality of feature vectors based on pooling the time-dependent features and selecting a plurality of time-dependent features according to a plurality of selection strategies. Next, a plurality of segmental classification units generates scores for the feature vectors. Each segmental classification unit (SCU) can be dedicated to a specific pooling interface unit (PIU) to form a PIU-SCU combination. Multiple PIU-SCU combinations can be further combined to form an ensemble of combinations, and the ensemble can be diversified by varying the pooling operations used by the PIU-SCU combinations. Based on the scores, the plurality of segmental classification units selects a class label and returns a result.

    摘要翻译: 本文公开了用于通过时间池来组合帧和段级处理用于语音分类的系统,方法和非暂时的计算机可读存储介质。 帧处理器单元接收输入并从输入中提取与时间相关的特征。 多个池化接口单元基于集合时间依赖特征并根据多个选择策略选择多个时间相关特征来生成多个特征向量。 接下来,多个分段分类单元生成特征向量的得分。 每个分段分类单元(SCU)可专用于特定的汇聚接口单元(PIU)以形成PIU-SCU组合。 可以进一步组合多个PIU-SCU组合以形成组合的集合,并且可以通过改变PIU-SCU组合使用的合并操作来使集合多样化。 基于分数,多个分段分类单元选择分类标签并返回结果。