- 专利标题: METHOD AND SYSTEM FOR CONVERSATION TRANSCRIPTION WITH METADATA
-
申请号: US18743562申请日: 2024-06-14
-
公开(公告)号: US20240331702A1公开(公告)日: 2024-10-03
- 发明人: Kiersten L. BRADLEY , Ethan COEYTAUX , Ziming YIN
- 申请人: SoundHound AI IP, LLC.
- 申请人地址: US CA Santa Clara
- 专利权人: SoundHound AI IP, LLC.
- 当前专利权人: SoundHound AI IP, LLC.
- 当前专利权人地址: US CA Santa Clara
- 主分类号: G10L15/26
- IPC分类号: G10L15/26 ; G06F40/134 ; G06F40/166 ; G06F40/284 ; G10L15/02 ; G10L15/06 ; G10L15/07
摘要:
Methods and systems for enabling an efficient review of meeting content via a metadata-enriched, speaker-attributed transcript are disclosed. By incorporating speaker diarization and other metadata, the system can provide a structured and effective way to review and/or edit the transcript. One type of metadata can be image or video data to represent the meeting content. Furthermore, the present subject matter utilizes a multimodal diarization model to identify and label different speakers. The system can synchronize various sources of data, e.g., audio channel data, voice feature vectors, acoustic beamforming, image identification, and extrinsic data, to implement speaker diarization.
公开/授权文献
- US3212114A Multiple part fastener assembly machine 公开/授权日:1965-10-19
信息查询