Method and system for conversation transcription with metadata

发明授权

US12125487B2 Method and system for conversation transcription with metadata 有权

请登陆查看更多内容

专利标题： Method and system for conversation transcription with metadata
申请号： US17450551

申请日： 2021-10-11
公开(公告)号： US12125487B2

公开(公告)日： 2024-10-22
发明人: Kiersten L. Bradley , Ethan Coeytaux , Ziming Yin
申请人： SoundHound, Inc.
申请人地址： US CA Santa Clara
专利权人： SoundHound AI IP, LLC.
当前专利权人： SoundHound AI IP, LLC.
当前专利权人地址： US CA Santa Clara
代理机构： Platinum Intellectual Property
主分类号： G10L15/26
IPC分类号： G10L15/26 ; G06F40/134 ; G06F40/166 ; G06F40/284 ; G10L15/02 ; G10L15/06 ; G10L15/07

Method and system for conversation transcription with metadata

摘要：

Methods and systems for enabling an efficient review of meeting content via a metadata-enriched, speaker-attributed and multiuser-editable transcript are disclosed. By incorporating speaker diarization and other metadata, the system can provide a structured and effective way to review and/or edit the transcript by one or more editors. One type of metadata can be image or video data to represent the meeting content. Furthermore, the present subject matter utilizes a multimodal diarization model to identify and label different speakers. The system can synchronize various sources of data, e.g., audio channel data, voice feature vectors, acoustic beamforming, image identification, and extrinsic data, to implement speaker diarization.

公开/授权文献

US20220115019A1 METHOD AND SYSTEM FOR CONVERSATION TRANSCRIPTION WITH METADATA 公开/授权日：2022-04-14

信息查询

Espacenet

IPC分类:

G	物理
G10	乐器；声学
G10L	语音分析或合成；语音识别；语音或声音处理；语音或音频编码或解码
G10L15/00	语音识别（G10L17/00优先）
G10L15/26	.语音—正文识别系统（G10L15/08优先）