Jump to content

SoundHound AI IP, LLC. (20250014582). METHOD AND SYSTEM FOR CONVERSATION TRANSCRIPTION WITH METADATA

From WikiPatents

METHOD AND SYSTEM FOR CONVERSATION TRANSCRIPTION WITH METADATA

Organization Name

SoundHound AI IP, LLC.

Inventor(s)

Kiersten L. Bradley of Santa Clara CA US

Ethan Coeytaux of Boulder CO US

Ziming Yin of Toronto CA

METHOD AND SYSTEM FOR CONVERSATION TRANSCRIPTION WITH METADATA

This abstract first appeared for US patent application 20250014582 titled 'METHOD AND SYSTEM FOR CONVERSATION TRANSCRIPTION WITH METADATA

Original Abstract Submitted

methods and systems for enabling an efficient review of meeting content via a metadata-enriched, speaker-attributed and multiuser-editable transcript are disclosed. by incorporating speaker diarization and other metadata, the system can provide a structured and effective way to review and/or edit the transcript by one or more editors. one type of metadata can be image or video data to represent the meeting content. furthermore, the present subject matter utilizes a multimodal diarization model to identify and label different speakers. the system can synchronize various sources of data, e.g., audio channel data, voice feature vectors, acoustic beamforming, image identification, and extrinsic data, to implement speaker diarization.