Patents by Inventor Ethan Coeytaux

Ethan Coeytaux has filed for patents to protect the following inventions. This listing includes patent applications that are pending as well as patents that have already been granted by the United States Patent and Trademark Office (USPTO).

VIDEO CONFERENCE CAPTIONING

Publication number: 20230245661

Abstract: A video conferencing system, such as one implemented with a cloud server, receives audio streams from a plurality of endpoints. The system uses automatic speech recognition to transcribe speech in the audio streams. The system multiplexes the transcriptions into individual caption streams and sends them to the endpoints, but the caption stream to each endpoint omits the transcription of audio from the endpoint. Some systems allow muting of audio through an indication to the system. The system then omits sending the muted audio to other endpoints and also omits sending a transcription of the muted audio to other endpoints.

Type: Application

Filed: April 10, 2023

Publication date: August 3, 2023

Applicant: SoundHound, Inc.

Inventor: Ethan COEYTAUX
METHOD AND SYSTEM FOR CONVERSATION TRANSCRIPTION WITH METADATA

Publication number: 20220115020

Abstract: Methods and systems for enabling an efficient review of meeting content via a metadata-enriched, speaker-attributed transcript are disclosed. By incorporating speaker diarization and other metadata, the system can provide a structured and effective way to review and/or edit the transcript. One type of metadata can be image or video data to represent the meeting content. Furthermore, the present subject matter utilizes a multimodal diarization model to identify and label different speakers. The system can synchronize various sources of data, e.g., audio channel data, voice feature vectors, acoustic beamforming, image identification, and extrinsic data, to implement speaker diarization.

Type: Application

Filed: October 11, 2021

Publication date: April 14, 2022

Applicant: SoundHound, Inc.

Inventors: Kiersten L. BRADLEY, Ethan COEYTAUX, Ziming YIN
METHOD AND SYSTEM FOR CONVERSATION TRANSCRIPTION WITH METADATA

Publication number: 20220115019

Abstract: Methods and systems for enabling an efficient review of meeting content via a metadata-enriched, speaker-attributed and multiuser-editable transcript are disclosed. By incorporating speaker diarization and other metadata, the system can provide a structured and effective way to review and/or edit the transcript by one or more editors. One type of metadata can be image or video data to represent the meeting content. Furthermore, the present subject matter utilizes a multimodal diarization model to identify and label different speakers. The system can synchronize various sources of data, e.g., audio channel data, voice feature vectors, acoustic beamforming, image identification, and extrinsic data, to implement speaker diarization.

Type: Application

Filed: October 11, 2021

Publication date: April 14, 2022

Applicant: SoundHound, Inc.

Inventors: Kiersten L. BRADLEY, Ethan COEYTAUX, Ziming YIN
VIDEO CONFERENCE CAPTIONING

Publication number: 20210074298

Abstract: Aspects include adding text captioning to a video conference. Multiple conferencing endpoints participate in a video conference. An endpoint locally captures an audio stream and transcribes human speech included in the audio stream into a caption stream. The endpoint can multiplex the caption stream with the audio stream and/or with a captured video stream into a transport stream. The endpoint sends the transport stream to the one or more other conferencing endpoints. To increase reliability and effectiveness, the conferencing endpoint can send the caption stream redundantly. A receiving endpoint can receive and demultiplex the transport stream. The receiving endpoint can coordinate output of the caption stream, the audio stream, and the video stream at corresponding output interfaces.

Type: Application

Filed: September 11, 2019

Publication date: March 11, 2021

Applicant: SoundHound, Inc.

Inventor: Ethan Coeytaux

VIDEO CONFERENCE CAPTIONING

METHOD AND SYSTEM FOR CONVERSATION TRANSCRIPTION WITH METADATA

METHOD AND SYSTEM FOR CONVERSATION TRANSCRIPTION WITH METADATA

VIDEO CONFERENCE CAPTIONING