Patents by Inventor Aarav Gupta

Aarav Gupta has filed for patents to protect the following inventions. This listing includes patent applications that are pending as well as patents that have already been granted by the United States Patent and Trademark Office (USPTO).

  • Patent number: 12670688
    Abstract: Systems, methods, and media for navigating video content are disclosed. Systems, methods, and media can receive at least a first video content item; select a first subset of video frames; identify, using a first computer vision model, a plurality of sets of visual features; generate, using a first language model, a plurality of sets of caption data for the first subset of video frames; generate, using a first speech recognition model, recognized speech data based at least on the audio data of the first video content item; receive first user query data; generate, using a second language model, at least a first textual response to the first user query data; determine, using a third language model, first relevant video frames of the first subset of video frames which are associated with respective time positions; and cause one or more selectable links to the respective time positions to be presented.
    Type: Grant
    Filed: February 23, 2024
    Date of Patent: June 30, 2026
    Inventors: Anurag Gupta, Premith Kumar Chilukuri, Aarav Gupta, Ansh Gupta
  • Publication number: 20250272948
    Abstract: Systems, methods, and media for navigating video content are disclosed. Systems, methods, and media can receive at least a first video content item; select a first subset of video frames; identify, using a first computer vision model, a plurality of sets of visual features; generate, using a first language model, a plurality of sets of caption data for the first subset of video frames; generate, using a first speech recognition model, recognized speech data based at least on the audio data of the first video content item; receive first user query data; generate, using a second language model, at least a first textual response to the first user query data; determine, using a third language model, first relevant video frames of the first subset of video frames which are associated with respective time positions; and cause one or more selectable links to the respective time positions to be presented.
    Type: Application
    Filed: February 23, 2024
    Publication date: August 28, 2025
    Inventors: Anurag Gupta, Premith Kumar Chilukuri, Aarav Gupta, Ansh Gupta