AUTOMATIC RECOGNITION OF VISUAL AND AUDIO-VISUAL CUES

    公开(公告)号:US20230169795A1

    公开(公告)日:2023-06-01

    申请号:US17539652

    申请日:2021-12-01

    Applicant: ADOBE INC.

    Abstract: A method for detecting a cue (e.g., a visual cue or a visual cue combined with an audible cue) occurring together in an input video includes: presenting a user interface to record an example video of a user performing an act including the cue; determining a part of the example video where the cue occurs; applying a feature of the part to a neural network to generate a positive embedding; dividing the input video into a plurality of chunks and applying a feature of each chunk to the neural network to generate a plurality of negative embeddings; applying a feature of a given one of the chunks to the neural network to output a query embedding; and determining whether the cue occurs in the input video from the query embedding, the positive embedding, and the negative embeddings.

Patent Agency Ranking