-
公开(公告)号:US20230169795A1
公开(公告)日:2023-06-01
申请号:US17539652
申请日:2021-12-01
Applicant: ADOBE INC.
Inventor: JIYOUNG LEE , Justin Jonathan SALAMOM , Dingzeyu LI
CPC classification number: G06V40/20 , G06V10/82 , G06V20/41 , G06V20/46 , G06V20/49 , G06N3/08 , G06N3/0454
Abstract: A method for detecting a cue (e.g., a visual cue or a visual cue combined with an audible cue) occurring together in an input video includes: presenting a user interface to record an example video of a user performing an act including the cue; determining a part of the example video where the cue occurs; applying a feature of the part to a neural network to generate a positive embedding; dividing the input video into a plurality of chunks and applying a feature of each chunk to the neural network to generate a plurality of negative embeddings; applying a feature of a given one of the chunks to the neural network to output a query embedding; and determining whether the cue occurs in the input video from the query embedding, the positive embedding, and the negative embeddings.