Patent search ap:("Adobe Inc.") AND inv:"Xue Bai" Page 1

1.

发明授权
Interacting with hierarchical clusters of video segments using a metadata panel 有权

公开(公告)号：US11995894B2

公开(公告)日：2024-05-28

申请号：US17017353

申请日：2020-09-10

Applicant: ADOBE INC.

Inventor： Seth Walker , Joy Oakyung Kim , Hijung Shin , Aseem Agarwala , Joel R. Brandt , Jovan Popović , Lubomira Dontcheva , Dingzeyu Li , Xue Bai

IPC: G06V20/40 , G06T7/10

CPC classification number: G06V20/49 , G06T7/10 , G06V20/41 , G06T2207/10016

Abstract: Embodiments are directed to techniques for interacting with a hierarchical video segmentation using a metadata panel with a composite list of video metadata. The composite list is segmented into selectable metadata segments at locations corresponding to boundaries of video segments defined by a hierarchical segmentation. In some embodiments, the finest level of a hierarchical segmentation identifies the smallest interaction unit of a video—semantically defined video segments of unequal duration called clip atoms, and higher levels cluster the clip atoms into coarser sets of video segments. One or more metadata segments can be selected in various ways, such as by clicking or tapping on a metadata segment or by performing a metadata search. When a metadata segment is selected, a corresponding video segment is emphasized on the video timeline, a playback cursor is moved to the first video frame of the video segment, and the first video frame is presented.

2.

发明授权
Segmentation and hierarchical clustering of video 有权

公开(公告)号：US11450112B2

公开(公告)日：2022-09-20

申请号：US17017344

申请日：2020-09-10

Applicant: ADOBE INC.

Inventor： Hijung Shin , Xue Bai , Aseem Agarwala , Joel R. Brandt , Jovan Popović , Lubomira Dontcheva , Dingzeyu Li , Joy Oakyung Kim , Seth Walker

IPC: G11B27/00 , G11B27/19 , G06V20/40 , G06K9/62 , G10L25/78 , H04N7/00 , H04N21/00

Abstract: Embodiments are directed to segmentation and hierarchical clustering of video. In an example implementation, a video is ingested to generate a multi-level hierarchical segmentation of the video. In some embodiments, the finest level identifies a smallest interaction unit of the video—semantically defined video segments of unequal duration called clip atoms. Clip atom boundaries are detected in various ways. For example, speech boundaries are detected from audio of the video, and scene boundaries are detected from video frames of the video. The detected boundaries are used to define the clip atoms, which are hierarchically clustered to form a multi-level hierarchical representation of the video. In some cases, the hierarchical segmentation identifies a static, pre-computed, hierarchical set of video segments, where each level of the hierarchical segmentation identifies a complete set (i.e., covering the entire range of the video) of disjoint (i.e., non-overlapping) video segments with a corresponding level of granularity.

3.

发明申请
INTERACTING WITH HIERARCHICAL CLUSTERS OF VIDEO SEGMENTS USING A METADATA SEARCH 有权

公开(公告)号：US20220075820A1

公开(公告)日：2022-03-10

申请号：US17017370

申请日：2020-09-10

Applicant: ADOBE INC.

Inventor： Seth Walker , Joy Oakyung Kim , Morgan Nicole Evans , Najika Skyler Halsema Yoo , Aseem Agarwala , Joel R. Brandt , Jovan Popovic , Lubomira Dontcheva , Dingzeyu Li , Hijung Shin , Xue Bai

IPC: G06F16/738 , G06T13/80 , G06F3/0482 , G06F3/0484 , G06F16/74 , G06F16/735 , G06F16/75

Abstract: Embodiments are directed to techniques for interacting with a hierarchical video segmentation by performing a metadata search. Generally, various types of metadata can be extracted from a video, such as a transcript of audio, keywords from the transcript, content or action tags visually extracted from video frames, and log event tags extracted from an associated temporal log. The extracted metadata is segmented into metadata segments and associated with corresponding video segments defined by a hierarchical video segmentation. As such, a metadata search can be performed to identify matching metadata segments and corresponding matching video segments defined by a particular level of the hierarchical segmentation. Matching metadata segments are emphasized in a composite list of the extracted metadata, and matching video segments are emphasized on the video timeline. Navigating to a different level of the hierarchy transforms the search results into corresponding coarser or finer segments defined by the level.

4.

发明授权
Face-aware speaker diarization for transcripts and text-based video editing 有权

公开(公告)号：US12125501B2

公开(公告)日：2024-10-22

申请号：US17967399

申请日：2022-10-17

Applicant: Adobe Inc.

Inventor： Fabian David Caba Heilbron , Xue Bai , Aseem Omprakash Agarwala , Haoran Cai , Lubomira Assenova Dontcheva

IPC: G11B27/031 , G06V20/40

CPC classification number: G11B27/031 , G06V20/41

Abstract: Embodiments of the present invention provide systems, methods, and computer storage media for face-aware speaker diarization. In an example embodiment, an audio-only speaker diarization technique is applied to generate an audio-only speaker diarization of a video, an audio-visual speaker diarization technique is applied to generate a face-aware speaker diarization of the video, and the audio-only speaker diarization is refined using the face-aware speaker diarization to generate a hybrid speaker diarization that links detected faces to detected voices. In some embodiments, to accommodate videos with small faces that appear pixelated, a cropped image of any given face is extracted from each frame of the video, and the size of the cropped image is used to select a corresponding active speaker detection model to predict an active speaker score for the face in the cropped image.

5.

发明授权
Snap point video segmentation identifying selection snap points for a video 有权

公开(公告)号：US12033669B2

公开(公告)日：2024-07-09

申请号：US17330702

申请日：2021-05-26

Applicant: ADOBE INC.

Inventor： Seth Walker , Hijung Shin , Cristin Ailidh Fraser , Aseem Agarwala , Lubomira Dontcheva , Joel Richard Brandt , Jovan Popović , Joy Oakyung Kim , Justin Salamon , Jui-hsien Wang , Timothy Jeewun Ganter , Xue Bai , Dingzeyu Li

IPC: G11B27/00 , G06F3/0482 , G06F3/0486 , G11B27/02 , G11B27/036 , G11B27/10 , G11B27/031 , G11B27/36

CPC classification number: G11B27/036 , G06F3/0482 , G06F3/0486

Abstract: Embodiments are directed to a snap point segmentation that defines the locations of selection snap points for a selection of video segments. Candidate snap points are determined from boundaries of feature ranges of the video indicating when instances of detected features are present in the video. In some embodiments, candidate snap point separations are penalized for being separated by less than a minimum duration corresponding to a minimum pixel separation between consecutive snap points on a video timeline. The snap point segmentation is computed by solving a shortest path problem through a graph that models different snap point locations and separations. When a user clicks or taps on the video timeline and drags, a selection snaps to the snap points defined by the snap point segmentation. In some embodiments, the snap points are displayed during a drag operation and disappear when the drag operation is released.

6.

发明授权
Zoom and scroll bar for a video timeline 有权

公开(公告)号：US11899917B2

公开(公告)日：2024-02-13

申请号：US17969536

申请日：2022-10-19

Applicant: Adobe Inc.

Inventor： Seth Walker , Joy O Kim , Aseem Agarwala , Joel Richard Brandt , Jovan Popovic , Lubomira Dontcheva , Dingzeyu Li , Hijung Shin , Xue Bai

IPC: G06F3/04847 , G06F3/0485 , G06F3/04845

CPC classification number: G06F3/04847 , G06F3/0485 , G06F3/04845 , G06F2203/04806

Abstract: Embodiments are directed to techniques for interacting with a hierarchical video segmentation using a video timeline. In some embodiments, the finest level of a hierarchical segmentation identifies the smallest interaction unit of a video—semantically defined video segments of unequal duration called clip atoms, and higher levels cluster the clip atoms into coarser sets of video segments. A presented video timeline is segmented based on one of the levels, and one or more segments are selected through interactions with the video timeline. For example, a click or tap on a video segment or a drag operation dragging along the timeline snaps selection boundaries to corresponding segment boundaries defined by the level. Navigating to a different level of the hierarchy transforms the selection into coarser or finer video segments defined by the level. Any operation can be performed on selected video segments, including playing back, trimming, or editing.

7.

发明授权
Interacting with semantic video segments through interactive tiles 有权

公开(公告)号：US11887629B2

公开(公告)日：2024-01-30

申请号：US17330689

申请日：2021-05-26

Applicant: ADOBE INC.

Inventor： Seth Walker , Hijung Shin , Cristin Ailidh Fraser , Aseem Agarwala , Lubomira Dontcheva , Joel Richard Brandt , Jovan Popović , Joy Oakyung Kim , Justin Salamon , Jui-hsien Wang , Timothy Jeewun Ganter , Xue Bai , Dingzeyu Li

IPC: G11B27/00 , G11B27/036 , G06F3/0486 , G06F3/0482 , G11B27/02

CPC classification number: G11B27/036 , G06F3/0482 , G06F3/0486

Abstract: Embodiments are directed to interactive tiles that represent video segments of a segmentation of a video. In some embodiments, each interactive tile represents a different video segment from a particular video segmentation (e.g., a default video segmentation). Each interactive tile includes a thumbnail (e.g., the first frame of the video segment represented by the tile), some transcript from the beginning of the video segment, a visualization of detected faces in the video segment, and one or more faceted timelines that visualize a category of detected features (e.g., a visualization of detected visual scenes, audio classifications, visual artifacts). In some embodiments, interacting with a particular interactive tile navigates to a corresponding portion of the video, adds a corresponding video segment to a selection, and/or scrubs through tile thumbnails.

8.

发明申请
ZOOM AND SCROLL BAR FOR A VIDEO TIMELINE 有权

公开(公告)号：US20230043769A1

公开(公告)日：2023-02-09

申请号：US17969536

申请日：2022-10-19

Applicant: Adobe Inc.

Inventor： Seth WALKER , Joy O KIM , Aseem AGARWALA , Joel Richard Brandt , Jovan POPOVIC , Lubomira DONTCHEVA , Dingzeyu LI , Hijung SHIN , Xue Bai

IPC: G06F3/04847 , G06F3/0485 , G06F3/04845

Abstract: Embodiments are directed to techniques for interacting with a hierarchical video segmentation using a video timeline. In some embodiments, the finest level of a hierarchical segmentation identifies the smallest interaction unit of a video—semantically defined video segments of unequal duration called clip atoms, and higher levels cluster the clip atoms into coarser sets of video segments. A presented video timeline is segmented based on one of the levels, and one or more segments are selected through interactions with the video timeline. For example, a click or tap on a video segment or a drag operation dragging along the timeline snaps selection boundaries to corresponding segment boundaries defined by the level. Navigating to a different level of the hierarchy transforms the selection into coarser or finer video segments defined by the level. Any operation can be performed on selected video segments, including playing back, trimming, or editing.

9.

发明申请
INTERACTING WITH HIERARCHICAL CLUSTERS OF VIDEO SEGMENTS USING A METADATA PANEL 有权

公开(公告)号：US20220076024A1

公开(公告)日：2022-03-10

申请号：US17017353

申请日：2020-09-10

Applicant: ADOBE INC.

Inventor： Seth Walker , Joy Oakyung Kim , Hijung Shin , Aseem Agarwala , Joel R. Brandt , Jovan Popovic , Lubomira Dontcheva , Dingzeyu Li , Xue Bai

IPC: G06K9/00 , G06T7/10

Abstract: Embodiments are directed to techniques for interacting with a hierarchical video segmentation using a metadata panel with a composite list of video metadata. The composite list is segmented into selectable metadata segments at locations corresponding to boundaries of video segments defined by a hierarchical segmentation. In some embodiments, the finest level of a hierarchical segmentation identifies the smallest interaction unit of a video—semantically defined video segments of unequal duration called clip atoms, and higher levels cluster the clip atoms into coarser sets of video segments. One or more metadata segments can be selected in various ways, such as by clicking or tapping on a metadata segment or by performing a metadata search. When a metadata segment is selected, a corresponding video segment is emphasized on the video timeline, a playback cursor is moved to the first video frame of the video segment, and the first video frame is presented.

10.

发明申请
INTERACTING WITH HIERARCHICAL CLUSTERS OF VIDEO SEGMENTS USING A VIDEO TIMELINE 有权

公开(公告)号：US20220075513A1

公开(公告)日：2022-03-10

申请号：US17017366

申请日：2020-09-10

Applicant: ADOBE INC.

Inventor： Seth Walker , Joy Oakyung Kim , Aseem Agarwala , Joel R. Brandt , Jovan Popovic , Lubomira Dontcheva , Dingzeyu Li , Hijung Shin , Xue Bai

IPC: G06F3/0484 , G06F3/0485

Abstract: Embodiments are directed to techniques for interacting with a hierarchical video segmentation using a video timeline. In some embodiments, the finest level of a hierarchical segmentation identifies the smallest interaction unit of a video—semantically defined video segments of unequal duration called clip atoms, and higher levels cluster the clip atoms into coarser sets of video segments. A presented video timeline is segmented based on one of the levels, and one or more segments are selected through interactions with the video timeline. For example, a click or tap on a video segment or a drag operation dragging along the timeline snaps selection boundaries to corresponding segment boundaries defined by the level. Navigating to a different level of the hierarchy transforms the selection into coarser or finer video segments defined by the level. Any operation can be performed on selected video segments, including playing back, trimming, or editing.

Search Results

Country/Region

Patent validity

Application date

Publication (announcement) day

applicant

The country/region where the applicant is located

Inventor

IPC

IPC Department

IPC class

IPC subclass

IPC group

IPC team

Appearance classification