Patent search ap:("Beijing Baidu Netcom Science AND Technology Co. Page Ltd.") AND inv:"Zhizhen Chi"

1.

发明授权
Method, device, apparatus for predicting video coding complexity and storage medium 有权

公开(公告)号：US11259029B2

公开(公告)日：2022-02-22

申请号：US16797911

申请日：2020-02-21

Applicant: Beijing Baidu Netcom Science and Technology Co., Ltd.

Inventor： Zhichao Zhou , Dongliang He , Fu Li , Xiang Zhao , Xin Li , Zhizhen Chi , Xiang Long , Hao Sun

IPC: H04N19/14 , H04N19/12 , H04N19/625 , G06N3/08

Abstract: A method, device, apparatus for predicting a video coding complexity and a computer-readable storage medium are provided. The method includes: acquiring an attribute feature of a target video; extracting a plurality of first target image frames from the target video; performing a frame difference calculation on the plurality of the first target image frames, to acquire a plurality of first frame difference images; determining a histogram feature for frame difference images of the target video according to a statistical histogram of each first frame difference image; and inputting a plurality of features of the target video into a coding complexity prediction model to acquire a coding complexity prediction value of the target video. Through the above method, the BPP prediction value can be acquired intelligently.

2.

发明申请
METHOD AND APPARATUS FOR CLASSIFYING VIDEO 有权

公开(公告)号：US20210019531A1

公开(公告)日：2021-01-21

申请号：US16830895

申请日：2020-03-26

Applicant: Beijing Baidu Netcom Science and Technology Co., Ltd.

Inventor： Xiang Long , Dongliang He , Fu Li , Zhizhen Chi , Zhichao Zhou , Xiang Zhao , Ping Wang , Hao Sun , Shilei Wen , Errui Ding

IPC: G06K9/00 , G06N3/08 , G06K9/62

Abstract: a method and an apparatus for classifying a video are provided. The method may include: acquiring a to-be-classified video; extracting a set of multimodal features of the to-be-classified video; inputting the set of multimodal features into a post-fusion model corresponding to each modal respectively, to obtain multimodal category information of the to-be-classified video; and fusing the multimodal category information of the to-be-classified video, to obtain category information of the to-be-classified video. This embodiment improves the accuracy of video classification.

3.

发明授权
Method and apparatus for classifying video 有权

公开(公告)号：US11256920B2

公开(公告)日：2022-02-22

申请号：US16830895

申请日：2020-03-26

Applicant: Beijing Baidu Netcom Science and Technology Co., Ltd.

Inventor： Xiang Long , Dongliang He , Fu Li , Zhizhen Chi , Zhichao Zhou , Xiang Zhao , Ping Wang , Hao Sun , Shilei Wen , Errui Ding

IPC: G06K9/00 , G06K9/62 , G06N3/08

Abstract: A method and an apparatus for classifying a video are provided. The method may include: acquiring a to-be-classified video; extracting a set of multimodal features of the to-be-classified video; inputting the set of multimodal features into a post-fusion model corresponding to each modal respectively, to obtain multimodal category information of the to-be-classified video; and fusing the multimodal category information of the to-be-classified video, to obtain category information of the to-be-classified video. This embodiment improves the accuracy of video classification.

4.

发明申请
METHOD, DEVICE, APPARATUS FOR PREDICTING VIDEO CODING COMPLEXITY AND STORAGE MEDIUM 审中-公开

公开(公告)号：US20200374526A1

公开(公告)日：2020-11-26

申请号：US16797911

申请日：2020-02-21

Applicant: Beijing Baidu Netcom Science and Technology Co., Ltd

Inventor： Zhichao Zhou , Dongliang He , Fu Li , Xiang Zhao , Xin Li , Zhizhen Chi , Xiang Long , Hao Sun

IPC: H04N19/14 , H04N19/12 , H04N19/625 , G06N3/08

Abstract: A method, device, apparatus for predicting a video coding complexity and a computer-readable storage medium are provided. The method includes: acquiring an attribute feature of a target video; extracting a plurality of first target image frames from the target video; performing a frame difference calculation on the plurality of the first target image frames, to acquire a plurality of first frame difference images; determining a histogram feature for frame difference images of the target video according to a statistical histogram of each first frame difference image; and inputting a plurality of features of the target video into a coding complexity prediction model to acquire a coding complexity prediction value of the target video. Through the above method, the BPP prediction value can be acquired intelligently.

Patent Agency Ranking