Video copy detection method and system based on soft cascade model sensitive to deformation
A technology of video copy detection and soft cascading, applied in the field of computer network, to achieve the effect of improving copy detection speed, reducing space-time cost, and simple model
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Publication Date
- 2015-07-01
- Estimated Expiration
- Not applicable · inactive patent
Smart Images
Figure 1 Figure 2 Figure 3
Abstract
Description
Technical field
[0001] The present invention provides a video copy detection method and system based on a deformation-sensitive soft cascade model, which can accurately and quickly identify whether a query video is a copy of a given reference video library, in digital rights management, advertising tracking, video content There are important applications in fields such as filtration. The invention belongs to the field of computer network technology. Background technique
[0002] With economic and cultural development and technological progress, the global film and television industry has been growing steadily in recent years. In 2011 alone, my country's movie box office exceeded 13.1 billion yuan, an increase of 28.93% over 2010, and the global movie box office hit a new high of 32.6 billion US dollars. The film and television industry has become one of the pillar industries in many countries. For example, the film and television industry in the United States created an output ...
Examples
Embodiment Construction
[0060] The present invention will be described in detail below in conjunction with the embodiments and drawings.
[0061] A video copy detection method based on a deformation-sensitive soft cascade model. For the overall process, see image 3 . Among them, the preprocessing operation includes the following steps:
[0062] Step 11: Extract visual key frames; the present invention extracts visual key frames at equal intervals at a frequency of 3 frames per second. The sampling rate of 3 frames per second can discard most of the video frames while maintaining the main visual content of the video, saving the time and space cost of visual frame retrieval.
[0063] Step 12: Extract audio frames; to do this, first divide the audio track of the video into 90 milliseconds of audio words, with a 60 millisecond overlap between adjacent audio words, and then, 198 consecutive audio words form a 6-second long Audio frames, adjacent audio frames share 178 audio words, that is, there is an overlap...