Video Metadata Extraction for Retail Foot Traffic Analysis

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current video surveillance systems and retail merchandising methods lack efficient means to quickly search and analyze data on object movements within a specified area, and there is a need for quantitative measurement of floor plans and merchandising effectiveness.

Innovation Solution

A system and method for capturing and storing video metadata that tracks the movements of objects, particularly people, within a Region Of Interest (ROI), allowing for real-time analysis and generation of reports on dwell time, directional analysis, and shopper traffic patterns, which can be displayed graphically for informed decision-making.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Loss of time

If video data is stored and searched using traditional methods (VCR tapes or basic digital recorders), then the system can record and store video footage, but the searching process becomes extremely time-consuming and labor-intensive requiring personnel to review tapes manually

Engineering Contradiction:
Improvesearching timeVSAvoidsearch efficiency
Core Design Contradiction:
Loss of timeVSProductivity

Solution Approach 1:

The patent extracts and separates the analysis function from the storage function by generating video metadata that contains object movement information independent of the raw video data. This extracted metadata can be searched and analyzed without reviewing the actual video footage, dramatically reducing search time and improving productivity.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The patent introduces video metadata as an intermediary layer between the raw video data and the search/analysis process. This metadata acts as a mediator that contains essential information about object movements, allowing users to search and analyze video content without manually reviewing the actual video footage.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Measurement precision

If raw video data is processed for analysis, then complete information about object movements is obtained, but the processing time and computational resources required are significantly high

Engineering Contradiction:
Improvemovement tracking accuracyVSAvoidprocessing time
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The patent performs preliminary action by generating video metadata during or immediately after video recording, extracting object movement information in advance. This pre-processing allows subsequent analysis to work with the already-extracted metadata rather than processing raw video data, reducing processing time while maintaining measurement precision.

Inventive Principle:
Principle #10Preliminary action

3Loss of information

If detailed tracking of all objects in video footage is performed, then comprehensive movement data is captured, but the storage requirements and data transmission bandwidth become excessively large

Engineering Contradiction:
Improvemovement data completenessVSAvoiddata storage volume
Core Design Contradiction:
Loss of informationVSQuantity of substance

Solution Approach 1:

The patent extracts only the essential movement information from the video data and stores it as compact metadata. This extracted metadata contains object tracking data, location information, and movement patterns in a condensed format that requires significantly less storage space while maintaining data completeness for analysis purposes.

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS8964036B2System and method for capturing, storing, analyzing and displaying data related to the movements of objects
Publication Date: 2015.02.24 COGNYTE TECH ISRAEL LTD
  • US8964036B2 patent drawing
  • US8964036B2 patent drawing
  • US8964036B2 patent drawing

AI summary

A system and method for the capture and storage of data relating to the movements of objects, in a specified area and enables this data to be displayed in a graphically meaningful and useful manner. Video data is collected and video metadata is generated relating to objects (persons) appearing in the video data and their movements over time. The movements of the objects are then analyzed to detect the movements within a region of interest. This detection of movement allows a user, such as a manager of a store, to make informed decisions as to the infrastructure and operation of the store. One detection method relates to the number of people that are present in a region of interest for a specified time period. A second detection method relates to the number of people that remain or dwell in a particular area for a particular time period. A third detection method determines the flow of people and the direction they take within a region of interest. A fourth detection method relates to the number of people that enter a certain area by crossing a virtual line, a tripwire.