Video Analytics Pipeline Framework for Custom Model Training

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current video analytics systems in the surveillance industry face challenges due to data and scene variability, limited model applicability, and a disconnect between researchers and end-users, leading to high false alarms and limited performance in real-world scenarios.

Innovation Solution

A pipeline framework that allows users to annotate datasets, train custom computer vision algorithms, and perform analytics through various modules for preprocessing, pattern recognition, and statistical analysis, enabling scalable and flexible video analysis on multiple streams.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If vision algorithms are designed and optimized on datasets to encapsulate real world scenarios, then the algorithms can achieve good performance on training data, but the performance is unknown in new scenarios leading to higher false alarms

Engineering Contradiction:
Improvealgorithm performanceVSAvoidfalse alarm rate
Core Design Contradiction:
Measurement precisionVSReliability

Solution Approach 1:

The system dynamically adapts vision algorithms by allowing users to adjust parameters and retrain models based on specific scene conditions. The framework enables dynamic modification of algorithm behavior to match varying real-world scenarios, transitioning from static pre-trained models to adaptive systems that can handle data and scene variability.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The framework allows users to change algorithm parameters and retrain models with custom datasets specific to their scenarios. By modifying training data and algorithm parameters, the system optimizes performance for specific applications while reducing false alarms in new scenarios.

Inventive Principle:
Principle #35Parameter changes

2Ease of manufacture

If a black boxed analytic based on one method is used, then the system is simple to deploy, but the applicability is limited in other scenarios

Engineering Contradiction:
Improvedeployment simplicityVSAvoidscenario applicability
Core Design Contradiction:
Ease of manufactureVSAdaptability or versatility

Solution Approach 1:

The framework provides a universal platform that can accommodate multiple vision algorithms and methods within a single system. Users can select and switch between different algorithms (e.g., density-based crowd counting, detection-based crowd counting) based on scenario requirements, making the system versatile across various applications while maintaining ease of deployment through a unified interface.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The system transitions from a static black-box approach to a dynamic configurable framework where users can adjust algorithm selection, parameter settings, and model configurations based on specific scenario needs, enabling the same system to adapt to diverse applications.

Inventive Principle:
Principle #15Dynamics

3Adaptability or versatility

If users are given the power to build, customize and perform analytics, then the system becomes more adaptable to specific tasks, but the device complexity increases

Engineering Contradiction:
Improvecustomization capabilityVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The framework enables users to independently annotate datasets, train custom models, and configure analytics pipelines without requiring expert knowledge from researchers or software developers. The system provides self-service capabilities through intuitive interfaces that guide users through the customization process, reducing the complexity burden despite increased adaptability.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The framework acts as an intermediary layer between raw vision algorithms and end-users, providing abstraction through modular pipelines and configuration interfaces. This mediator simplifies the interaction complexity while preserving full customization capability, allowing users to leverage powerful algorithms without directly managing their complexity.

Inventive Principle:
Principle #24Intermediary (Mediator)

4Measurement precision

If data driven algorithms are trained on annotated datasets to accomplish specific tasks, then the algorithms achieve task-specific performance, but retraining requires specific data that may not be available to users

Engineering Contradiction:
Improvetask-specific performanceVSAvoidretraining accessibility
Core Design Contradiction:
Measurement precisionVSEase of operation

Solution Approach 1:

The framework empowers users to independently annotate their own datasets and retrain models for their specific tasks without relying on pre-packaged datasets created by researchers. Users can collect, annotate, and use their own data to train algorithms, making the retraining process accessible and eliminating dependency on external data availability.

Inventive Principle:
Principle #25Self-service

Data Source

PatentUS11532158B2Methods and systems for customized image and video analysis
Publication Date: 2022.12.20 UNIV HOUSTON SYST
  • US11532158B2 patent drawing
  • US11532158B2 patent drawing
  • US11532158B2 patent drawing

AI summary

Preferred embodiments described herein relate to a pipeline framework that allows for customized analytic processes to be performed on multiple streams of videos. An analytic takes data as input and performs a set of operations and transforms it into information. The methods and systems disclosed herein include a framework (1) that allows users to annotate and create variable datasets, (2) to train computer vision algorithms to create custom models to accomplish specific tasks, (3) to pipeline video data through various computer vision modules for preprocessing, pattern recognition, and statistical analytics to create custom analytics, and (4) to perform analysis using a scalable architecture that allows for running analytic pipelines on multiple streams of videos.