Vehicle Gaze Detection Using Virtual Camera Space Training

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional approaches for training machine learning models for object identification require a significant amount of labeled training data, which is time-consuming and costly to create, especially in dynamic environments like vehicle interiors where camera positions can vary.

Innovation Solution

The use of a virtual camera space and a coordinate propagation mechanism to bridge the physical world with image data captured from arbitrary camera positions, allowing for the generation of ground truth data and training of neural networks across various vehicle configurations.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If conventional supervised training approaches are used to train machine learning models for object identification, then the model can achieve accurate object detection, but a significant amount of labeled training data is required which is time-consuming and costly to create

Engineering Contradiction:
Improveobject detection accuracyVSAvoidtraining data creation time
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The patent uses image synthesis networks to generate synthetic training images that copy the visual characteristics and patterns of real images. These synthesized images serve as training data substitutes, eliminating the need to manually create large amounts of labeled real images while maintaining the statistical properties needed for accurate model training

Inventive Principle:
Principle #26Copying

Solution Approach 2:

The system performs self-service by automatically generating its own training data through the image synthesis network. Instead of requiring external manual annotation efforts, the model can be trained using synthetically generated images that are created algorithmically, making the training process self-sufficient and eliminating the bottleneck of manual data preparation

Inventive Principle:
Principle #25Self-service

2Measurement precision

If conventional supervised training approaches are used to train machine learning models, then the model can achieve accurate object detection, but the process becomes too expensive for various uses

Engineering Contradiction:
Improveobject detection accuracyVSAvoidtraining cost
Core Design Contradiction:
Measurement precisionVSEase of manufacture

Solution Approach 1:

The patent uses image synthesis networks to generate synthetic training images that copy the visual characteristics and patterns of real images. These synthesized images serve as training data substitutes, eliminating the need to manually create large amounts of labeled real images while maintaining the statistical properties needed for accurate model training

Inventive Principle:
Principle #26Copying

Solution Approach 2:

The system replaces expensive, time-intensive manual data annotation processes with inexpensive automated synthetic image generation. The synthesized images act as disposable training data that can be generated on-demand without the recurring costs of human annotators, making the training process economically viable

Inventive Principle:
Principle #27Cheap short-living objects (Disposable)

3Measurement precision

If conventional supervised training approaches are used to train machine learning models, then the model can achieve accurate object detection, but an insufficient amount of training data may result

Engineering Contradiction:
Improveobject detection accuracyVSAvoidtraining data quantity
Core Design Contradiction:
Measurement precisionVSQuantity of substance

Solution Approach 1:

The patent uses image synthesis networks to generate synthetic training images that copy the visual characteristics and patterns of real images. These synthesized images serve as training data substitutes, eliminating the need to manually create large amounts of labeled real images while maintaining the statistical properties needed for accurate model training

Inventive Principle:
Principle #26Copying

Solution Approach 2:

The system performs preliminary action by pre-generating large quantities of synthetic training images before the actual model training begins. This advance preparation ensures that sufficient training data is available, eliminating the constraint of limited real labeled images and enabling comprehensive model training

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS12236351B2Gaze detection using one or more neural networks
Publication Date: 2025.02.25 NVIDIA CORP
  • US12236351B2 patent drawing
  • US12236351B2 patent drawing
  • US12236351B2 patent drawing

AI summary

Apparatuses, systems, and techniques are described to determine locations of objects using images including digital representations of those objects. In at least one embodiment, a gaze of one or more occupants of a vehicle is determined independently of a location of one or more sensors used to detect those occupants.