Augmenting Video Streams with Geospatial Context

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing video capturing technologies are limited to displaying only what the operator sees, lacking additional information about points of interest or geographical context, which restricts the depth and relevance of the visual content.

Innovation Solution

A mobile computing device equipped with GPS and video processing capabilities captures video streams, extracts visual and geographical information, and transmits it to a server for augmentation, overlaying relevant data such as object recognition, user profiles, and location-based information back onto the video stream.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Loss of information

If video capturing devices are used to capture video streams, then the operator can record visual content, but the content is limited to only what is visible to the operator without additional contextual information

Engineering Contradiction:
Improvevisual informationVSAvoidsystem complexity
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The patent combines multiple data sources including visual information from video streams, geographical data from GPS, and contextual information from databases to create an augmented video stream. This merging of data types enriches the visual content with additional context while managing system complexity through integrated processing architecture.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The patent introduces a server computer as an intermediary that receives visual and geographical information from mobile devices, processes this data against stored reference data, and returns augmented video streams. This intermediary approach allows complex processing to be performed remotely, reducing the complexity burden on the mobile device itself.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Loss of information

If geographical information and visual information are combined to augment video streams, then contextual relevance is improved, but processing time and system complexity increase

Engineering Contradiction:
Improvecontextual informationVSAvoidprocessing time
Core Design Contradiction:
Loss of informationVSLoss of time

Solution Approach 1:

The patent performs preliminary actions by pre-storing reference data in databases before the actual video capture occurs. This allows the system to quickly match and retrieve relevant contextual information during video processing without performing complex computations in real-time, thereby reducing processing time while maintaining rich contextual information.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent extracts only the essential visual features and geographical data from the video streams and transmits these condensed representations to the server for processing. This extraction approach reduces the amount of data that needs to be processed and transmitted, significantly decreasing processing time while preserving the necessary information for accurate augmentation.

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS8953054B2System to augment a visual data stream based on a combination of geographical and visual information
Publication Date: 2015.02.10 HEWLETT PACKARD DEVELOPMENT COMPANY LP
  • US8953054B2 patent drawing
  • US8953054B2 patent drawing
  • US8953054B2 patent drawing

AI summary

According to an example, a computing device includes a memory on which is stored machine readable instructions that may cause a processor to access a video stream, generate geographical information associated with frames of the video stream, extract features of points of interest in the frames, transmit the extracted features of the points of interest and the geographical information to a server computer that is to use the extracted features of the points of interest and the geographical information to identify augment information, receive the augment information from the server computer, and augment the video stream with the augment information to generate an augmented video stream.