Augmenting Video Streams with Geospatial Context
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing video capturing technologies are limited to displaying only what the operator sees, lacking additional information about points of interest or geographical context, which restricts the depth and relevance of the visual content.
Innovation Solution
A mobile computing device equipped with GPS and video processing capabilities captures video streams, extracts visual and geographical information, and transmits it to a server for augmentation, overlaying relevant data such as object recognition, user profiles, and location-based information back onto the video stream.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If video capturing devices are used to capture video streams, then the operator can record visual content, but the content is limited to only what is visible to the operator without additional contextual information
Solution Approach 1:
The patent combines multiple data sources including visual information from video streams, geographical data from GPS, and contextual information from databases to create an augmented video stream. This merging of data types enriches the visual content with additional context while managing system complexity through integrated processing architecture.
Solution Approach 2:
The patent introduces a server computer as an intermediary that receives visual and geographical information from mobile devices, processes this data against stored reference data, and returns augmented video streams. This intermediary approach allows complex processing to be performed remotely, reducing the complexity burden on the mobile device itself.
2Loss of information
If geographical information and visual information are combined to augment video streams, then contextual relevance is improved, but processing time and system complexity increase
Solution Approach 1:
The patent performs preliminary actions by pre-storing reference data in databases before the actual video capture occurs. This allows the system to quickly match and retrieve relevant contextual information during video processing without performing complex computations in real-time, thereby reducing processing time while maintaining rich contextual information.
Solution Approach 2:
The patent extracts only the essential visual features and geographical data from the video streams and transmits these condensed representations to the server for processing. This extraction approach reduces the amount of data that needs to be processed and transmitted, significantly decreasing processing time while preserving the necessary information for accurate augmentation.
Data Source
AI summary
According to an example, a computing device includes a memory on which is stored machine readable instructions that may cause a processor to access a video stream, generate geographical information associated with frames of the video stream, extract features of points of interest in the frames, transmit the extracted features of the points of interest and the geographical information to a server computer that is to use the extracted features of the points of interest and the geographical information to identify augment information, receive the augment information from the server computer, and augment the video stream with the augment information to generate an augmented video stream.


