AR Virtual Assistant for Context-Aware Task Guidance

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current augmented reality systems lack the capability to effectively mentor users in completing complex physical tasks by providing real-time, context-aware guidance that integrates visual and auditory feedback seamlessly, limiting their ability to perform tasks accurately and efficiently.

Innovation Solution

An augmented reality virtual assistant system that utilizes a combination of computer vision, natural language processing, and machine learning to analyze user actions and provide interactive guidance through a head-mounted display, correlating video and audio inputs to offer step-by-step instructions and feedback, enhancing user understanding and task completion.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Manufacturing precision

If augmented reality systems provide real-time visual feedback through head-mounted displays, then user guidance and task completion accuracy are improved, but system complexity and computational requirements increase

Engineering Contradiction:
Improvetask completion accuracyVSAvoidsystem complexity
Core Design Contradiction:
Manufacturing precisionVSDevice complexity

Solution Approach 1:

The system segments complex tasks into discrete steps with specific visual markers, breaking down the overall task complexity into manageable segments that can be processed and guided individually through the AR interface

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system performs preliminary actions by pre-processing video feeds, extracting visual markers, and preparing guidance content before it is needed, reducing real-time computational complexity while maintaining accurate task guidance

Inventive Principle:
Principle #10Preliminary action

2Adaptability or versatility

If the system integrates multiple sensors and processing modules for comprehensive scene understanding, then context-aware guidance is improved, but device complexity and power consumption increase

Engineering Contradiction:
Improvecontext-aware guidanceVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The system merges video processing, audio processing, and sensor data into a unified context understanding framework, integrating multiple information sources to achieve comprehensive scene understanding while managing system complexity through unified architecture

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The system implements multi-functional processing modules that can handle multiple types of data (video, audio, sensor) and perform multiple functions (marker detection, speech recognition, context analysis) using shared computational resources and algorithms

Inventive Principle:
Principle #6Universality (Multi-functionality)

3Loss of information

If the system provides detailed step-by-step instructions and visual overlays, then user understanding and task accuracy are improved, but information processing load and response time increase

Engineering Contradiction:
Improveuser understandingVSAvoidresponse time
Core Design Contradiction:
Loss of informationVSLoss of time

Solution Approach 1:

The system applies local quality by providing detailed visual overlays and instructions only at specific locations and moments in the task sequence, rather than uniformly throughout, reducing overall information processing load while maintaining user understanding where needed

Inventive Principle:
Principle #3Local quality

Solution Approach 2:

The system uses partial action by selectively providing guidance information based on the user's current task state and needs, offering only the necessary subset of available information rather than complete task documentation, reducing response time while maintaining adequacy

Inventive Principle:
Principle #16Partial or excessive action

Data Source

PatentUS10824310B2Augmented reality virtual personal assistant for external representation
Publication Date: 2020.11.03 SRI INTERNATIONAL
  • US10824310B2 patent drawing
  • US10824310B2 patent drawing
  • US10824310B2 patent drawing

AI summary

A computing system for virtual personal assistance includes technologies to, among other things, correlate an external representation of an object with a real world view of the object, display virtual elements on the external representation of the object and/or display virtual elements on the real world view of the object, to provide virtual personal assistance in a multi-step activity or another activity that involves the observation or handling of an object and a reference document.