Agent Device Loudness-Based Display Control

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing agent technologies in vehicles lack user-friendliness, as they do not effectively adapt to varying vocal input loudness, leading to inconsistent and sometimes intrusive agent interactions.

Innovation Solution

An agent device and system that utilize a display controller and controller to dynamically adjust the display of an agent image based on vocal input loudness, allowing for private or public realization of the agent service, ensuring user intention is reflected and enhancing user experience.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If the agent image is always displayed on the first display, then the agent service is always visible to all users, but this may be intrusive when the user intends private interaction

Engineering Contradiction:
Improveuser-friendlinessVSAvoidintrusive interaction
Core Design Contradiction:
Ease of operationVSObject-affected harmful factors

Solution Approach 1:

The patent applies dynamics by making the agent image display adaptive rather than static. The controller dynamically switches between displaying the agent image on the first display (public mode) and the second display (private mode) based on the detected loudness of the user's voice. When soft speech is detected, the system transitions to private mode, displaying the agent image only on the user's terminal device, thereby reducing intrusiveness while maintaining service availability.

Inventive Principle:
Principle #15Dynamics

2Adaptability or versatility

If the agent image is displayed on the second display when voice loudness is low, then private interaction is supported, but the agent image may not be visible to other users who need to see it

Engineering Contradiction:
Improveprivate interaction supportVSAvoidagent image visibility
Core Design Contradiction:
Adaptability or versatilityVSLoss of information

Solution Approach 1:

The patent applies segmentation by dividing the display function across multiple devices. The first display (in-vehicle display) serves public visibility needs, while the second display (terminal device display) serves private interaction needs. The controller segments the agent image display based on voice loudness detection, routing it to appropriate displays, thus supporting both private interaction and public visibility without information loss.

Inventive Principle:
Principle #1Segmentation

3Measurement precision

If the system detects voice loudness to control display, then user intention is accurately reflected, but the system complexity increases

Engineering Contradiction:
Improvevoice loudness detection accuracyVSAvoidcontrol system complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent applies feedback by implementing a closed-loop control system. The voice recognition device continuously monitors the user's voice loudness, and the controller uses this feedback to automatically adjust the agent image display mode. This feedback mechanism enables accurate detection of user intention (soft speech indicating private interaction desire) while managing system complexity through automated control logic rather than manual intervention.

Inventive Principle:
Principle #23Feedback

Data Source

PatentUS11518399B2Agent device, agent system, method for controlling agent device, and storage medium
Publication Date: 2022.12.06 HONDA MOTOR CO LTD
  • US11518399B2 patent drawing
  • US11518399B2 patent drawing
  • US11518399B2 patent drawing

AI summary

An agent device includes a display controller configured to cause a first display to display an agent image when an agent providing a service including causing an output device to output response of voice in response to an utterance of a user is activated, and a controller configured to execute particular control for causing a second display to display the agent image according to loudness of a voice received by an external terminal receiving a vocal input.