Automated Character Control via Audio and Visual Input
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing systems for controlling human inhabited characters, such as avatars and animations, require significant time and effort for users to learn complex input commands, which can distract from real-time interactions.
Innovation Solution
A method and system for automated control of human inhabited characters using multiple input devices, including microphones, cameras, and hand-held controllers, to receive audio and image data, determine appearance states, and display the characters accordingly, reducing the need for extensive training.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If traditional input controllers with multiple buttons are used to control human inhabited characters, then the character can be controlled with precise command input, but the user requires significant time and effort to learn the complex input commands
Solution Approach 1:
The patent replaces the mechanical button-based input controller with automated sensing systems including audio data capture from microphones and image data capture from cameras. The system automatically determines appearance states and character behaviors based on this data, eliminating the need for users to learn complex button combinations while maintaining precise control capability.
Solution Approach 2:
The system performs self-service by automatically interpreting user intentions from audio and image data without requiring explicit manual commands. The automated control system processes the captured data to determine appropriate character appearances and behaviors, allowing the system to serve itself rather than requiring extensive user training.
2Loss of time
If multiple input devices are used to automatically control character appearance states, then training time is reduced, but the device complexity increases
Solution Approach 1:
The patent implements multi-functionality by using a single integrated control system that can process multiple types of input data (audio from microphones, images from cameras, and controller inputs) to determine character appearance states. This universal approach consolidates what would otherwise be separate complex subsystems into one coordinated system, reducing overall complexity despite the increased number of input devices.
Data Source
AI summary
Aspects of the present disclosure provide systems and methods for automated control of human inhabited characters. In an example, control of human inhabited character may be achieved via a plurality of input devices, including, but not limited to, a microphone, a camera, or a hand-held controller, that can modify and trigger changes in the appearance and/or the behavioral response of a character during the live interactions with humans. In an example, a computing device may include a neural network that receives the input from the microphone and/or the camera and changes the appearance and/or the behavioral response of the character according to the input. Further, input from the hand-held controller may be used to adjust a mood of the character or, in other words, emphasize or deemphasize the changes to the appearance and/or the behavioral response of the character.


