AI Avatar Emotion Recognition via Learning Model
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing avatar technologies are limited in providing realistic, interactive, and emotionally responsive interfaces in non-face-to-face environments, lacking the ability to effectively communicate and interact with humans using advanced AI and sensor technologies.
Innovation Solution
An avatar-based interaction service method and apparatus that utilizes a computer system to provide an AI avatar capable of interacting with users by analyzing and reflecting the image and voice of a service provider, training responses based on a learning model, and generating an AI avatar to provide services in fields like customer service, counseling, and education, while recognizing and responding to user emotions through facial expressions, gestures, and voice tones.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If traditional two-dimensional avatars are used in online chats and games, then the implementation is simple and cost-effective, but the realism and immersive experience are poor
Solution Approach 1:
The patent creates a virtual copy of the user through avatar technology, replicating their appearance, voice, and behavioral characteristics in the virtual space. This allows the avatar to serve as a realistic representation of the user without requiring the user to physically be present, thereby improving realism while maintaining operational simplicity through automated copying processes
Solution Approach 2:
The patent replaces traditional mechanical avatar control methods with AI-driven automated systems. Instead of manual animation or pre-recorded sequences, the system uses machine learning models to automatically generate realistic avatar movements, expressions, and responses based on user input and contextual understanding, thereby improving realism without proportionally increasing complexity
2Adaptability or versatility
If AI technology and sensor technology are integrated into avatars to enable practical interaction and communication with humans, then the interaction capability is improved, but the system complexity increases
Solution Approach 1:
The patent integrates multiple functions into a single avatar system, including image processing, voice recognition, emotional analysis, and natural language generation. Rather than separate systems for each function, the avatar serves as a universal interface that handles various interaction modes (visual, auditory, emotional) through unified AI processing, thereby improving adaptability while managing complexity through functional integration
Solution Approach 2:
The patent introduces an AI processing layer as an intermediary between users and the avatar system. This intermediary handles the complex tasks of interpreting user inputs, generating appropriate responses, and coordinating multiple sensor and actuator systems, thereby enabling versatile interaction capability while shielding users from the underlying system complexity
3Ease of operation
If avatars are used in non-face-to-face conversation environments to reflect service provider image and voice, then user accessibility is improved, but the emotional connection and trust are reduced
Solution Approach 1:
The patent employs dynamic visual modifications of the avatar's appearance, including facial expressions, gestures, and visual effects that change in response to emotional context. These visual variations allow the avatar to convey emotional states and intentions, thereby maintaining emotional connection and trust in non-face-to-face environments while preserving the accessibility benefits of remote interaction
Solution Approach 2:
The patent replaces traditional text-based or static image communication with AI-generated dynamic visual and auditory responses. The system automatically generates realistic voice modulations, facial expressions, and body language that convey emotional nuance, thereby restoring emotional connection in remote interactions without requiring physical presence or complex user setup
Data Source
AI summary
Provided is an avatar-based interaction service method performed by a computer system including: providing an interaction service to a first user terminal through an avatar of a service provider reflecting an image and a voice of the service provider in a non-face-to-face conversation environment between the service provider and a first user; training a response of the service provider to the first user based on a pre-stored learning model; and providing the interaction service to a second user terminal by generating an artificial intelligence (AI) avatar based on the trained learning model.


