3D Avatar UI with Selectable Speaking Actions for Text

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current avatar user interfaces lack the ability to provide dynamic and customizable speaking actions for selected text, limiting user interaction and engagement with graphical representations.

Innovation Solution

A method is implemented where a client computing device determines user-selected text and presents a UI element with selectable speaking actions, such as 'Speak,' 'Pronounce,' and 'AI-View,' which trigger animations and synchronized speech audio for a 3D avatar, utilizing a text-to-speech module and sentiment analysis to generate appropriate audio and visual responses.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If a basic text-to-speech function is provided for selected text, then the system maintains simplicity and ease of operation, but user engagement and interaction depth are limited

Engineering Contradiction:
Improveease of operationVSAvoiduser engagement
Core Design Contradiction:
Ease of operationVSAdaptability or versatility

Solution Approach 1:

The system transitions from a static, single-function text-to-speech interface to a dynamic interface that adapts based on user selection and context. Multiple speaking actions (Speak, Pronounce, AI-View) are dynamically presented through a popover menu, allowing the system to adjust its functionality based on user needs while maintaining ease of operation through intuitive selection.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The avatar system is designed to perform multiple functions beyond basic text-to-speech. It can speak text aloud, demonstrate pronunciation with visual mouth movements, provide AI-generated analysis of selected text, and offer customizable appearance options. This multi-functionality increases user engagement while maintaining a unified, easy-to-use interface.

Inventive Principle:
Principle #6Universality (Multi-functionality)

2Adaptability or versatility

If multiple speaking actions and customizable avatar features are added, then user engagement and interaction depth improve, but system complexity increases

Engineering Contradiction:
Improveuser engagementVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The system segments complex functionality into distinct, manageable actions (Speak, Pronounce, AI-View) that are presented separately in a popover menu. Each action corresponds to a specific function, allowing the system to offer comprehensive capabilities while keeping the interface organized and not overwhelming users with complexity.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The popover menu serves as an intermediary element that mediates between the simple text selection action and the multiple complex speaking actions. It provides a clean, organized presentation of available functions without exposing the underlying system complexity to the user, maintaining ease of operation while enabling advanced functionality.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Device complexity

If static avatar representations are used, then the system maintains simplicity, but user engagement and emotional connection are limited

Engineering Contradiction:
Improvesystem simplicityVSAvoiduser engagement
Core Design Contradiction:
Device complexityVSAdaptability or versatility

Solution Approach 1:

The system replaces static avatar images with dynamic 3D models that can animate in response to user actions. The avatar performs synchronized mouth movements during pronunciation, displays different facial expressions, and responds to user selection, creating a more engaging and emotionally connective experience while maintaining system simplicity through efficient rendering and animation techniques.

Inventive Principle:
Principle #15Dynamics

Data Source

PatentUS20240087199A1Avatar UI with Multiple Speaking Actions for Selected Text
Publication Date: 2024.03.14 SAMSUNG ELECTRONICS CO LTD
  • US20240087199A1 patent drawing
  • US20240087199A1 patent drawing
  • US20240087199A1 patent drawing

AI summary

In one embodiment, a method includes determining, by a client computing device, that a user has selected text displayed on a display of the client computing device. The method further includes presenting, in response to the determination, a UI element on the display of the client computing device. The UI element includes a plurality of selectable portions, each associated with a distinct speaking action for a 3D avatar to perform with respect to the selected text. In response to the user's selection, the method includes presenting on the display of the client computing device an animation of the 3D avatar performing a speaking action corresponding to the selected portion; and providing, by the client computing device, speech audio synchronized with the speaking action of the animated 3D avatar.