Text-Based Virtual Object Animation Generation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing virtual object animation generation technologies require significant labor and time, relying on specific voice actors and artists for dubbing and expression production, limiting versatility and increasing production costs.

Innovation Solution

A text-based virtual object animation generation method that acquires text information, analyzes emotional features and rhyme boundaries, performs speech synthesis, and generates synchronized virtual object animations, eliminating the need for specific voice actors and reducing manual intervention.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If traditional dubbing and animation production methods are used, then professional quality can be achieved, but labor costs and time costs are significant

Engineering Contradiction:
Improveprofessional qualityVSAvoidproduction efficiency
Core Design Contradiction:
ReliabilityVSProductivity

Solution Approach 1:

The system enables automatic generation of virtual object animations from text input, where the virtual object itself performs the animation without requiring human voice actors or animators. The text-to-speech synthesis and automatic animation generation processes allow the system to serve itself, eliminating the need for professional dubbing artists and reducing both labor costs and production time while maintaining acceptable quality.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent replaces the mechanical process of manual dubbing and animation production with automated text-to-speech synthesis and computer-generated animation. Instead of requiring human voice actors and animators to manually create the content, the system uses algorithmic processes to generate speech and synchronize virtual object movements, thereby significantly reducing labor requirements and production time.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Reliability

If specific voice actors are used for dubbing, then emotional expression quality is improved, but versatility is limited

Engineering Contradiction:
Improveemotional expression qualityVSAvoidversatility
Core Design Contradiction:
ReliabilityVSAdaptability or versatility

Solution Approach 1:

The text-to-speech synthesis system can generate speech with various emotional expressions and characteristics from a single text input, making the system universally applicable to different scenarios and audiences. Instead of being limited to specific voice actors, the system can adapt its synthesis to create different emotional tones and speech styles, thereby achieving both quality and versatility.

Inventive Principle:
Principle #6Universality (Multi-functionality)

3Manufacturing precision

If manual fixation of actor movement is performed, then animation precision is improved, but time consumption increases

Engineering Contradiction:
Improveanimation precisionVSAvoidtime consumption
Core Design Contradiction:
Manufacturing precisionVSLoss of time

Solution Approach 1:

The patent replaces the manual process of animators fixing and adjusting actor movements with automated animation generation algorithms. The system automatically synchronizes virtual object movements with the generated speech and text content, achieving precise animation without requiring time-consuming manual intervention. The automatic generation process maintains animation precision while dramatically reducing the time required for production.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Data Source

PatentUS11908451B2Text-based virtual object animation generation method, apparatus, storage medium, and terminal
Publication Date: 2024.02.20 MOFA (SHANGHAI) INFORMATION TECH CO LTD
  • US11908451B2 patent drawing
  • US11908451B2 patent drawing
  • US11908451B2 patent drawing

AI summary

A text-based virtual object animation generation includes acquiring text information, where the text information includes an original text of a virtual object animation to be generated; analyzing an emotional feature of the text information; performing speech synthesis according to the emotional feature, a rhyme boundary, and the text information to obtain audio information, where the audio information includes emotional speech obtained by conversion based on the original text; and generating a corresponding virtual object animation based on the text information and the audio information, where the virtual object animation is synchronized in time with the audio information.