Interactive Dubbing Aggregation for Multi-Character Reading

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing book reading applications only allow users to dub for selected characters without interaction, lacking a comprehensive and engaging dubbing experience.

Innovation Solution

A dubbing interaction method that enables multi-user and multi-character combined dubbing by associating user-generated and AI-generated dubbing audios with character information, allowing for interactive dubbing experiences through chat interfaces and dialect conversions.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If a general user dubbing function is provided that only allows users to dub for selected characters individually, then the implementation is simple, but the dubbing experience and user engagement are insufficient

Engineering Contradiction:
Improvedubbing experienceVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent combines dubbing audios from multiple users and multiple characters into a single integrated dubbing work. The system merges individual dubbing contributions into a cohesive whole, allowing users to experience complete story dubbing rather than isolated character dubbing. This resolving the contradiction by enabling rich dubbing experiences through aggregation while managing complexity through automated assembly processes.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The dubbing system is designed to handle multiple functions: individual character dubbing, multi-user collaboration, automatic audio aggregation, and integrated playback. The system serves both individual users and groups, supporting both simple and complex dubbing scenarios within a single platform, thus improving versatility without proportionally increasing complexity.

Inventive Principle:
Principle #6Universality (Multi-functionality)

2Productivity

If multi-user and multi-character combined dubbing is enabled, then user engagement and interaction are enhanced, but the system complexity and processing requirements increase

Engineering Contradiction:
Improveuser engagementVSAvoidsystem complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The system segments the dubbing process into independent modules: individual character selection, separate dubbing recording for each character, independent audio processing, and modular aggregation. This segmentation allows multiple users to work on different characters simultaneously without interfering with each other, enhancing user engagement while managing system complexity through divided processing tasks.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system introduces an intermediary aggregation module that collects, synchronizes, and integrates dubbing audios from multiple users and characters. This intermediary component manages the complexity of multi-user interactions by providing a centralized coordination layer, allowing high user engagement while containing system complexity within the mediation function.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Reliability

If dubbing audios from multiple users and characters are aggregated into a complete dubbing work, then the dubbing experience is comprehensive, but the processing and management complexity increases

Engineering Contradiction:
Improvedubbing completenessVSAvoidprocessing complexity
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The system performs preliminary organization of dubbing audios by character and user before final aggregation. Each user's dubbing for each character is pre-labeled, pre-synchronized with corresponding text, and pre-formatted. This preliminary action ensures complete and reliable dubbing works while reducing processing complexity during the final aggregation stage, as the system only needs to assemble pre-processed segments rather than manage raw audio data.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS12512125B2Dubbing interaction method and apparatus, computer device, and storage medium
Publication Date: 2025.12.30 BEIJING ZITIAO NETWORK TECH CO LTD
  • US12512125B2 patent drawing
  • US12512125B2 patent drawing
  • US12512125B2 patent drawing

AI summary

A dubbing interaction method and an apparatus, a computer device, and a storage medium are provided. The method includes: showing character information associated with a text to be dubbed; in response to a selection operation for first character information, obtaining a first dubbing audio of a first user, and associating the first dubbing audio with the first character information, wherein the first dubbing audio is a dub for a text fragment of the text to be dubbed that is associated with the first character information; and obtaining an aggregated dubbing audio corresponding to the text to be dubbed based on the first dubbing audio associated with the first character information and an obtained second dubbing audio respectively corresponding to second character information associated with the text to be dubbed, and showing a first audio identifier corresponding to the aggregated dubbing audio.