Interactive Dubbing Aggregation for Multi-Character Reading
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing book reading applications only allow users to dub for selected characters without interaction, lacking a comprehensive and engaging dubbing experience.
Innovation Solution
A dubbing interaction method that enables multi-user and multi-character combined dubbing by associating user-generated and AI-generated dubbing audios with character information, allowing for interactive dubbing experiences through chat interfaces and dialect conversions.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If a general user dubbing function is provided that only allows users to dub for selected characters individually, then the implementation is simple, but the dubbing experience and user engagement are insufficient
Solution Approach 1:
The patent combines dubbing audios from multiple users and multiple characters into a single integrated dubbing work. The system merges individual dubbing contributions into a cohesive whole, allowing users to experience complete story dubbing rather than isolated character dubbing. This resolving the contradiction by enabling rich dubbing experiences through aggregation while managing complexity through automated assembly processes.
Solution Approach 2:
The dubbing system is designed to handle multiple functions: individual character dubbing, multi-user collaboration, automatic audio aggregation, and integrated playback. The system serves both individual users and groups, supporting both simple and complex dubbing scenarios within a single platform, thus improving versatility without proportionally increasing complexity.
2Productivity
If multi-user and multi-character combined dubbing is enabled, then user engagement and interaction are enhanced, but the system complexity and processing requirements increase
Solution Approach 1:
The system segments the dubbing process into independent modules: individual character selection, separate dubbing recording for each character, independent audio processing, and modular aggregation. This segmentation allows multiple users to work on different characters simultaneously without interfering with each other, enhancing user engagement while managing system complexity through divided processing tasks.
Solution Approach 2:
The system introduces an intermediary aggregation module that collects, synchronizes, and integrates dubbing audios from multiple users and characters. This intermediary component manages the complexity of multi-user interactions by providing a centralized coordination layer, allowing high user engagement while containing system complexity within the mediation function.
3Reliability
If dubbing audios from multiple users and characters are aggregated into a complete dubbing work, then the dubbing experience is comprehensive, but the processing and management complexity increases
Solution Approach 1:
The system performs preliminary organization of dubbing audios by character and user before final aggregation. Each user's dubbing for each character is pre-labeled, pre-synchronized with corresponding text, and pre-formatted. This preliminary action ensures complete and reliable dubbing works while reducing processing complexity during the final aggregation stage, as the system only needs to assemble pre-processed segments rather than manage raw audio data.
Data Source
AI summary
A dubbing interaction method and an apparatus, a computer device, and a storage medium are provided. The method includes: showing character information associated with a text to be dubbed; in response to a selection operation for first character information, obtaining a first dubbing audio of a first user, and associating the first dubbing audio with the first character information, wherein the first dubbing audio is a dub for a text fragment of the text to be dubbed that is associated with the first character information; and obtaining an aggregated dubbing audio corresponding to the text to be dubbed based on the first dubbing audio associated with the first character information and an obtained second dubbing audio respectively corresponding to second character information associated with the text to be dubbed, and showing a first audio identifier corresponding to the aggregated dubbing audio.


