Speech Modification Assistance Apparatus for User-Driven Text Editing
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing technologies for modifying speech data lack user intervention, making it difficult for users to easily adjust recorded speech data.
Innovation Solution
A speech modification assistance apparatus that allows users to select target recorded speech data, designate change target character strings, and generate modified speech data through a user-friendly interface.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If automatic extraction and automatic synthesis are used to modify speech data, then productivity is improved, but ease of operation deteriorates
Solution Approach 1:
The speech modification process is segmented into distinct phases: automatic extraction of partial character strings from recorded speech, automatic synthesis of modified speech, and user review/adjustment stages. This segmentation allows the system to handle different aspects of modification through appropriate methods (automatic vs. manual) at different stages, resolving the contradiction between productivity and ease of operation.
Solution Approach 2:
The system introduces an intermediary review and adjustment interface between the automatic extraction/synthesis processes and the final output. This intermediary stage allows users to review automatically generated modifications and make adjustments if needed, bridging the gap between automated processing and user control.
2Device complexity
If automatic speech modification is performed without user intervention, then device complexity is reduced, but ease of operation worsens
Solution Approach 1:
The system implements a dynamic modification process that can adapt between automatic and manual intervention based on user needs. The process starts with automatic extraction and synthesis but dynamically transitions to user review and adjustment stages, allowing the system to maintain simplicity while providing user control when necessary.
Data Source
AI summary
A speech modification assistance apparatus (10) includes one or more hardware processors configured to function as a first reception unit (22A), a display control unit (21), and a second reception unit (22B). The first reception unit (22A) receives selection of target recorded speech data, which is basic recorded speech data (70) to be processed, from among one or more pieces of basic recorded speech data (70) that are recorded. The display control unit (21) converts the target recorded speech data into a basic character string and displays the basic character string. The second reception unit (22B) receives designation of a change target character string to be changed in the displayed basic character string. The generation control unit (24) generates modified speech data corresponding to the target recorded speech data and the change target character string.


