Dynamic Voice Response Pace Control
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing voice processing systems do not effectively adapt the response pace to match the user's spoken pace, leading to suboptimal user experience in interactions.
Innovation Solution
A method where server devices analyze the user's spoken pace and compare it to a normal pace based on demographics and historical data, adjusting the response pace accordingly to improve interaction efficiency.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If the voice processing system uses a fixed standard pace for all responses, then the system operation is simple and consistent, but the user experience deteriorates because the response pace does not match the user's input pace
Solution Approach 1:
The voice processing system dynamically adjusts the response pace based on the detected user speaking pace. The system transitions from a static fixed pace to a dynamic adaptive pace by measuring the user's speech rate and modifying the response generation parameters accordingly, ensuring the response matches the user's input tempo
Solution Approach 2:
The system implements feedback by detecting the user's speaking pace from the input voice signal and using this information to control the output response pace. The detected pace serves as feedback that closes the loop between user input and system response, allowing real-time adaptation of the response characteristics
2Measurement precision
If the system analyzes user speaking pace in detail, then the response accuracy and user experience improve, but the processing time and system complexity increase
Solution Approach 1:
The system applies partial action by focusing measurement resources on the most critical aspect - the temporal spacing between phonemes or words. Rather than analyzing all speech characteristics in detail, the system concentrates computational effort on detecting speaking pace, achieving sufficient accuracy without excessive processing overhead
Data Source
AI summary
A system is configured to obtain a first voice request, from a client, to access a voice processing system that processes voice communications received from clients; determine a first pace at which terms, associated with the first voice request, are spoken by a user of the client; determine a second pace, associated with the user, based on terms, associated with other voice requests, spoken by the user and users of the clients prior to receiving the first voice request by using a weighted average of the pace associated with the user of the client and a pace associated with the users of the clients other than the client; compare the first pace to the second pace; determine a third pace based on the comparison; and send, to the client, a voice response to be outputted at the third pace.


