Personalized Voice Interaction Using User Voice Imitation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing intelligent devices provide limited personalized voice interaction experiences, often relying on fixed responses and failing to recognize individual users, resulting in suboptimal user engagement.

Innovation Solution

An electronic device that recognizes user voice information to initiate a voice conversation by imitating the voice and mannerisms of a specific user, enhancing interaction performance and providing a more personalized experience.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If the electronic device uses fixed patterned voice replies based on a set voice mode, then the device complexity is reduced, but the adaptability or versatility of voice interaction deteriorates

Engineering Contradiction:
Improvevoice interaction adaptabilityVSAvoidvoice processing complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent applies voice cloning technology to replicate the voice characteristics, tone, and mannerisms of specific users. The system creates virtual copies of user voices and uses them to generate personalized responses, enabling the device to adapt its voice output to match individual users without requiring complex real-time synthesis for each interaction scenario

Inventive Principle:
Principle #26Copying

Solution Approach 2:

The system dynamically adjusts voice parameters including pitch, tone, speed, and accent based on the identified user and conversation context. The voice synthesis engine adapts its output in real-time to match the characteristics of different users, transforming fixed patterned replies into dynamic, personalized voice interactions

Inventive Principle:
Principle #15Dynamics

2Ease of operation

If the electronic device provides personalized voice interaction, then the user experience is improved, but the device complexity increases

Engineering Contradiction:
Improveuser interaction qualityVSAvoidvoice processing complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The system incorporates feedback mechanisms that analyze user responses and conversation flow to refine voice generation. The device monitors interaction patterns and adjusts its voice output accordingly, creating a feedback loop that improves personalization over time while managing computational resources efficiently

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The system performs preliminary voice analysis and user identification before generating responses. By pre-processing voice characteristics and storing user profiles, the device can quickly retrieve and apply appropriate voice parameters during actual interaction, reducing real-time computational complexity

Inventive Principle:
Principle #10Preliminary action

3Measurement precision

If the electronic device recognizes and imitates user voices, then the measurement precision of voice identification is improved, but the loss of information increases

Engineering Contradiction:
Improvevoice recognition accuracyVSAvoidvoice data privacy
Core Design Contradiction:
Measurement precisionVSLoss of information

Solution Approach 1:

The system extracts only the essential voice characteristics needed for identification and synthesis, such as pitch contours, tone patterns, and rhythmic features. By extracting only these key parameters rather than storing complete voice recordings, the system achieves accurate voice recognition while minimizing data retention and privacy risks

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The patent introduces voice synthesis as an intermediary layer between voice recognition and response generation. Instead of directly analyzing and storing raw user voice data, the system converts recognized voice patterns into synthesized responses, reducing the need to store and process sensitive original voice information

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS12494199B2Voice interaction method and electronic device
Publication Date: 2025.12.09 HUAWEI TECH CO LTD
  • US12494199B2 patent drawing
  • US12494199B2 patent drawing
  • US12494199B2 patent drawing

AI summary

This application provides a voice interaction method and an electronic device, and relates to the field of artificial intelligence (AI) technologies and the field of voice processing technologies. An example solution includes: An electronic device receiving first voice information sent by a second user, and the electronic device recognizing the first voice information in response to receiving the first voice information. The first voice information is used to request a voice conversation with a first user. The electronic device may have, on a basis that the electronic device recognizes that the first voice information is voice information of the second user, a voice conversation with the second user by imitating a voice of the first user and in a mode in which the first user has a voice conversation with the second user.