Voiceprint Recognition for Personalized Information Pushing
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Intelligent voice chat products struggle to recognize multiple users in a home setting and provide personalized services, as existing systems lack the ability to differentiate between users based on voice characteristics.
Innovation Solution
A method and apparatus that extracts voiceprint characteristics from awakening voice information, matches them with a preset registration voiceprint information set, and pushes targeted audio information based on user behavior data, allowing for personalized service delivery by establishing and managing user voiceprint information using a pre-trained universal background model.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If voiceprint recognition is implemented to recognize different users, then personalized service can be provided, but system complexity increases
Solution Approach 1:
The system performs preliminary voiceprint registration and characterization before actual user recognition. Voiceprint characteristic information is extracted and stored in advance during a registration phase, allowing the system to quickly match against pre-stored templates during operation, thereby reducing real-time processing complexity while enabling personalized service
Solution Approach 2:
A voiceprint characteristic extraction module serves as an intermediary between voice input and user identification. This intermediate processing layer transforms raw voice signals into standardized voiceprint features, simplifying the subsequent matching process and enabling personalized service through a structured intermediate representation
2Measurement precision
If voiceprint characteristic information is extracted and stored for multiple users, then user recognition accuracy improves, but information storage requirements increase
Solution Approach 1:
The system extracts only the essential voiceprint characteristic information from complete voice signals, separating and storing only the discriminative features needed for user identification. This extraction process removes redundant information while preserving recognition accuracy, thereby reducing storage requirements
Solution Approach 2:
Voiceprint characteristic information is transformed into compressed parameter representations suitable for efficient storage. By converting voice data into extracted feature parameters rather than storing raw audio, the system maintains recognition precision while significantly reducing the quantity of stored information
3Speed
If the system processes and analyzes voice information in real-time, then user interaction responsiveness improves, but energy consumption increases
Solution Approach 1:
Voiceprint characteristic information is extracted and prepared in advance during registration, creating pre-processed templates that can be quickly matched during real-time interaction. This preliminary processing shifts computational burden away from real-time operation, improving response speed while reducing energy consumption during active use
Solution Approach 2:
The system extracts only the necessary voiceprint features from incoming speech signals rather than processing complete audio data in real-time. This selective extraction reduces computational load and energy consumption while maintaining the responsiveness needed for natural user interaction
Data Source
AI summary
The present disclosure discloses a method and apparatus for pushing information. A specific embodiment of the method comprises: receiving voice information sent through a terminal by a user, the voice information including awakening voice information and querying voice information; extracting a voiceprint characteristic from the awakening voice information to obtain voiceprint characteristic information; matching the voiceprint characteristic information and a preset registration voiceprint information set, each piece of registration voiceprint information in the registration voiceprint information set including registration voiceprint characteristic information, and user behavior data of a registration user corresponding to the registration voiceprint characteristic information; and pushing, in response to the voiceprint characteristic information successfully matching the registration voiceprint characteristic information in the registration voiceprint information set, audio information to the terminal based on the querying voice information and user behavior data corresponding to the successfully matched registration voiceprint characteristic information.


