Voiceprint Recognition for Personalized Information Pushing

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Intelligent voice chat products struggle to recognize multiple users in a home setting and provide personalized services, as existing systems lack the ability to differentiate between users based on voice characteristics.

Innovation Solution

A method and apparatus that extracts voiceprint characteristics from awakening voice information, matches them with a preset registration voiceprint information set, and pushes targeted audio information based on user behavior data, allowing for personalized service delivery by establishing and managing user voiceprint information using a pre-trained universal background model.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If voiceprint recognition is implemented to recognize different users, then personalized service can be provided, but system complexity increases

Engineering Contradiction:
Improvepersonalized service capabilityVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The system performs preliminary voiceprint registration and characterization before actual user recognition. Voiceprint characteristic information is extracted and stored in advance during a registration phase, allowing the system to quickly match against pre-stored templates during operation, thereby reducing real-time processing complexity while enabling personalized service

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

A voiceprint characteristic extraction module serves as an intermediary between voice input and user identification. This intermediate processing layer transforms raw voice signals into standardized voiceprint features, simplifying the subsequent matching process and enabling personalized service through a structured intermediate representation

Inventive Principle:
Principle #24Intermediary (Mediator)

2Measurement precision

If voiceprint characteristic information is extracted and stored for multiple users, then user recognition accuracy improves, but information storage requirements increase

Engineering Contradiction:
Improveuser recognition accuracyVSAvoidinformation storage volume
Core Design Contradiction:
Measurement precisionVSQuantity of substance

Solution Approach 1:

The system extracts only the essential voiceprint characteristic information from complete voice signals, separating and storing only the discriminative features needed for user identification. This extraction process removes redundant information while preserving recognition accuracy, thereby reducing storage requirements

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

Voiceprint characteristic information is transformed into compressed parameter representations suitable for efficient storage. By converting voice data into extracted feature parameters rather than storing raw audio, the system maintains recognition precision while significantly reducing the quantity of stored information

Inventive Principle:
Principle #35Parameter changes

3Speed

If the system processes and analyzes voice information in real-time, then user interaction responsiveness improves, but energy consumption increases

Engineering Contradiction:
Improveresponse speedVSAvoidenergy consumption
Core Design Contradiction:
SpeedVSUse of energy by moving object

Solution Approach 1:

Voiceprint characteristic information is extracted and prepared in advance during registration, creating pre-processed templates that can be quickly matched during real-time interaction. This preliminary processing shifts computational burden away from real-time operation, improving response speed while reducing energy consumption during active use

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system extracts only the necessary voiceprint features from incoming speech signals rather than processing complete audio data in real-time. This selective extraction reduces computational load and energy consumption while maintaining the responsiveness needed for natural user interaction

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS10832686B2Method and apparatus for pushing information
Publication Date: 2020.11.10 BAIDU ONLINE NETWORK TECH (BEIJIBG) CO LTD
  • US10832686B2 patent drawing
  • US10832686B2 patent drawing
  • US10832686B2 patent drawing

AI summary

The present disclosure discloses a method and apparatus for pushing information. A specific embodiment of the method comprises: receiving voice information sent through a terminal by a user, the voice information including awakening voice information and querying voice information; extracting a voiceprint characteristic from the awakening voice information to obtain voiceprint characteristic information; matching the voiceprint characteristic information and a preset registration voiceprint information set, each piece of registration voiceprint information in the registration voiceprint information set including registration voiceprint characteristic information, and user behavior data of a registration user corresponding to the registration voiceprint characteristic information; and pushing, in response to the voiceprint characteristic information successfully matching the registration voiceprint characteristic information in the registration voiceprint information set, audio information to the terminal based on the querying voice information and user behavior data corresponding to the successfully matched registration voiceprint characteristic information.