User Terminal Volume Control Through Sound Scene Recognition
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing mobile phone volume adjustment technologies inaccurately determine scenarios due to reliance on ambient noise levels, leading to inappropriate volume settings, which can be mistaken for human activity, resulting in suboptimal user experience.
Innovation Solution
A method and apparatus that analyze sound signals to differentiate between blank sound, human sound, and noise, determining a scene mode and adjusting volume accordingly, using mel-frequency cepstral coefficients to quantify human presence and adjust ring tone and earpiece volumes based on pre-stored coefficients.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If volume is adjusted according to ambient noise decibels only, then volume adjustment is simple, but accuracy of volume adjustment is low
Solution Approach 1:
The patent segments the ambient sound signal into three distinct components: blank sound (silence), human sound (speech), and noise (other sounds). By analyzing the proportions of these different sound types separately, the system achieves more accurate scene recognition than simply measuring overall noise decibels, while maintaining automated operation.
Solution Approach 2:
The patent changes the parameter used for volume adjustment from a single parameter (noise decibels) to multiple parameters (proportions of blank sound, human sound, and noise). This multi-parameter approach allows the system to distinguish between different sound sources and adjust volume more accurately according to the actual scene.
2Extent of automation
If environment sound decibels are used to determine scene mode, then volume adjustment is automated, but scene determination accuracy is low
Solution Approach 1:
The patent automatically segments the ambient sound into three categories (blank sound, human sound, noise) and calculates their respective proportions. This automated segmentation enables accurate scene determination by analyzing the composition of ambient sounds rather than relying on a single decibel measurement.
Solution Approach 2:
The system continuously monitors ambient sound, analyzes the proportions of different sound types, and adjusts volume based on the determined scene mode. This closed-loop feedback mechanism ensures automated volume adjustment adapts to changing environmental conditions while maintaining high accuracy through ongoing sound composition analysis.
Data Source
AI summary
A volume adjustment method and apparatus, and a terminal is presented. Perform analysis on the collected sound signal surrounding a user terminal, to obtain composition information, where the composition information includes sound types included in the sound signal and proportions of sounds of the various types, and the sound types include blank sound, human sound, and noise; determine a current scene mode of the user terminal according to the composition information; and adjust volume of the user terminal according to the determined scene mode, thereby significantly reducing occurrence of a case, caused by mistaken determining of the scenario, in which play volume adjustment does not conform to the scenario, and enhancing user experience.


