Voice Control via System Layer Extraction for Smart Terminals
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current voice control technologies for smart terminal devices require pre-integrated communication SDKs and specific call interfaces, limiting the application scope of voice control.
Innovation Solution
A method and apparatus that receive voice information and element information from a displayed page, perform voice recognition, match the recognition result with element content information, and generate page control information to control the page, allowing voice control without additional APP development.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If voice control requires pre-integrated communication SDK and specific call interfaces, then the control function can be implemented, but the application scope is limited
Solution Approach 1:
The patent extracts the voice recognition and matching functionality from the APP layer and places it in the system layer (notification bar or status bar). This extraction allows voice control to work across multiple APPs without requiring each APP to integrate communication SDKs, thereby expanding application scope while reducing device complexity
Solution Approach 2:
The patent creates a universal voice control mechanism that can control multiple different APPs through a common interface. The notification bar or status bar serves as a universal controller that can issue voice commands to various APPs, making the voice control system multi-functional and applicable across different contexts without requiring APP-specific integration
2Ease of operation
If voice control requires APP to provide call interface or pre-integrate communication SDK, then control function is achieved, but development complexity increases
Solution Approach 1:
The system provides self-service voice control capability through the notification bar or status bar, which automatically performs voice recognition and matching without requiring APP developers to implement additional interfaces. The system serves itself by handling the voice control logic centrally, eliminating the need for each APP to provide call interfaces or integrate communication SDKs
Data Source
AI summary
Embodiments of the present disclosure disclose a method and apparatus for controlling a page. A specific embodiment of the method comprises: receiving voice information from a terminal and element information of at least one element in a displayed page; performing voice recognition on the voice information to acquire a voice recognition result, in response to determining the voice information being used for controlling the displayed page; matching the voice recognition result with the element content information of the at least one element; and generating page control information in response to determining successfully matching the voice recognition result with the element content information of the at least one element, and sending the page control information to the terminal to allow the terminal to control the displayed page based on the page control information.


