Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

680 results about "Speech control" patented technology

Large model tool calling method and system based on risk assessment

The invention discloses a large model tool calling method and system based on risk assessment, and the method comprises the following steps: receiving a natural language instruction of a user, and obtaining a calling request of the user based on a large model; generating a risk score based on the user call request; generating a confirmation strategy based on the risk score, the user preference adjustment factor and the emergency degree adjustment factor; executing a corresponding confirmation process based on the confirmation strategy; based on the confirmation process, the controller executes tool calling; by the adoption of the technical scheme, the problem of balance between safety and user experience in tool calling is solved, a safer and more intelligent interaction mechanism is provided for voice control and large language model driven tool calling, a risk self-adaptive confirmation process and a multi-stage confirmation mechanism provide sufficient safety guarantee, and meanwhile the safety of a user is improved. And the user interaction experience is improved to the greatest extent.
Owner:福建汉特云智能科技有限公司

Audio-based user engagement detection

A system can operate a speech-controlled device to perform user engagement detection (UED) processing to detect when speech represented in audio data is directed to the device. For example, the device may extract audio features from the audio data and process these audio features using a classifier to estimate an orientation of the user's head, which may be used as a proxy for user engagement. Thus, if the head orientation is within an engagement zone (which varies based on distance to the user), the device may determine that the user is engaged with the device and perform language processing on input speech. In contrast, if the head orientation is outside of the engagement zone, the device may determine that the user is not engaged and ignore the input speech. To enable additional functionality, the classifier may optionally output a coarse estimate of the head orientation along with the UED determination.
Owner:AMAZON TECH INC

Voice control method, device, equipment, medium and product

The invention discloses a voice control method and device, equipment, a medium and a product. The method comprises the following steps: acquiring first voice information; performing intention recognition based on the first voice information to obtain the target intention; according to the target intention, a target audio playing mode is determined from multiple preset audio playing modes, and the multiple audio playing modes comprise playing through an audio playing assembly in a vehicle cabin and playing through an audio playing assembly outside the vehicle cabin; based on the target intention and the first voice information, determining a target audio responding to the first voice information; and controlling the vehicle to play the target audio according to the target audio playing mode. The accuracy of voice control can be improved.
Owner:ZHEJIANG GEELY HLDG GRP CO LTD +1

Voice control method and device for household electrical appliance

The invention discloses a voice control method and device for a household electrical appliance, and the method comprises the following steps: S1, collecting a user voice signal, and carrying out the preprocessing of the user voice signal; s2, performing local wake-up detection and local voice recognition on the preprocessed voice signal to obtain a text instruction; s3, performing complexity analysis on the text instruction; s4, based on the result of the complexity analysis, processing the text instruction through a local path or a cloud path to generate a control instruction; and S5, controlling the target household electrical appliance to execute corresponding operation according to the control instruction. According to the invention, the recognition efficiency and the recognition accuracy can be improved; the natural language expression of the user can be converted into accurate equipment linkage operation in combination with voice context information; a simple control instruction is subjected to localization processing, only necessary data are uploaded to the cloud, and the risk of privacy disclosure is reduced; personalized services are provided by binding voiceprint recognition with user preferences.
Owner:江苏宇石智能科技有限公司

Vehicle control method and device and vehicle

The invention discloses a vehicle control method and device and a vehicle, and relates to the technical field of vehicle control and voice interaction. The vehicle comprises a vehicle machine screen. The method comprises the following steps: determining an operation object on the vehicle machine screen according to a voice operation instruction input by a user; and when the number of the operation objects is greater than 1, identifying an intention operation object of the voice operation instruction in all the determined operation objects according to the historical operation of the user on the vehicle machine screen and / or the pointing information contained in the voice operation instruction and / or the vehicle machine application where the operation objects are located. According to the method, accurate matching of the voice operation instruction can be achieved, especially under the condition that multiple operation objects exist on the vehicle-mounted terminal screen, the operation object intended by the user can be accurately recognized, and therefore the accuracy and reliability of voice control or voice interaction are improved, and the user has good voice interaction experience.
Owner:GREAT WALL MOTOR CO LTD

Voice control in a healthcare facility

Systems for voice control of medical devices in a healthcare facility are disclosed herein. The systems employ continuous speech processing software, voice recognition software, natural language processing software, and other software to permit voice control of the medical devices. Systems are also provided for distinguishing which medical device from among multiple medical devices in a patient room is the particular medical device to be controlled by voice input from a caregiver or a patient.
Owner:HILL ROM SERVICES INC

Robot control method based on voice analysis

The invention relates to the field of voice analysis, in particular to a robot control method based on voice analysis, and the method comprises the steps: obtaining image information in a space region where a target robot is located, determining semantic tags, determining a potential association semantic tag group, and screening out a semantic guiding corpus group for the target robot; when voice control data is received, instruction fuzzy parameters of the voice control data are determined so as to judge instruction fuzzy tendency, optimization is carried out on the voice control data with the instruction fuzzy tendency, specifically, confidence centralized clusters and semantic discrete clusters are determined, semantic expansion is carried out on the semantic discrete clusters, and then an expansion text is obtained; and screening the expanded text based on the semantic oriented corpus group to obtain a confidence instruction text. According to the method, the semantic oriented corpus group is constructed in combination with the image information, the analysis of the voice control data with the instruction fuzzy tendency is guided, the analysis accuracy of the voice control data under the semantic fuzzy condition is improved, and the control instruction recognition precision is ensured.
Owner:厦门工学院

AI glasses automatic shooting method and system based on voice control

The invention provides an AI glasses automatic shooting method and system based on voice control, and relates to the technical field of intelligent wearable equipment, accurate interaction is realized through a multi-channel directional microphone array and an end-cloud collaborative voice recognition engine, the microphone array optimizes the pickup angle based on the wearing position characteristics of a user, and the user experience is improved. In combination with real-time voice activity detection, environmental noise is filtered out, it is ensured that a clear voice instruction can still be captured in a noisy environment, an end-cloud cooperation mode operates a lightweight model locally to guarantee the off-line response speed, a cloud large model is called when a network is available to improve the complex instruction analysis capability, and a composite statement containing parameter adjustment can be recognized; and an operation intention and parameters are automatically bound through a natural language processing engine, so that one-step execution of the instruction is realized.
Owner:MIODAO CLOUD COMPUTING (HANGZHOU) CO LTD

Interactive three-dimensional holographic display system for smart farm

The invention provides an interactive three-dimensional holographic display method for a smart farm, and relates to the technical field of smart agriculture and holographic visualization. Comprising the following steps: firstly, performing multi-angle aerial photography in different growth stages of a rice field through an unmanned aerial vehicle to collect image data; processing the data by using a three-dimensional reconstruction algorithm, and constructing a three-dimensional model library of the rice field; on the basis, a voice interaction module is integrated to a visual platform, so that a user can select and switch models and inquire growth information through voice control; further converting the selected three-dimensional model data into structured data suitable for holographic reconstruction, and generating a high-fidelity digital hologram by applying a complex-valued convolutional neural network point cloud gridding algorithm; and finally, performing optical reconstruction on the hologram by driving holographic display equipment. According to the method, the limitation of a traditional two-dimensional image in space and time sequence expression is overcome, dynamic visualization of the rice field growth process is achieved, and the monitoring and decision-making effects are improved.
Owner:YANGZHOU UNIV

Intelligent equipment integrated control method and system based on cloud-side cooperation

The embodiment of the invention provides an intelligent equipment integrated control method and system based on cloud-side cooperation, and relates to the field of equipment control. The method is applied to intelligent integrated equipment integrating at least one sub-control system of short-range communication, voice control and IOT control, and is associated with a client. The method comprises the following steps: receiving a first control instruction which is sent by an IoT PaaS platform and comprises a voice instruction or a physical instruction sent by a client; the voice instruction comprises an instruction determined by the voice cloud platform based on first voice data, and the first voice data comprises data sent by the intelligent integrated device or the client and obtained according to second voice data input by a user; determining a target device, protocol information thereof and a second control instruction according to the first control instruction; and sending a second control instruction to the target device through the sub-control system or the IoT SaaS platform based on the protocol information, so that the target device executes the second control instruction. The equipment control efficiency can be improved.
Owner:GUANGZHOU HEMI TECHNOLOGY CO LTD

Voice control method and voice control system

The invention provides a voice control method and a voice control system.The method comprises the steps that after a voice control instruction sent by a user in a physical space is sensed, voice content in the voice control instruction is converted into a first text; the application arrangement system classifies the voice control instruction of the user based on the content in the first text to determine an instruction type; according to the space identifier, obtaining a controlled device list corresponding to the physical space, and according to the instruction type, obtaining target device information, related to the instruction type, of each controlled device in the physical space; inputting the first text, the controlled equipment list and the target equipment information into a first large language model, outputting equipment control parameters matched with the control intention of the user through the first large language model, and transmitting the equipment control parameters to a central control host; and the central control host controls the target controlled equipment according to the control parameters. Through the method, the problem of error control is avoided.
Owner:BEIJING PILOT ZHILIAN INTERNET OF THINGS TECH CO LTD

Bluetooth earphone voice control method, system and equipment based on AI

The invention relates to the technical field of Bluetooth headset control, and discloses an AI-based Bluetooth headset voice control method, system and equipment, and the method comprises the steps: obtaining real-time audio and voice signals, and obtaining frequency distribution characteristics through Fourier transform; and obtaining a matched peak value position set through multi-layer convolution operation and peak value detection. Single sound source feature subsets are extracted and clustered, and classified sound source groups are obtained; and carrying out fusion analysis on the resonance feature vectors, generating a preliminary tuning parameter set and an alternative list, and determining an optimal set. And a dynamic tuning scheme is obtained through smooth processing and weight adjustment. And after the voice signal instruction is matched, activating the scheme, adjusting the real-time audio, and outputting an optimized signal. According to the method, intelligent real-time audio optimization and multi-band adaptive adjustment based on the voice instruction can be realized, and the tone quality performance and auditory experience of the Bluetooth earphone in different scenes are improved.
Owner:SHENZHEN SHINETEK TECH CO LTD

Electric cooker voice safety control system and method based on voiceprint verification

The invention relates to the technical field of intelligent household electrical appliance control and biological characteristic recognition, and discloses an electric cooker voice safety control system and method based on voiceprint verification, the system comprises a multi-mode sensing array, a storage unit and a core processing unit, the method comprises the following steps: generating and storing equipment physical fingerprints representing unique physical characteristics of equipment, the method comprises the following steps: generating and storing a multi-dimensional feature template comprising an acoustic voiceprint, living body interaction and a spatial vector template for an authorized user, when a voice instruction is received, synchronously acquiring a to-be-detected acoustic and structural vibration signal by a system, executing layered collaborative authentication, and sequentially performing spatial domain verification, physical interaction verification and user identity verification, the physical interaction verification identifies the living body attribute of the sound source by comparing the real-time sound vibration coupling characteristic with the equipment physical fingerprint, so that the non-living body attack is effectively identified and intercepted, and the safety and reliability of voice control are improved.
Owner:LINGNAN NORMAL UNIV

Label printing method based on voice control

The invention relates to the technical field of label printing, in particular to a label printing method based on voice control, and the method comprises the steps: obtaining user voice data through an interaction terminal, recognizing the voice data, and generating a recognition text; semantic analysis is conducted on the recognition text, and printing content and printing parameters are extracted; automatically generating a corresponding label printing layout according to the printing content and the printing parameters; and based on the label printing layout, driving the printing equipment to execute printing operation. According to the application, user voice can be recognized, and an engine for the dialect is called to select a voice recognition model for the specific dialect and a natural language understanding model for accurate recognition; according to the invention, the first step of conversion from voice to text is ensured to have extremely high accuracy, and the problems of label content error, format disorder and the like caused by misrecognition are fundamentally avoided, so that the waste of consumables such as label paper, colored tapes and the like caused by printing errors is reduced, and the overall economic benefit is improved.
Owner:新疆恒信技术服务有限公司

Voice control method and device, electronic equipment, computer readable storage medium and computer program product

The invention provides a voice control method and device, electronic equipment, a computer readable storage medium, a computer program product and a computer readable storage medium. The method comprises the following steps: in response to an instruction generation request of a user for a display interface, obtaining a control voice stream carried by the instruction generation request and an interface screenshot corresponding to the display interface; voice recognition is conducted on the control voice stream to obtain a recognition text, the interface screenshot is analyzed to obtain interface element description, and the interface element description is used for describing a plurality of interface elements existing in the interface screenshot; performing intention recognition based on the recognition text and the interface element description to obtain a control intention of the user for a target interface element in the plurality of interface elements; and generating a control instruction for the target interface element based on the control intention, and executing a corresponding control operation for the target interface element based on the control instruction. According to the invention, the universality of voice control for the electronic equipment can be improved.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

System and method for automated voice-based healthcare planning using a supplemental clinician user interface

A system and method providing automated voice-based healthcare plan delivery. The system interfaces with the patient through a voice-controlled personal assistant appliance via the internet. Some embodiments are implemented as voice applications or skills within the eco-system of an existing voice-controlled personal assistant device / appliance (hereinafter VCPAD). The system and method allow a clinician to create, digitally input, and modify a healthcare plan for a patient using a simple computer / web interface and without the need of any software development or computer programming skills, and allows a patient to interface with and query that healthcare plan using a largely conventional voice-controlled digital assistant appliance simply and efficiently to extract information about his / her individual healthcare plan in the patient's own language.
Owner:CHRISTIANA CARE HEALTH SYSTEM INC

Voice control method, multimedia system, vehicle, and storage medium

Voice control method, multimedia system, vehicle, and storage medium. This method includes performing semantic analysis on a received target voice command to obtain a semantic analysis result (S101), when the semantic analysis result matches a preset keyword, responding to the target voice command by using the application corresponding to the preset keyword (S102), when the semantic analysis result does not match the preset keyword, determining a target service type according to the semantic analysis result and obtaining an application list corresponding to the target service type (S103), determining a recommended application from the application list by using a first recommendation policy (S104), and responding to the target voice command by using the recommended application (S105).
Owner:BYD CO LTD

Display apparatus and voice control method

The invention provides a display device and a voice control method. The display device comprises a display, a sound collector and a controller connected with the display and the sound collector. Wherein the display is configured to display an image picture and a user interface; the sound collector is configured to collect a voice control instruction of a user, and the controller is configured to determine at least one scrolling control contained in a current display interface; constructing a voice scrolling control word list of a scrolling control in the current display interface; the voice scrolling control word list is used for representing a corresponding relationship between the scrolling direction of the scrolling control and the semantic control word; and in response to a voice control instruction of a user, controlling a target scrolling control in the current display interface to execute scrolling operation based on the voice scrolling control word list and a voice control text corresponding to the voice control instruction. Therefore, the voice control of the scroll control is realized, the flexibility and convenience of the control mode of the display equipment are improved, and the user experience is improved.
Owner:HISENSE VISUAL TECH CO LTD

Information processing device, information processing method, and information processing program

This technology uses AI to represent or reproduce characters and provides information that effectively utilizes those characters. [Solution] The information processing device according to the present invention is characterized by comprising: a setting unit that receives settings for advertising information and reading voice from an advertiser; a display control unit that displays reading information based on the advertising information as a message on the terminal device of a user who has added the advertiser's account as a friend in a messenger app; and a voice control unit that outputs voice data of the reading information read aloud with the set reading voice to the user's terminal device and reads the reading information aloud with the reading voice.
Owner:LY CORP

Prompt information output method and system, electronic device, and readable storage medium

Embodiments of the present application provide a prompt information output method and system, an electronic device and a readable storage medium. The method comprises: obtaining first device data of a vehicle-mounted device; obtaining second device data of the vehicle-mounted device; comparing the first device data and the second device data to determine whether the second device data has changed relative to the first device data; if there is a change, determining whether a voice control instruction corresponding to the vehicle-mounted device exists in a voice instruction set; if it exists, determining whether the voice control instruction matches the change data between the first device data and the second device data; and if it matches, generating and outputting a pop-up prompt information corresponding to the vehicle-mounted device based on the voice control instruction. The embodiments of the present application can improve the pertinence of voice broadcast and pop-up prompt, and improve the user experience.
Owner:SHANGHAI PATEO INTERNET TECH SERVICE CO LTD

Voice-controlled oxygenation pump

The utility model discloses a voice control oxygenation pump relates to oxygenation pump technical field, including the casing and set up in the casing, the air pumping assembly includes the air drum subassembly and is used for driving the air drum subassembly to move to produce the electromagnetic drive subassembly of oxygen, is equipped with the circuit board above the air pumping subassembly, and the circuit board is equipped with the air drum subassembly and the electromagnetic drive subassembly. The circuit board is integrated with a control unit, a voice recognition module used for converting voice into electric signals and a wireless communication module used for communicating with external equipment, the control unit comprises a microcontroller and a driving circuit, and the microcontroller outputs control signals to the electromagnetic driving assembly through the driving circuit. One side of the circuit board is electrically connected with a power supply unit, and the surface of the shell is provided with a display unit and a control adjusting part which are electrically connected with the circuit board. According to the oxygen increasing pump, the voice control structure is additionally arranged on the oxygen increasing pump, so that a user can conveniently operate the oxygen increasing pump through voice instructions, the use convenience of the user is improved, and the use experience of the user is greatly improved.
Owner:FUJIAN ANXI JIAHAO OUTDOOR PROD

Voice code scanning all-in-one computer

The utility model relates to the technical field of all-in-one computers, in particular to a voice code scanning all-in-one computer. The voice code scanning all-in-one computer comprises a base, a display screen and a host. The base is mounted on the back of the host and used for supporting the all-in-one computer; the display screen is installed on the front of the host. A control circuit board is arranged in the host, and the display screen is electrically connected with the control circuit board; a label hole is formed in the edge of the display screen, a sound receiving module corresponding to the label hole is arranged in the host, and the sound receiving module is electrically connected with the control circuit board and is used for receiving a voice instruction, so that the control circuit board can execute the instruction according to voice control; a code scanning module is arranged in the host and is electrically connected with the control circuit board; the host is provided with a code scanning port corresponding to the code scanning module, and the code scanning module scans a code through the code scanning port to obtain bar code information.
Owner:SHENZHEN HAILAN ELECTRONICS

Dynamic lamp language control method and system

The invention discloses a dynamic lamp language control method and system, and the method comprises the steps: recognizing a scene where a vehicle is located through a scene recognition module, and generating a lamp language control instruction through a lamp language generation engine according to the recognition result of the scene recognition module; and the LED lamp group ECU receives the lamp language control instruction and controls the working state of each LED particle in the LED lamp group according to the lamp language control instruction. According to the scheme, multi-mode lamp language switching is achieved, the lamp language can be controlled and adjusted according to the fault state, the lamp language output control requirements of the vehicle lamp in different states are met, and the user experience is improved.
Owner:ATECH AUTOMOTIVE WUHU

Intelligent lifting storage cabinet based on voice recognition

The invention belongs to the technical field of storage cabinets, and particularly relates to an intelligent lifting storage cabinet based on voice recognition, the intelligent lifting storage cabinet comprises a base and a storage cabinet body, the upper surface of the base is fixedly connected with a stand column through bolts, one side of the stand column is provided with a T-shaped sliding groove in the length direction of the stand column, and the T-shaped sliding groove is slidably connected with a T-shaped sliding block; the T-shaped sliding block is fixedly connected with two symmetrically-distributed L-shaped connecting plates through bolts, one side of each L-shaped connecting plate is fixedly connected with a supporting plate, and the two storage cabinet bodies are fixed to the corresponding supporting plates. According to the voice-controlled lifting storage cabinet, when the cabinet is used, the cabinet body can be operated to integrally ascend and descend in a voice control mode, manual control is not needed, actual use is facilitated, the cabinet body ascends to release the bottom space when the cabinet body is idle, the cabinet body is moved downwards when objects are stored, and the problems that an existing storage cabinet is inconvenient to ascend and descend and poor in flexibility are solved.
Owner:SHANDONG ZHENGCHENG NEW MATERIAL CO LTD

Fabric-covered electronic device

The present disclosure relates to fabric-covered electronic devices. An electronic device, such as a voice-controlled speaker device, can have a housing characterized by a vertical longitudinal axis. A flexible substrate, such as a flexible mesh substrate having component support regions coupled by flexible sections, can be wrapped around the housing and the vertical axis. The housing can include a surface region having compound curvature. The flexible substrate can conform to the region having compound curvature. A fabric spacer layer can be interposed between the flexible substrate and the housing. Electronic components, such as input-output devices, can be mounted to the component support regions. A display can be formed from an array of light-emitting devices mounted on respective component support regions. Light from the light-emitting devices can pass through the fabric spacer layer toward the housing and exit away from the housing. An outer fabric layer can cover the mesh.
Owner:APPLE INC

Remote device screen projection interaction system

The application discloses a remote device screen projection interaction system, and relates to the technical field of interaction control. The system comprises a remote execution device, a control relay server, a mobile relay terminal and a display terminal. The remote execution device collects and encodes screen audio and video data, establishes a long connection with the control relay server and executes control instructions. The control relay server manages sessions, generates connection information, relays data and instructions. The mobile relay terminal scans a code to establish a two-way connection, collects near-end voice instructions and converts the voice instructions into structured control instructions, and relays audio and video data streams. The display terminal displays connection information and screen projection content. The scheme realizes convenient screen projection and natural interaction across networks through mobile relay and voice control, and improves remote collaboration experience.
Owner:SHANGHAI JIUCHI NETWORK TECH CO LTD

An embedded fuzzy voice control method and system based on language operator quantization

The application relates to an embedded fuzzy voice control method and system based on language operator quantification, and relates to the field of embedded voice control, which comprises the following steps: collecting a voice instruction of a user; performing voice recognition on the voice instruction to generate corresponding text instructions; performing semantic analysis on the text instructions to extract control variables, action directions and language operators; determining a basic fuzzy set based on the control variables and the action directions; determining a deformation operator according to the language operator; combining the deformation operator and the basic fuzzy set to generate an input fuzzy quantity, and collecting a current physical state value; and obtaining an aggregated fuzzy set based on the input fuzzy quantity, the current physical state value and a fuzzy rule library; performing defuzzification calculation on the aggregated fuzzy set to determine a physical control signal value, and driving a physical device to perform corresponding actions according to the physical control signal value. The application has the effect of meeting the actual needs of users for the delicacy of device adjustment.
Owner:NINGBO LADDER EDUCATION TECH CO LTD

Device control method, device control apparatus, device, storage medium, and program product

PendingCN122454975APathPingTarget control
The application relates to a device control method and device, equipment, a storage medium and a program product. The method comprises the following steps: obtaining voice control text information, target running parameters of each candidate Internet of Things device, and target environment state information; determining a dynamic physical capability boundary based on the target running parameters and the target environment state information; based on the voice control text information and the dynamic physical capability boundary, at least one of a candidate action set of a large language model, a tool definition, a generation path, a candidate word element probability distribution or a generation action space is constrained, so that the large language model generates control information conforming to the dynamic physical capability boundary and containing a target control instruction; if the target control instruction and a target Internet of Things device contained in the target control instruction are parsed from the control information, the target control instruction is subjected to a security check; and if the check is passed, the target Internet of Things device executes the target control instruction. The method can improve the accuracy and safety of device control.
Owner:GUANGZHOU ANYKA MICROELECTRONICS CO LTD

Computer with voice-controlled graphical user interface

1. The name of the design product: computer with voice control graphical user interface. 2. The use of the design product: for running programs, displaying information. 3. The design points of the design product: the interface content of the graphical user interface displayed by the computer, the computer is an existing design. 4. The picture or photo that best indicates the design points: front view. 5. The computer is an existing design, omitting the rear view, left view, right view, top view, and bottom view. 6. The use of the graphical user interface: for inputting voice instructions. 7. The human-computer interaction mode of the graphical user interface: click the microphone button in the upper right corner of the front view interface, display the voice instruction input interface on the interface, and present the interface change from the front view to the change state diagram 1; when inputting voice instructions on the change state diagram 1, display the voice instruction content on the voice instruction input interface, and present the interface change from the change state diagram 1 to the change state diagram 2.
Owner:POWERCHINA HUADONG ENG CORP LTD