Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

1476 results about "Gesture recognition" patented technology

Gesture recognition is a topic in computer science and language technology with the goal of interpreting human gestures via mathematical algorithms. Gestures can originate from any bodily motion or state but commonly originate from the face or hand. Current focuses in the field include emotion recognition from face and hand gesture recognition. Users can use simple gestures to control or interact with devices without physically touching them. Many approaches have been made using cameras and computer vision algorithms to interpret sign language. However, the identification and recognition of posture, gait, proxemics, and human behaviors is also the subject of gesture recognition techniques. Gesture recognition can be seen as a way for computers to begin to understand human body language, thus building a richer bridge between machines and humans than primitive text user interfaces or even GUIs (graphical user interfaces), which still limit the majority of input to keyboard and mouse and interact naturally without any mechanical devices. Using the concept of gesture recognition, it is possible to point a finger at this point will move accordingly. This could make conventional input on devices such and even redundant.

Distributed intelligent warehouse scheduling system based on artificial intelligence

The invention discloses a distributed intelligent warehouse scheduling system based on artificial intelligence, and belongs to the technical field of warehouse scheduling. Comprising a multi-source environment sensing module, a dynamic inventory management module, a distributed task scheduling module, an intelligent path planning module, a resource dynamic allocation module, an anomaly detection and emergency response module, an energy consumption optimization module, a supply chain collaboration module and a man-machine interaction and visualization module. A warehouse digital twinborn model is constructed, immersive display of a storage state and a scheduling strategy is realized, an AR scene is superposed through a color coding path, a thermodynamic diagram and a particle flow form, a manager can intuitively master inventory distribution, task progress and an abnormal region, eye movement tracking and a gesture recognition technology support an interactive decision, and the workload of the manager is reduced. The AR marking function can mark an abnormal area and synchronize the abnormal area to a decision making system, and through combination of AR and AI, a brand new interaction normal form is provided for intelligence and humanization of warehouse management.
Owner:SUZHOU SHUHONG INTELLIGENT TECHNOLOGY CO LTD

Multi-sensory autonomous multimodal emotion-synchronized environmental control architecture and regulation system (amesecar)

An autonomous environmental regulation and behavioral monitoring system is disclosed, configured to adapt temperature, lighting, and acoustic conditions based on real-time emotional and physiological data. The system includes a dual-redundant central processor, hierarchical communication networks, multi-angle visual acquisition units, infrared thermometers, and modular environmental subsystems. It detects posture, gestures, facial expressions, and thermal signals to classify user states and apply individualized airflow, light, and sound modulation without relying on external internet connectivity. The system also monitors connected appliances using voltage-based pressure analysis to forecast device degradation. With integrated gesture recognition, privacy-preserving data handling, and predictive adaptation, the invention enables multi-user personalization, long-term learning, and uninterrupted operation within residential, administrative, or healthcare infrastructures.
Owner:SEYEDKHAMOUSHI FAEZEHALSADAT +1

Multi-mode-based AI digital human intelligent interaction method, system and equipment

The invention relates to the technical field of computer vision and human-computer interaction, and discloses an AI digital human intelligent interaction method, system and equipment based on multiple modalities, and the method comprises the steps: pre-awakening a digital human when a human face is detected, and further thoroughly awakening the digital human based on recognized preset voice information or preset gesture information; voice and video information of a user in the interaction process is obtained, a keyword extraction result, a gesture recognition result and an emotional state tag are generated, a pre-constructed knowledge base is utilized to retrieve related information, a big language generation model module is combined to generate an answer text, and the answer text is input into a preset voice synthesis model to generate emotional voice output. And based on the current emotional state label of the user, driving the digital human animation to be output in an emotional manner. According to the method and the system, the digital human for understanding the emotion of the user, generating personalized answers, providing voices with rich emotions and displaying natural expressions and actions can be created, better interaction with the user can be realized, and more humanized and effective services can be provided.
Owner:BEI JING WAN JIE SHU JU KE JI YOU XIAN ZE REN GONG SI WU HAN FEN GONG SI +1

Computer-implemented system and method for providing VR / AR visual experiences to users by pupil-directed retinal projection in near-eye displays

A computer-implemented system and method for pupil-directed retinal projection in near-eye displays are disclosed. The computer-implemented system provides smart glasses with directed physical pixels that project light beams / signals directly onto a user's retina based on pupil position and size. The glasses comprise a frame, lenses with directed pixel layers, and sensors for tracking pupil movement. Each directed pixel may generate multiple virtual pixels by rapidly changing its emission angle. The glasses function as prescription lenses, virtual reality displays, and augmented reality devices without traditional optical systems. Additional features include depth sensors, cameras, and connectivity to peripheral devices. The glasses enable a seamless blend of virtual and real-world experiences, creating an immersive “Mixverse” environment. Various input methods, including gesture recognition and brain-computer interfaces, allow for intuitive control and interaction.
Owner:OSKUI ALI MIZANI

Anti-interference optimized gesture recognition method

The invention relates to the technical field of gesture recognition, in particular to an anti-interference optimized gesture recognition method. Comprising the following steps: acquiring a gesture video stream through a camera, constructing a dynamic background model by using a frame difference method and a Gaussian mixture model, eliminating a static background and interference, and extracting a target area image; performing local brightness histogram analysis on the target region image, and optimizing the image quality by adopting a region adaptive compensation algorithm and a multi-scale edge enhancement technology; positioning a gesture area in real time by using a color histogram and a feature matching algorithm, and dynamically updating a gesture track in combination with Kalman filtering; and extracting gesture shapes, tracks and dynamic mode features through deep learning, comparing the features with a standard model library, and outputting gesture categories and corresponding function instructions. According to the method, a multi-level optimization strategy is adopted for a complex background, a dynamic target and a changeable illumination environment, so that the anti-interference capability and the recognition precision of gesture recognition are improved.
Owner:GUANGZHOU LANGO ELECTRONICS TECH CO LTD

Touch screen control method and device and electronic equipment

The invention relates to the technical field of man-machine interaction, and discloses a touch screen control method and device and electronic equipment. The method comprises the steps that S1, it is detected that a user generates a touch starting event in a screen edge area, touch track data are collected, and the track data comprise touch position coordinates, pressure values and corresponding timestamps; s2, constructing the touch track data into a track function with continuous time, and obtaining speed information, acceleration information and track steering change characteristics of a track; and S3, according to the speed information, the acceleration information and the track steering change characteristics, calculating a group of track dynamic characteristic indexes representing track naturalness. By adopting a combined judgment technology of track naturalness and structural complexity, the naturalness and the complexity of the touch track can be accurately recognized, and a high-precision gesture recognition effect is achieved. Compared with the scheme depending on the touch position and speed in the prior art, the method can effectively distinguish mistaken touch and effective touch, and the problem of frequent misoperation is solved.
Owner:BEIJING HUACAN ELECTRONICS CO LTD

Micro-posture recognition method based on multi-modal feature fusion and fine adjustment

The invention discloses a micro posture recognition method based on multi-modal feature fusion and fine adjustment, and relates to the field of computer vision and action recognition. The method is characterized in that a universal cross-modal knowledge fusion framework is provided, and multi-modal features are respectively extracted through a fine tuning network by using three kinds of modal information of video, skeleton and text. Meanwhile, a video-skeleton and text-skeleton fusion module is introduced to enhance the interactivity between modals. Compared learning is adopted to align feature distribution, and model training is optimized in combination with a freezing-fine tuning strategy, so that the calculation complexity is reduced, and the recognition efficiency is improved. According to the method provided by the invention, the problem of single-mode information loss can be solved, the perception capability of tiny attitude change is enhanced, the recognition precision and robustness are improved, and the method is suitable for multiple application scenes such as behavior monitoring, human-computer interaction and safety protection.
Owner:CHINA UNIV OF PETROLEUM (EAST CHINA)

Real-time interactive image generation system based on multi-point touch canvas

The invention relates to the technical field of computer graphic interactive processing, in particular to a real-time interactive image generation system based on a multi-point touch canvas. The input acquisition unit is used for acquiring original touch data from an operating system and preprocessing the original touch data; the gesture recognition and analysis unit is used for performing high-level semantic behavior analysis on the contact data sequence processed by the preprocessing module; the interaction control and parameter mapping unit is used for receiving the semantic event output by the gesture recognition and analysis unit, analyzing the semantic event into an image control command and generating an executable command sequence; and the image generation unit generates interactive image content dynamically responded in real time based on an internal graph state management mechanism. By introducing the adaptive Kalman filtering and trajectory prediction auxiliary mechanism, the filtering intensity of the contact data can be dynamically adjusted, the efficient suppression of finger jitter and the consistent reconstruction of the contact ID are realized, and the input stability and data continuity under the multi-point touch operation are remarkably improved.
Owner:HUNAN VOCATIONAL COLLEGE OF SCI & TECH

Virtual reality interaction and content generation method and system based on large model

The invention provides a virtual reality interaction and content generation method and system based on a large model. According to the method, original point cloud data are collected to be aligned with semantic tags, and a topological graph containing object positions, categories and spatial constraints is constructed; dynamically calibrating the point cloud and mapping the point cloud to a unified coordinate system, and extracting a shielding boundary to generate a parallax mapping table; recognizing a blank area delimited by a user through a gesture, and generating a new object parameter candidate set of physical compliance in combination with the topological graph; screening parameters by using a large model, and reserving physical effective parameters after collision response, gravity and illumination verification; and generating the geometry and texture of the virtual object in the occlusion area based on the incremental neural radiation field, optimizing the depth and the rendering precision in combination with the mapping table, outputting the depth and the rendering precision to an interactive interface, and updating topological parameters. According to the method and the device, a new object which is physically consistent and rendered and optimized is dynamically generated based on user interaction in a virtual reality scene, and the sense of reality and interaction experience are improved.
Owner:LUSTER LIGHTWAVE CO LTD

VR interactive new energy automobile high-voltage electrical principle teaching method, system and equipment

The invention belongs to the technical field of new energy automobile teaching, and particularly relates to a VR interactive new energy automobile high-voltage electrical principle teaching method, system and equipment, and the method comprises the steps: obtaining the physical parameter data and dynamic circuit simulation data of a new energy automobile high-voltage electrical system; constructing a high-precision 3D model according to the physical parameter data, inputting dynamic circuit simulation data into the model, and generating interactive principle annotations and particle effect animations; based on a gesture recognition model and a new energy automobile VR interaction technology, circuit state changes in the particle effect animation are obtained, a new energy automobile high-voltage system principle and a signal flow are displayed, a practical operation simulation training scene is constructed, student operation data are recorded according to the practical operation simulation training scene, an AI evaluation model is trained, and an evaluation report is obtained. And adjusting interaction feedback parameters of the training scene, and meanwhile, generating a personalized learning report of the trainee. Therefore, the problems of safety risk, teaching lag, abstraction obstacle and the like in the prior art are solved.
Owner:MINGZHEN INTELLIGENT TECH (BEIJING) CO LTD

Robot sign language communication method, related device and storage medium

The invention discloses a robot sign language communication method, a related device and a storage medium. The method comprises the steps of collecting each frame of gesture image of a current user and performing feature extraction; inputting the features of each frame of gesture image into a trained gesture recognition model, and recognizing a current gesture recognition result; based on the mapping relation between the gesture symbols and vocabularies, mapping the current gesture recognition result to obtain a vocabulary sequence; forming the vocabulary sequence into a current complete recognition text according to a preset grammar rule; correcting the current complete recognition text by combining the multi-modal large language model with the current complete recognition text, the current scene understanding information and the current action understanding information; inputting the corrected current complete recognition text, the current scene picture and the historical communication information into a third visual language model, and analyzing a current response text; converting the current response text into a current sign language sequence; and driving the robot to simulate sign language actions according to the current sign language word order.
Owner:DIGITAL HUAXIA (SHENZHEN) TECHNOLOGY CO LTD

Intelligent control faucet and faucet body

The invention relates to the technical field of electromagnetic valves, in particular to an intelligent control faucet and a faucet body, the faucet carries a control system, and the system comprises an environment interference suppression module, a gesture recognition processing module, an electromagnetic hysteresis compensation module, a multi-parameter coupling optimization module and a residual magnetism elimination control module. According to the method, the infrared compensation proportion is dynamically adjusted by fusing the environment light intensity and the hand position information, the interference suppression capability is improved, the recognition stability under complex illumination is enhanced, the action consistency is extracted in combination with the gesture track waveform and the dynamic response feature, and the recognition accuracy is improved; electromagnetic hysteresis is dynamically corrected based on real-time current and displacement deviation combined with a temperature factor, motion feedback precision is improved, multi-parameter coupling is fused with multi-physical field data to construct a stable control reference, system sensitivity is enhanced, a reverse pulse sequence is designed by capturing a current attenuation waveform, residual magnetism interference is effectively eliminated, and system performance is improved. And multi-dimensional improvement of identification precision, control stability and response efficiency is integrally realized.
Owner:XIAMEN SANCHANG SANITARY WARE TECH CO LTD

Intelligent remote controller control method, device and equipment based on gesture recognition

The invention provides an intelligent remote controller control method, device and equipment based on gesture recognition, and the method comprises the steps: obtaining multi-modal sensor data through an integrated time of flight (ToF) camera, a millimeter wave radar and an inertial measurement unit (IMU), carrying out the feature extraction to generate gesture spatial-temporal features, and carrying out the dimension exchange of the spatial-temporal features to generate exchange dimension features; constructing a candidate gesture probability superposition state based on the exchange dimension features, establishing a semantic negative hypothesis and generating a reverse verification parameter; performing weight correction on the probability superposition state by using the inversion verification parameter to generate a correction probability state, and determining a target gesture through probability aggregation; re-extracting verification feature data based on the target gesture, selecting a grading classification model according to a preset grading condition to perform secondary verification, and outputting a gesture recognition result; and detecting a communication environment state, constructing a satellite flash and UWB dual-mode communication race state, aggregating into an optimal communication mode, and completing intelligent control of the remote controller.
Owner:WUXI WEIDA INTELLIGENT ELECTRONICS CO LTD

Centimeter-level VR real-scene real-time synchronous rendering method and system for multi-source data fusion and physical field coupling

The invention relates to the technical field of virtual reality, in particular to a centimeter-level VR real-scene real-time synchronous rendering method for multi-source data fusion and physical field coupling. Comprising the following steps: eliminating cloud layer interference and expanding Kalman filtering to improve positioning precision through multi-source acquisition of satellite remote sensing data, meteorological data and an equipment sensor; dynamically adjusting process noise by using Kalman filtering, and complementing delay data; constructing a ray tracing and U-Net noise reduction model, carrying out hierarchical loading of textures in combination with dynamic texture management, modeling dynamic illumination and wind field simulation based on a solar declination angle, and realizing scene rendering; a virtual ground is aligned through WebXR, gesture recognition and a force feedback mechanism are integrated, and fingertip curvature and movement speed features are extracted in combination with a Leap Motion sensor. According to the invention, through multi-source data fusion, efficient rendering and accurate interactive design, the authenticity and real-time synchronization capability of the VR scene are significantly improved.
Owner:张雨

Gesture recognition method and device based on deep learning

The invention discloses a gesture recognition method and device based on deep learning, and the method comprises the steps: collecting a video stream, and obtaining a hand key point data set in the video stream; performing data enhancement and preprocessing on the hand key point data set; establishing a gesture recognition model based on deep learning, and performing training optimization; and carrying out lightweight processing on the trained gesture recognition model, and outputting a gesture recognition result in real time based on the lightweight gesture recognition model. According to the method, through pre-segmentation of the video image, a progressive polynomial attenuation pruning strategy, collaborative optimization of pruning and quantification and a comprehensive callback mechanism, real-time, accurate and efficient operation of a gesture recognition model on mobile equipment and an embedded system is achieved, and the high-standard requirement of a modern intelligent interaction system is met.
Owner:ANHUI UNIV

Intelligent household electrical appliance interaction control system based on embedded software

The invention discloses an intelligent household electrical appliance interaction control system based on embedded software. A system operation process specifically comprises the following steps: acquiring a user input signal; analyzing the user input signal based on an embedded software architecture to generate a standardized control instruction; establishing communication connection with a plurality of intelligent household electrical appliances through an Internet of Things protocol, and obtaining real-time environment data and real-time operation states of the intelligent household electrical appliances; performing conjoint analysis on the user historical behavior data and the user input signal; adjusting operation parameters of the intelligent household electrical appliance; and returning interaction response information to the user through at least one mode of a voice feedback module, a touch interface, a gesture recognition interface or a mobile phone APP. The method has the following advantages and effects: multi-mode instructions can be adaptively fused, the equipment capability and the user intention can be dynamically matched, and cross-equipment and cross-protocol efficient and accurate control is realized through standardized instruction generation, real-time load sensing and priority dynamic calculation.
Owner:SHENZHEN XINYINGDA TECH CO LTD

Real-time dynamic trajectory tracking method and system for millimeter wave radar gesture recognition

The invention discloses a real-time dynamic trajectory tracking method and system for millimeter wave radar gesture recognition, and relates to the technical field of gesture recognition tracking, and the method comprises the steps: receiving an echo signal reflected by a gesture, extracting a potential target point cloud, and carrying out the clustering generation of a gesture point cloud sequence; establishing a multi-modal motion model library, dynamically selecting an optimal motion model by adopting graph matching, and generating a prediction state in combination with a gesture point cloud sequence; on the basis of a Poisson multi-Bernoulli hybrid filtering framework, according to the signal-to-noise ratio and spatial distribution of the current gesture point cloud sequence, dynamically adjusting the observation weight, optimizing the observation point cloud, and carrying out optimal association by combining Mahalanobis distance with dynamic time warping; a multi-hypothesis tracking strategy is adopted to maintain trajectory hypothesis, an optimal trajectory is selected through a trajectory scoring mechanism, and Kalman filtering smoothing processing is performed on the optimal trajectory. According to the method, high-precision and low-delay tracking of gesture motion is realized, gesture habits of different users and complex environment interference can be adapted, and meanwhile, relatively high track precision is kept.
Owner:SHENZHEN YUNENG WIRELESS TECH CO LTD

High-precision multi-modal twin model assembly method based on large language model and multiple docking optimization mechanism

The invention discloses a high-precision multi-modal twin model assembly method based on a large language model and a multiple docking optimization mechanism. The whole system framework is composed of three modules: a database, an interaction system and human factors. Wherein the database module integrates a model library module, a user interaction data recording module and a temporary data recording module, and is used for storing and managing various data information required for constructing a virtual scene; the interactive system module is used as a core part, integrates a model dynamic loading module, a model accurate positioning module and a model adaptive assembly module, realizes instruction analysis and model generation by relying on a large language model, and combines a multi-docking optimization mechanism of eye movement tracking, gesture recognition, automatic docking and logic rule binding to realize multi-docking. Efficient and accurate interaction between the user and the database is effectively guaranteed; the method not only provides a solid foundation for virtual experiments and virtual-real symbiosis, but also significantly improves the construction efficiency and precision of digital twin model assembly.
Owner:SOUTHEAST UNIV

Method and system for preventing pressing plate of transformer substation control screen cabinet from being touched by mistake

The invention discloses a transformer substation control screen cabinet pressing plate mistaken touch prevention method and system, and the method comprises the steps: the system automatically detects the approaching of the hand of an operator through infrared induction and gesture recognition, and judges whether the recognition operation is effective or not; according to environmental changes such as high humidity or temperature, the system automatically adjusts the sensitivity of the touch screen to optimize the operation experience; before key operation, the system starts a multiple confirmation mechanism, and an operator needs to confirm an operation intention through a touch screen button and voice recognition; for high-risk operation, the system pops up an alarm and carries out secondary confirmation; meanwhile, the system dynamically adjusts the authority according to the identity of the operator and the task requirement, and the low-authority operator is limited to access the key component and only can execute the operation within the authority range of the low-authority operator; according to the method and the system for preventing the pressing plate of the transformer substation control screen cabinet from being touched mistakenly, the accuracy and the flexibility of preventing mistakenly touching are effectively improved, and the defects in the prior art are overcome through an intelligent means.
Owner:GUANGZHOU KAJUN MASCH EQUIP CO LTD

Skeleton sign language recognition method of double-flow space-time dynamic graph convolutional network fused with residual learning

The invention discloses a skeleton sign language recognition method of a double-flow space-time dynamic graph convolutional network fused with residual learning, and belongs to the technical field of artificial intelligence and gesture recognition. According to the method, an input gesture skeleton sequence relative to a face is divided into two data streams, namely a hand form data stream and a wrist track data stream through double-reference-system differential homeomorphic mapping; the method comprises the following steps: firstly, processing hand posture data, and capturing a hand joint spatial topological relation by combining a spatial-temporal dynamic graph convolutional network (STDGCNN) with a residual convolutional block; meanwhile, a Finsler trajectory dynamics encoder (FTDE) is adopted to carry out differential geometric modeling on the wrist trajectory, and the direction sensitivity characteristic of the trajectory is captured through a multi-scale causal convolutional network. Then, mutual enhancement of double-flow features is realized through a bidirectional cross feature enhancer (BCFE), and the problem of geometric inconsistency of a heterogeneous feature space is solved through a geometric-driven optimal transmission fusion device (Geometric-OT). The method solves the challenge that a traditional sign language recognition method processes complex space-time correlation of gesture forms and motion tracks at the same time, the technical problems of insufficient relation between hands and faces, insufficient feature expression ability and low space-time feature extraction efficiency, and the problems of geometric inconsistency, single reference system and the like. And the identification accuracy and the real-time performance are obviously improved. Experiments show that the accuracy rate of the method in complex hundreds of sign language vocabulary recognition tasks reaches 95% or above on average, the reasoning speed is only 17ms on average, high-precision real-time sign language recognition is achieved, and the method has higher robustness in complex environments such as noise and shielding.
Owner:刘良锦

SEMG gesture recognition method based on lightweight deep separable residual attention network

The invention discloses an sEMG gesture recognition method based on a lightweight deep separable residual attention network, and relates to the technical field of intelligent recognition. According to the invention, DSRANet and a lightweight deep separable residual attention network are provided, so that the spatial-temporal characteristics of sEMG gestures can be effectively captured while the calculation efficiency (0.459 M parameter, 0.1 G FLOPs) is maintained; a DSR block is designed to serve as an innovative architecture component, through combination of depth separable convolution and residual connection and efficient extraction of discriminative features, an ablation experiment shows that the calculation complexity is reduced by 45% compared with that of standard convolution, meanwhile, an MACA mechanism is introduced, key features are enhanced in a self-adaptive cross-channel and time-frequency dimension mode, and a contrast experiment verifies that the model precision can be improved by 6.41%.
Owner:CHONGQING UNIV OF TECH

Diagnosis and treatment assisting system and method based on voice gesture interaction technology

The invention belongs to the technical field of man-machine interaction, and discloses a diagnosis and treatment auxiliary system and method based on a voice gesture interaction technology, and the system comprises a voice recognition module which is used for collecting voice information; converting the voice information into text information; inputting the text information into the large language model to obtain a voice system instruction, and sending the voice system instruction to the data processing and interaction module; the gesture recognition module is used for collecting diagnosis and treatment scene information; identifying a preset operation gesture in the diagnosis and treatment scene information; mapping a preset operation gesture into a gesture system instruction and sending the gesture system instruction to the data processing and interaction module; the data processing and interaction module is used for receiving and executing a voice system instruction or a gesture system instruction; generating a calling instruction according to the voice system instruction or the gesture system instruction and sending the calling instruction to the network communication module; converting the received medical record data into diagnosis and treatment data in a display format; the network communication module is connected with a hospital information system; the calling module is used for calling medical record data after receiving a calling instruction of the data processing and interaction module and sending the medical record data to the data processing and interaction module; the display module is used for displaying the diagnosis and treatment data. The problem that a doctor cannot conveniently operate a computer to check patient data due to the sterility principle in the diagnosis and treatment process is solved, the diagnosis and treatment efficiency is improved, infection control is facilitated, and the sterility principle is met.
Owner:BEIJING STOMATOLOGY HOSPITAL CAPITAL MEDICAL UNIV

Cursor mode switching

Methods and systems for processing input from an image-capture device for gesture-recognition. The method further includes computationally interpreting user gestures in accordance with a first mode of operation; analyzing the path of movement of an object to determine an intent of a user to change modes of operation; and, upon determining an intent of the user to change modes of operation, subsequently interpreting user gestures in accordance with the second mode of operation.
Owner:SIM IP HXR LLC

Multi-mode virtual-real fusion interaction method and system

The invention provides a multi-mode virtual-real fusion interaction method and system, and particularly relates to the technical field of man-machine interaction. According to the method, hand three-dimensional coordinate data are collected through a depth perception camera, voice data are collected through a microphone array, the hand data are input into a gesture recognition model to output operation intention parameters, real-time state data of virtual equipment are obtained according to the operation intention parameters, tactile feedback parameters are generated, and the real-time state data of the virtual equipment are obtained after linear mapping association. And generating a vibration waveform and outputting to the touch execution unit. And meanwhile, analyzing the voice data to obtain a voice instruction, and performing space-time alignment on the voice instruction and the operation intention parameter to generate a virtual-real synchronous control instruction. According to the method, the technical problems of how to realize touch feedback of linkage of gesture operation and the equipment state through a low-cost and low-delay natural interaction mode and how to improve the multi-mode instruction cooperation capability so as to reduce the misoperation rate are solved, and accurate touch feedback of linkage of gesture operation and the virtual equipment state is realized; and the multi-mode instruction cooperation capability is improved, and the misoperation rate is remarkably reduced.
Owner:国网重庆市电力公司潼南供电分公司

Systems and methods for person classification and gesture recognition

A method includes receiving image data that includes at least two images of an environment associated with a vehicle, identifying at least one person of interest in the image data, and generating, using a pose estimation model and the image data, a representation of the person of interest. The method also includes determining at least one characteristic associated with the at least two images of the image data and providing, to a machine learning model, at least the representation of the person of interest and the at least one characteristic associated with the at least two images of the image data. The method also includes receiving, from the machine learning model, a gesture prediction indicating a predicted gesture being made by the person of interest, and causing the vehicle to take at least one action based on the gesture prediction.
Owner:VALEO SCHALTER & SENSOREN GMBH

Face posture recognition method, anti-dazzle lamp regulation and control method, lamp, equipment and medium

The invention discloses a face posture recognition method, an anti-dazzle lamp regulation and control method, a lamp, equipment and a medium, and relates to the technical field of lamp control. According to the method, firstly, face region recognition is performed based on the integral image corresponding to the video image, then the face image is intercepted from the video image according to the recognition result, and then face pose recognition is performed based on the trained CNN model and the trained backbone network model, so that recognition of each face pose in the target scene is realized, and the recognition efficiency is improved. And then the irradiation angle and brightness of the lamp are adjusted based on the recognized face posture, and adaptive adjustment of the glare problem is achieved.
Owner:CHONGQING UNIV

Multi-modal classroom teaching optimization system integrating space intelligent management and somatosensory annotation

The invention relates to the technical field, in particular to a multi-mode classroom teaching optimization system integrating space intelligent management and somatosensory annotation. According to the technical scheme, the system comprises a space intelligent management module which dynamically adjusts the environment through a genetic algorithm to enable the comfort level index to be optimal, a somatosensory annotation interaction module which is used for recognizing gesture tracks of teachers and students and automatically controlling teaching equipment according to the intention of a user, and a multi-modal data fusion module which is used for carrying out data fusion on the teaching equipment according to the intent of the user. The module improves the accuracy of data analysis and the robustness of system interaction, and the optimization scheduling and resource allocation module optimizes teaching resource allocation based on a Lagrange multiplier method and intelligently schedules resources. Through intelligent environment management, multi-modal data fusion, high-precision gesture recognition and an optimization algorithm, the intelligent level of classroom teaching is improved, and more advanced technical support is provided for future intelligent education.
Owner:NANJING CASSO SYST ENG CO LTD

Operator and occupant monitoring validation for autonomous and semi autonomous machines

In various examples, one or more validity checks that model one or more aspects of human physiology may be applied to frames of detected human features to detect and respond to the presence of faults. Example validity checks include human feature constraints derived from the kinematics of human motion, anatomical and spatial constraints, consistency across detection modalities, and / or others. The present techniques may be utilized to validate human features detected by various computer vision tasks, such as those involving pose estimation, facial detection, gesture recognition, and / or activity monitoring, to name a few examples.
Owner:NVIDIA CORP

General electromyographic signal processing method and system based on large self-supervised model

The invention discloses a general electromyographic signal processing method and system based on a large self-supervised model. The general electromyographic signal processing method comprises the following steps: step 1, acquiring a multi-source original multi-electrode channel EMG signal X from an electromyographic acquisition device; and finally, performing data unification processing, and finally converting into a space-time activity diagram with a fixed size of 224 * 224. On the basis of the space-time activity diagram and the fatigue state mark, constructing an AEMG for training according to heterogeneous unlabeled EMG data collected by a collection device; performing light-weight Adapter layer fine adjustment on the pre-trained large myoelectricity model to adapt to gesture recognition muscle force regression gait analysis or rehabilitation evaluation downstream tasks; aiming at the problem that the dimension and the structure of myoelectricity data are not matched due to different acquisition devices, acquisition parts and acquisition tasks, original signals are converted into space-time activity diagrams in a unified format through data unification processing, device differences are represented by combining a sensor embedding module, effective alignment of cross-source data is achieved, and the accuracy of the data is improved. And a basis is provided for large-scale data utilization.
Owner:SOUTH CHINA UNIV OF TECH

Method and system for intelligently capturing air gesture actions and processing interactive data

The invention relates to the technical field of gesture recognition, in particular to an air gesture action intelligent capture and interactive data processing method and system.According to the air gesture action intelligent capture and interactive data processing method and system, significant jumps are marked through timestamps, time sequence deconstruction of action streams is achieved, nonlinear alignment processing of time structures between paragraphs is introduced, and the time sequence deconstruction efficiency is improved. The fragment consistency recognition capability under the conditions of unequal action durations and different rhythms is remarkably improved, spatial topology among three points is modeled by adopting a graph structure, a graph convolutional network is introduced to process dynamic changes of connection relations among different nodes, the continuity and stability of structure judgment can be kept under a shielding or structure deformation scene, and the recognition efficiency is improved. Through a triple constraint mechanism of direction vector included angle amplitude, length ratio and slope fluctuation, the action is subjected to modeling support in three aspects of time dimension, space structure and direction continuity at the same time, and the dynamic stability of gesture recognition and the resolution precision of paragraph judgment are enhanced.
Owner:BAIGE ONLINE (XIAMEN) DIGITAL TECHNOLOGY CO LTD