Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

21674results about "Mechanical pattern convertion" patented technology

System and method for ai-driven multi-modal content generation and immersive interaction experiences

A system and method for creating complex, immersive, and interactive digital content is disclosed. The system integrates advanced artificial intelligence, multi-modal input processing, cloud-based shared environments, and immersive hardware to generate, optimize, and deliver rich interactive experiences. The platform supports content mashups, custom scenario generation, and adaptive AI behaviors, enabling the creation of unique and engaging digital environments across various media formats.
Owner:QOMPLX INC

Digital twinborn enabling intelligent pump station preventive operation and maintenance system

The invention discloses a digital twin enabling intelligent pump station preventive operation and maintenance system. Comprising a dynamic twin construction module, a multi-source heterogeneous multi-modal data acquisition module, an edge computing and cloud collaboration module, an equipment health degree evaluation module, a predictive maintenance decision module, a cross-system data fusion module, a self-evolution knowledge graph module, an intelligent diagnosis and early warning module, a self-adaptive maintenance decision module and a man-machine collaboration interaction module. And the dynamic twin construction module comprises a physical-virtual synchronous calibration mechanism and an equipment degradation parameter dynamic updating mechanism. According to the method, the limitation problem of a traditional static model is solved, the method can adapt to nonlinear changes under complex working conditions, the comprehensive judgment and prediction capability of the system on the equipment state can be enhanced, the energy utilization efficiency is improved, the energy consumption is reduced, the decision and verification mechanism is perfected, and the data acquisition and processing problem is improved; the problems of timeliness and flexibility of the model are solved, and the computing architecture and the response capability are optimized.
Owner:哈尔滨凯纳科技股份有限公司

Methods for navigating user interfaces

In some embodiments, an electronic device navigates between user interfaces based at least on detecting a gaze of the user. In some embodiments, an electronic device enhances interactions with control elements of user interfaces. In some embodiments, an electronic device scrolls representations of categories and subcategories in a coordinated manner. In some embodiments, an electronic device navigates back from user interfaces having different levels of immersion in different ways.
Owner:APPLE INC

Methods for sharing content and interacting with physical devices in a three-dimensional environment

In some embodiments, a computer system displays a user interface element corresponding to content shared with the computer system for interacting with the shared content in the three-dimensional environment. In some embodiments, a computer system displays one or more virtual control elements in a three-dimensional environment that are selectable to cause one or more corresponding operations involving a physical device to be performed.
Owner:APPLE INC

Community intelligent monitoring and emergency linkage method and system fusing BIM spatial semantics

The invention discloses a community intelligent monitoring and emergency linkage method and system fusing BIM spatial semantics, and the method comprises the steps: constructing a BIM scene map, and obtaining the attributes and mutual relationships of components and spatial regions in a BIM model; mapping a dynamic target detected in video monitoring into the BIM model, and obtaining spatial semantic information of the dynamic target; based on BIM spatial semantic information of a dynamic target, a target-environment interaction graph is constructed, a graph neural network model is used for training and reasoning, and specific complex events related to spatial contexts are recognized; taking the BIM model as a space-time reference, fusing multi-source heterogeneous data, and reconstructing by adopting a graph-based event association algorithm to form a complete event chain containing an atomic event sequence and an association relationship; and when an emergency event or an event chain is detected to indicate an emergency state, combining BIM preset information and real-time sensor data, dynamically generating an optimal emergency plan, and performing visual commanding and dispatching through a BIM three-dimensional scene and augmented reality.
Owner:ZHEJIANG LEISHENG CONSTRUCTION ENGINEERING CO LTD

Peripheral nerve injury personalized rehabilitation system and method based on multi-modal large model

The invention relates to the technical field of artificial intelligence assisted medical rehabilitation, in particular to a peripheral nerve injury personalized rehabilitation system and method based on a multi-modal large model, and the method comprises the steps: a feature fusion module employs space-time attention to fuse multi-modal time series data, and constructs an evaluation map; the personalized generation module is combined with historical data and reinforcement learning to generate a scheme containing virtual scene parameters; the interaction feedback module collects data through mixed reality and calculates action deviation; and the adaptive adjustment module adopts a meta-learning optimization model and a distributed iterative output scheme. According to the method, the cross-modal association precision of the motion features and the mechanical parameters is improved, dynamic matching of the training scene and the motion ability of the user is achieved, the virtual environment and the mechanical feedback threshold are optimized by dynamically adjusting the rehabilitation scheme parameters, the scheme optimization period is shortened based on an online iterative optimization mechanism, and the training efficiency is improved. The core defects of personalized adaptation lagging and low utilization efficiency of multi-modal data are overcome.
Owner:FIRST HOSPITAL AFFILIATED TO GENERAL HOSPITAL OF PLA

Robot anthropomorphic interaction method based on multi-modal emotion recognition and customized portrait generation

The invention discloses a robot anthropomorphic interaction method based on multi-modal emotion recognition and customized portrait generation. The method comprises the following steps: S1, dynamically fusing multi-modal emotions; the method comprises the following steps: S1, synchronously acquiring voice, visual and text signals through a multi-source heterogeneous sensor, capturing a user voice stream by a high-fidelity microphone array, and extracting acoustic characteristics such as intonation and speed, S2, performing cross-modal reasoning; s3, synchronously generating contents; step S4: style migration; step S5, anthropomorphic voice and expression generation; according to the method, man-machine interaction emotion is analyzed and generated by utilizing a large language model and multi-modal information fusion, the singleness of interaction emotion and the deficiency of emotional sharing ability are avoided, a strong emotion interaction characteristic is achieved, the image of the robot is obtained through a generative technology and can be migrated to any image, the limitation that a specific image is independently made is broken through, and the interaction effect of the robot is improved. The advantage that one robot can be suitable for different scenes is achieved.
Owner:JIANGSU YUNMU ZHIZAO TECH CO LTD

AR navigation system and method based on visual language model

The invention discloses an AR navigation system and method based on a visual language model, and the method comprises the steps: firstly carrying out video data collection, and constructing a memory database; secondly, querying a navigation target most related to a natural language request of a user in the constructed memory database to obtain a current AR equipment pose and a target pose; and then solving the shortest path from the current pose to the target pose by using the current AR equipment pose and the target pose according to the point cloud map, and optimizing the path direction to finally obtain an optimized path. And finally, guiding a user to move along the planned path through view superposition path indication and voice prompt in AR equipment by utilizing the optimized path, and updating the point cloud map and the memory database. According to the invention, high-precision real-time positioning and sparse point cloud map construction can be realized only by camera input in indoor and outdoor complex environments with low GPS precision, accurate navigation is carried out, and the flexibility and intelligent level of navigation are improved.
Owner:HANGZHOU DIANZI UNIV +1

Methods for manipulating objects in an environment

In some embodiments, while displaying a content entry user interface element and while a first content entry tool is selected, an electronic device detects a first movement of a predefined portion of a user while the predefined portion has a first shape. In some embodiments, in response to detecting the first movement, in accordance with a determination that a gaze of the user was directed toward the content entry user interface element when the first movement was detected, the electronic device enters first content, based on the first content entry tool, corresponding to the first movement into the content entry user interface element. In some embodiments, in accordance with a determination that the gaze of the user was directed toward a menu user interface element when the first movement was detected, the electronic device selects a second content entry tool for entering content into the content entry user interface element.
Owner:APPLE INC

Flight training evaluation system fusing electroencephalogram characteristics and physiological indexes

The invention relates to the technical field of flight training evaluation, and discloses an electroencephalogram feature and physiological index fused flight training evaluation system. The system comprises a physiological signal acquisition module which synchronously captures multichannel electroencephalogram original signals and body surface physiological index data, and the body surface physiological index data comprises an electrocardiograph R-R interval sequence, respiratory wave frequency amplitude and galvanic skin response amplitude; the multi-modal fusion module is used for analyzing an electrocardiograph R-R interval sequence to generate a heart rate variability feature vector and establishing dynamic association mapping of an electroencephalogram entropy value and a physiological feature vector; the cognitive state modeling module is used for generating a cognitive load index according to the dynamic association mapping and constructing a cognitive stability quantization matrix; the self-adaptive feedback module is used for receiving related data and dynamically adjusting simulated flight scene parameters; and the evaluation output module is used for integrating the data to generate a comprehensive training evaluation report containing a neurophysiological coordination degree score and an operation accuracy rating. According to the system, comprehensive evaluation and dynamic training adjustment of the cognitive state of the pilot are realized.
Owner:BEIJING AEROSPACE HUATENG TECH CO LTD

Four-foot robot mechanical arm tail end force feedback teleoperation control system and method

The invention belongs to the technical field of robot control, particularly provides a force feedback teleoperation control system and method for the tail end of a mechanical arm of a quadruped robot, and aims at the key challenges that control errors are caused by communication time delay and soft obstacles are difficult to recognize in a dynamic environment. Modeling and judgment are conducted on the contact state of the tail end of the mechanical arm in advance, and feedforward control and buffer adjustment oriented to communication time delay are achieved. And meanwhile, a dynamic semantic map is constructed in combination with multi-source sensing information, and soft obstacle reasoning and path optimization are performed by fusing a tail end force sense change trend, so that the recognition and avoidance capabilities of the system in a complex and invisible obstacle environment are remarkably improved. The system has good perspectiveness, self-adaptability and high redundancy safety characteristics, is suitable for multi-task inspection operation of industrial sites such as a thermal power plant, and is especially suitable for a remote man-machine cooperative operation scene in a narrow space.
Owner:武汉跨克信息技术有限公司

Method of displaying user interfaces in an environment and corresponding electronic device and computer readable storage medium

Methods for displaying user interfaces in a computer-generated environment provide for an efficient and intuitive user experience. In some embodiments, user interfaces can have different immersion levels. In some embodiments, a user interface can have a respective immersion level based on its location in the three-dimensional environment or distance from the user. In some embodiments, a user interface can have a respective immersion level based on the state of the user interface. In some embodiments, a user interface can switch from one immersion level to another in response to the user's interaction with the user interface.
Owner:APPLE INC

Three-dimensional monitoring system of precision servo press based on digital twinning

The invention relates to the technical field of press monitoring, in particular to a digital twinning-based three-dimensional monitoring system for a precision servo press, which comprises a physical layer sensing module for acquiring real-time operating parameters, environment variables and workpiece processing data of the press; the dynamic twin construction module constructs a total-factor digital twin, and simulates a force-heat-deformation coupling effect by using finite element analysis and a multi-body dynamics algorithm based on physical attributes and process parameters; the intelligent analysis center identifies a potential fault mode of the press machine and locates an abnormal source through multi-physics field simulation data in combination with an improved CNN-LSTM model; the three-dimensional visual interaction unit constructs an interactive immersive three-dimensional virtual scene, renders a running state and a processing process in real time, and generates a maintenance strategy; and the self-adaptive regulation and control unit predicts the residual life of the key component and dynamically adjusts parameters according to a maintenance strategy and real-time monitoring data. Therefore, the problems of single monitoring dimension, disjunction of maintenance strategies and the like in the prior art are solved.
Owner:XIANGSHAN YIDUAN PRECISION MACHINERY CO LTD

Hoisting construction safety monitoring and early warning system based on BIM

The invention discloses a BIM (Building Information Modeling)-based hoisting construction safety monitoring and early warning system. The system comprises a terminal sensing layer which is used for collecting environmental parameters and personnel behavior data in a closed space in real time; the edge computing layer is used for carrying out cleaning, compression and encrypted transmission on original data by utilizing an explosion-proof edge computing gateway; the cloud collaboration layer is used for storing full data based on a BIM digital twinborn platform, constructing a'danger mode-construction feature-disposal measure 'three-dimensional meta-knowledge graph by adopting an MAML + + algorithm, meanwhile, coupling a physical mechanism data enhancement engine with a multi-physics field coupling model and a physical constraint generative adversarial network, generating virtual data conforming to mass conservation and energy conservation, and sending the virtual data to the cloud collaboration layer; performing mixed training with real data; according to the intelligent decision-making layer, a space-time adaptive threshold evolutionary algorithm encodes a space-time context through a graph attention network and Transform, an alarm threshold is dynamically optimized through deep reinforcement learning, meanwhile, a digital twin deduction engine calculates a shortest safety path in real time, and rescue resource allocation is optimized.
Owner:POWERCHINA HUADONG ENG CORP LTD

Intelligent glasses AI voice interaction method

The invention provides a smart glasses AI voice interaction method, which comprises the following steps: simultaneously acquiring a user gesture image and a voice signal, acquiring a user historical interaction record, extracting key point coordinates and pointing direction information of the gesture image, acquiring hand distance information, generating a gesture motion path, simultaneously extracting frequency spectrum information and an intonation peak value of the voice signal, and generating a gesture motion path; forming a voice beat sequence; extracting a spatial semantic mode of the gesture and voice fusion data, and recognizing a core object of a user pointing instruction according to pointing coordinates of a gesture motion path and an intonation peak value of a voice beat sequence; after the core object pointing to the instruction is recognized, the user intention is determined in combination with the spatial semantic mode and the historical interaction record of the user; and the complete intention analysis result is output to the intelligent glasses display module to execute corresponding operation, and is fed back to the acquisition module to adjust the next capture parameters including the acquisition frequency and the recognition sensitivity, so that the response speed and the accuracy are improved.
Owner:SHENZHEN YAWELL LNTELLIGENT TECH CO LTD

An AR home experience method in a large scene

The invention discloses an AR home experience method in a large scene. On the basis of combination of a natural feature identification-based three-dimensional registration method and a binocular tracking positioning and local mapping method, the camera attitude is estimated by using feature points of a real-time scene and corresponding three-dimensional points thereof under the binocular trackingpositioning and local map construction technology. According to the mode, on-site environment features shot in real time are used as recognition tracking objects, a virtual home model can still be normally positioned and tracked under the condition that no identification graph exists, the problems that an existing AR home experience application is small in use range and poor in stability are solved, and therefore the AR home experience of virtual and real fusion can be met in a wider range and more truly.
Owner:MAANSHAN JUMEI YOUPIN DECORATION ENGINEERING CO LTD

Digital twinborn mixed cloud-side collaborative intelligent real estate building group operation and maintenance intelligent system

The invention relates to the field of building intellectualization, in particular to a digital twinborn mixed cloud edge collaborative intelligent house building group operation and maintenance intelligent system, which establishes a digital twinborn scene database by collecting building basic information, equipment operation data and environmental parameters, synchronizes the digital twinborn scene database to edge equipment, and uses BIM, GIS, Internet of Things, 5G and AI technologies to establish a digital twinborn scene database, so as to realize the intelligent operation and maintenance of a building group. A virtual-real combined digital intelligent building scene is constructed, real-time synchronization of a virtual scene and a physical environment is realized through AR / VR equipment, and an operation and maintenance module comprises multi-source heterogeneous data fusion, edge intelligent analysis decision, adaptive model training and iterative optimization, a predictive maintenance algorithm of virtual-real mapping and a multi-level collaborative decision and autonomous scheduling mechanism. And the monitoring module monitors the state and operation condition of the edge equipment, provides data service and supports visualization of management decisions, and the system effectively improves the intelligence and digitization level of operation and maintenance of the building group.
Owner:CETHIK GRP

Methods for adjusting and / or controlling immersion associated with user interfaces

In some embodiments, an electronic device emphasizes and / or deemphasizes user interfaces based on the gaze of a user. In some embodiments, an electronic device defines levels of immersion for different user interfaces independently of one another. In some embodiments, an electronic device resumes display of a user interface at a previously-displayed level of immersion after (e.g., temporarily) reducing the level of immersion associated with the user interface. In some embodiments, an electronic device allows objects, people, and / or portions of an environment to be visible through a user interface displayed by the electronic device. In some embodiments, an electronic device reduces the level of immersion associated with a user interface based on characteristics of the electronic device and / or physical environment of the electronic device.
Owner:APPLE INC

Game animation character display method based on virtual reality technology

The invention provides a game cartoon character display method based on a virtual reality technology. The method comprises the following steps: acquiring three-dimensional model data of a game cartoon character in a target display area through a user interaction terminal; the virtual reality content management platform determines a dynamic rendering precision level according to the model data, and generates a real-time rendering strategy including a model patch reduction coefficient, a texture compression rate and a skeleton animation updating frequency in combination with terminal performance parameters; the strategy is sent to a virtual reality supervision platform and a user interaction terminal, and a virtual reality rendering engine platform is instructed to execute real-time rendering; and the virtual reality supervision platform monitors the frame rate fluctuation data of the head-mounted display device, calculates the scene rendering stability, sends a rendering optimization instruction if the scene rendering stability is lower than a threshold value, and adjusts the model data acquisition frequency and the video memory cleaning period to optimize the performance. The rendering efficiency and the system stability can be improved, and the hardware load and the frame rate fluctuation are reduced.
Owner:JIANGSU JIUQU INTERACTIVE ENTERTAINMENT NETWORK TECHNOLOGY CO LTD

Multi-modal interaction method and system of digital human intelligent agent

The invention relates to the field of multi-modal interaction analysis, in particular to a multi-modal interaction method and system of a digital human agent. The method comprises the following steps: acquiring a real-time face image and a voice signal input stream of an interactive user based on an intelligent agent; performing real-time micro-expression recognition and deep emotion analysis based on the real-time facial image to obtain real-time emotion features of the user; performing time sequence evolution analysis on the real-time emotion characteristics of the user, performing holographic user emotion deep mining, and constructing a user emotion holographic characteristic spectrum; carrying out adaptive acoustic gain processing on the voice signal input stream, and carrying out voice-emotion association analysis based on the user emotion holographic characteristic spectrum to generate a voice-emotion linkage mapping spectrum; and carrying out eyeball fixation point migration tracking based on the user emotion holographic feature map and the real-time face image, and generating a user interaction depth intention signal. Through the real-time deep semantic understanding and emotion perception ability, the intelligent agent interaction intelligence and response accuracy are improved.
Owner:GUANGDONG HUITONG INFORMATION TECH CO LTD

Multi-mode-based AI digital human intelligent interaction method, system and equipment

The invention relates to the technical field of computer vision and human-computer interaction, and discloses an AI digital human intelligent interaction method, system and equipment based on multiple modalities, and the method comprises the steps: pre-awakening a digital human when a human face is detected, and further thoroughly awakening the digital human based on recognized preset voice information or preset gesture information; voice and video information of a user in the interaction process is obtained, a keyword extraction result, a gesture recognition result and an emotional state tag are generated, a pre-constructed knowledge base is utilized to retrieve related information, a big language generation model module is combined to generate an answer text, and the answer text is input into a preset voice synthesis model to generate emotional voice output. And based on the current emotional state label of the user, driving the digital human animation to be output in an emotional manner. According to the method and the system, the digital human for understanding the emotion of the user, generating personalized answers, providing voices with rich emotions and displaying natural expressions and actions can be created, better interaction with the user can be realized, and more humanized and effective services can be provided.
Owner:BEI JING WAN JIE SHU JU KE JI YOU XIAN ZE REN GONG SI WU HAN FEN GONG SI +1

Replaceable conductive marker tip

Embodiments of the invention provide a removable marker tip configured for application with an active pen-stylus. The removable marker tip comprises an antenna and an insulator and is designed to be held onto the pen-stylus' writing shaft by a set of crush ribs fitted into a posterior recess on the removable marker tip. The removable marker tip is designed to be a consumable part to improve the paper feeling received by users of the pen-stylus as they draw on the display of a tablet device. The removable marker tip is further designed to be hand-removable by users of the pen-stylus.
Owner:REMARKABLE AS

Method and system for managing expressway construction based on BIM (Building Information Modeling) technology

The invention discloses a method and system for managing expressway construction based on a BIM technology. The method comprises the following steps: acquiring multi-modal data in expressway construction; constructing an environment-load-response multi-modal data fusion risk matrix by using a space-time correlation model; constructing an AI prediction model by using bridge motion signals and humidity and CO2 concentration data in the multi-modal data; dynamic risk early warning and hierarchical response are realized; the bridge vibration signals and thermal infrared imager data are utilized to analyze and identify invisible faults in highway construction; constructing a fault and health management scheme based on dynamic risk early warning, hierarchical response and invisible faults; performing construction simulation, conflict elimination and extreme working condition deduction by using the digital twin environment of the BIM technology; and according to a deduction result, carrying out highway construction abnormal area management through VR visualization. According to the scheme of the invention, the safety early warning, fault prediction and health management levels of highway construction can be improved.
Owner:HENAN HIGHWAY ENG GROUP

Adaptive scene intelligent interaction system based on AI

The invention, which relates to the technical field of intelligent interaction, discloses an AI-based adaptive scene intelligent interaction system comprising a multi-modal data acquisition module, a modal preprocessing module, a multi-modal embedded coding module, an intention fusion and representation module, a service scene matching module and a service execution and reinforcement learning module. The method comprises the following steps: acquiring multi-modal original data in a user interaction process, including voice signals, text input and user behavior tracks, and synchronously recording an acquisition timestamp; according to the method, through a multi-modal unified embedding and dynamic weighting mechanism, the problem of characteristic dimension imbalance is effectively solved, and the user intention recognition accuracy is improved; meanwhile, reinforcement learning and a multi-factor scoring model are combined, personalized scene matching and dynamic response are achieved, the adaptive capacity and service accuracy of the system in a complex environment are improved, and therefore the stability and user experience of the intelligent interaction system are remarkably optimized.
Owner:HENAN CITIC BIG DATA TECH CO LTD

Gaze-based text entry in a three-dimensional environment

In some embodiments, while a keyboard is visible in a three-dimensional environment, the computer system detects a gaze of a user move from a position away from a first key of the keyboard to the first key. In some embodiments, in response to detecting the gaze of the user moving from the position away from the first key to the first key, in accordance with a determination the one or more criteria are satisfied, the computer system initiates a process to select a first character corresponding to the first key for entry. In some embodiments, in response to detecting the gaze of the user moving from the position away from the first key to the first key, in accordance with a determination that the one or more criteria are not satisfied, the computer system forgoes initiating the process.
Owner:APPLE INC

Virtual-real fusion exhibition display interaction system and multi-mode perception method

The invention discloses a virtual-real fusion exhibition display interaction system and a multi-mode perception method, and belongs to the technical field of exhibition display interaction. The system collects audience eyeball fixation points, gesture actions and ambient light data through AR / VR equipment, analyzes coordinates of a region of interest through an eyeball fixation point attention mechanism, a gesture space-time encoder and a multi-modal fusion unit, triggers holographic projection explanation and virtual exhibition stand light and shadow dynamic adjustment (including illumination intensity, color, Gaussian blur and the like) based on a threshold value, and performs real-time display on the virtual exhibition stand. And multi-user collaborative interaction is realized through federal learning. According to the method, reinforcement learning is adopted to optimize an event-driven threshold value, and virtual and real visual splitting is eliminated in combination with ambient light adaptive mapping. The problems of low participation degree, insufficient single-mode interaction information and multi-user cooperation of traditional exhibition are solved, interest analysis accuracy is improved through multi-mode fusion, personalized experience is enhanced through dynamic interaction, the method is suitable for multiple scenes such as museums and science and technology museums, and exhibition intellectualization, immersion and group interaction efficiency are effectively improved.
Owner:SUZHOU ART & DESIGN TECH INST