Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

52 results about "Visual placement" patented technology

Webpage data acquisition method and related device

The invention discloses a webpage data collection method and a related device, and relates to the technical field of data processing.The webpage data collection method comprises the steps that a webpage data collection knowledge base containing data collection modes of various webpages is constructed in advance; a data acquisition mode matched with the webpage is obtained from a webpage data acquisition knowledge base, and the data acquisition mode defines that an analysis mode of each field in the webpage is analyzed based on multi-modal features (at least two of DOM structure fingerprint features, visual position features and semantic features of the fields); according to the method, the analysis accuracy of the webpage data can be greatly improved, so that stable and efficient execution of a webpage data acquisition task is guaranteed.
Owner:ANHUI IFLYTEK INTELLIGENT SYST

Multi-modal visual position identification reordering method and system based on guidance

The invention relates to the technical field of visual position recognition, and particularly discloses a multi-modal visual position recognition reordering method and system based on guidance, and the method comprises the steps: obtaining a query image, and retrieving a plurality of candidate images based on a pre-trained visual basic model and the query image; constructing a composite multi-modal prompt object, wherein the composite multi-modal prompt object comprises an image pair formed by the query image and the current candidate image, and an instruction text used for guiding a multi-modal large language model to perform visual comparison; outputting a structured similarity judgment result, wherein the result comprises a quantitative similarity score; and sorting based on the similarity scores corresponding to all the candidate images, and determining the candidate image with the highest score as an optimal matching result. Through combination of guiding type prompt engineering and structured output, an intermediate text generation link is avoided fundamentally, and the calculation efficiency is improved while the fidelity of all original visual information is reserved.
Owner:SHENZHEN 1024 ROBOT TECHNOLOGY CO LTD

Security monitoring alarm linkage processing method and system

The invention relates to the technical field of security and protection monitoring, and discloses a security and protection monitoring alarm linkage processing method and system. The method comprises the following steps: constructing a three-dimensional scene graph of a monitoring area; obtaining visual position information and wireless positioning information of a target, and executing constrained Kalman filtering fusion to obtain a target position; mapping the target position to the three-dimensional scene graph for threat assessment to obtain an intelligent alarm signal; and performing linkage control in the three-dimensional scene graph based on the intelligent alarm signal to obtain an equipment activation instruction and a blocking strategy. According to the method, high-precision three-dimensional target positioning is realized in a complex environment, an alarm decision strategy can be dynamically optimized according to historical statistical data, the problem of high false alarm rate caused by a fixed threshold mode is effectively reduced, and the stability and the reliability in different environment conditions are improved.
Owner:SHENZHEN CHUANGSHI GRAIN NETWORK CO LTD

Metacosmic scenic spot multi-modal interaction special effect generation method and system based on big data

The invention discloses a big data-based universe scenic area multi-modal interaction special effect generation method and system, and relates to the technical field of virtual reality, and the method comprises the steps: obtaining universe scenic area information, obtaining historical tourist information according to the universe scenic area information, and generating a universe scenic area multi-modal interaction special effect according to the historical tourist information; and obtaining historical interaction time information and historical interaction tourist information corresponding to each interaction special effect. According to the method, the preferred tourists of each interaction special effect are accurately analyzed through the preferred tourist feature information, a data basis is provided for subsequent interaction special effect generation, the interaction deviation angle model is constructed through the special effect visual position and the tourist position, the special effect interaction degree is calculated in combination with the interaction time, single interaction duration judgment is replaced, and the interaction efficiency is improved. The tourist preference analysis accuracy is improved, the appropriate interaction special effect is selected by screening the interaction special effects corresponding to the meta universe scenic spots, the special effect individuation adaptation degree is improved, and the tourist immersion experience is optimized.
Owner:CHONGQING TOURISM CLOUD INFORMATION TECH CO LTD

In-vehicle target positioning method and device, vehicle and storage medium

The invention provides an in-vehicle target positioning method and device, a vehicle and a storage medium, the method is applied to the field of sound source localization, and the method comprises the steps: carrying out the hand image collection and audio collection of a user in the vehicle, and the audio collection is executed by a plurality of audio collectors; determining an acoustic position coordinate of the hand of the user according to a first acquisition time when each audio acquisition device acquires the sound; according to the first acquisition time, the acquired hand image and the acquisition time of the hand image, determining a visual position coordinate of the hand of the user and a confidence coefficient of the hand of the user in executing the preset gesture; and performing filtering correction on the acoustic position coordinates according to the visual position coordinates and the confidence coefficient. The method can effectively reduce the positioning error caused by a single sensor, is high in positioning precision, is high in anti-interference capability, and is suitable for various complex in-vehicle environments.
Owner:GREAT WALL MOTOR CO LTD

Indoor positioning and tracking method and system based on UWB and vision fusion, and storage medium

The embodiment of the invention relates to an indoor positioning and tracking method and system based on UWB and visual fusion and a storage medium, and the method comprises the steps: firstly carrying out the real-time detection through employing a camera, obtaining a plurality of visual position points, determining a target candidate point in the plurality of visual position points based on a position track, and carrying out the positioning and tracking of the target candidate point; determining a visual moving direction based on the visual track of the target pedestrian and the target candidate point, determining whether the visual moving direction is matched with the track direction of a first UWB track, if yes, updating the position point of the position track at the current moment based on the target candidate point, and if not, updating the position point of the position track at the current moment based on the target candidate point; and if yes, updating the position point of the position track at the current moment based on the position point of the position track at the previous moment and the position point of the first UWB track at the current moment. According to the method, visual positioning and UWB positioning are fused, track direction matching is carried out, a position track is updated based on a matching result, and the accuracy of positioning and tracking is improved.
Owner:DONGLAI INTELLIGENT TRANSPORTATION TECH (SHENZHEN) CO LTD

A visual position recognition method, system and device based on attention compression coding features

The application provides a visual position recognition method based on attention compression coding features. A hierarchical database is established based on a scene three-dimensional map; and global features are extracted by using an attention mechanism and codebook compression coding based on the hierarchical database to realize online position recognition. The application is used to reduce the calculation overhead of feature extraction in offline and online stages, improve global recall accuracy, and thus guarantee the service quality of user positioning.
Owner:HARBIN INST OF TECH

A micro-fan coil sealing structure, a coil packaging process and a micro-fan

The application belongs to the technical field of coil packaging, and particularly relates to a micro-fan coil sealing structure, a coil packaging process and a micro-fan. The micro-fan coil sealing structure comprises a sealing cover body, a plurality of independent sealing cavities for accommodating coils are uniformly and interval arranged along the circumferential direction on one side of the sealing cover body close to the coils; a positioning structure for fixing the coils is arranged in the sealing cavities. The independent sealing cavities are arranged on the sealing structure of the coils to provide independent mounting space and positioning reference. Meanwhile, the positioning structure is arranged in the sealing cavities, so that the radial / circumferential position of the coils in the sealing cavities is more easily limited. When batch assembly is performed, the randomness of'manual visual placement' can be reduced, and the consistency of the coil mounting position is improved. The stable position of the coils means that the relative position of the coils and the rotor magnetic steel / magnetic circuit is more stable, so that the problems of uneven magnetic field, torque fluctuation and the like caused by eccentricity and inclination are reduced. Furthermore, the fan operation is stable, and the probability of abnormal vibration or efficiency fluctuation is reduced.
Owner:SUZHOU XINGKAISHENG INTELLIGENT TECHNOLOGY CO LTD

Virtual reality video interaction method and device, electronic equipment and storage medium

The invention provides a virtual reality video interaction method and device, electronic equipment and a storage medium. According to the invention, a video display sphere associated with the position of a head-mounted display device is established in a virtual reality scene, and the video display sphere is used for playing a panoramic video containing a dynamic object and maintaining a fixed picture reference orientation; determining the real-time pose of the dynamic object under the current video playing progress from the pose change data sequence of the dynamic object during video production; setting a dynamic interaction body in a coordinate system corresponding to the video display sphere according to the real-time pose, and maintaining a synchronization relationship between the dynamic interaction body and the video display sphere; and when detecting that an interaction event occurs between the user interaction operation and the dynamic interaction body, generating interaction feedback in an area corresponding to the visual position of the dynamic object in the panoramic video picture.
Owner:BEIJING QIYI CENTURY SCI & TECH CO LTD

A cross-domain cross-source data alignment method, system and electronic device

The application discloses a cross-domain cross-source data alignment method and system and electronic equipment. The method first inputs a plurality of sets of table data to be aligned. Then, key-value pairs in the data and positions of the key-value pairs in the table are extracted. Next, a data multi-modal representation model is used to generate vector expressions of the keys, values and visual positions. Semantic distances of the vector expressions from different data are calculated. Finally, the semantic distances between different data are evaluated to determine an alignment result. In addition to using pairing between keys, the application considers matching of values, enhancing matching of keys in the prior art. In addition to the representation in the text, the application fuses a table visual structure as part of semantic representation of the key-value pairs, breaking the limitation of the prior art that only uses single-modal information for matching.
Owner:WUHAN UNIV

Video fusion method, system, device and program product

The invention provides a video fusion method, system and device and a program product, and the method comprises the steps: obtaining a real-time video stream to be fused to a three-dimensional scene model, and pre-calibrating a display plane and a three-dimensional scene reference point in the three-dimensional scene model; determining a coordinate of a visual position point of a three-dimensional scene reference point on a display plane under a visual angle of a current world camera corresponding to a current frame original image in the real-time video stream; mapping the visual position point into a calibrated pixel coordinate range to obtain a coordinate of a target image reference point; calculating a coordinate transformation relation between the original image and the target image according to the coordinate of the target image reference point and the calibrated coordinate of the original image reference point; performing transformation processing on the current frame original image based on the corresponding coordinate transformation relation to obtain a corresponding target image; and rendering the target image on the display plane. According to the invention, a video fusion angle adaptive mechanism is added, and the video fusion effect is effectively improved.
Owner:SUZHOU KEYUAN SOFTWARE TECH DEV +1

Intelligent tool traceability method and system

The invention relates to the technical field of industrial Internet of Things and edge intelligence, and provides an intelligent tool tracing method and system. According to the implementation scheme, the method comprises the steps of sensing an instrument occurrence event and constructing each observation data node; generating a spatial connectivity hypothesis based on visual similarity among the observation data nodes so as to dynamically construct a tool movement trajectory diagram; clustering the environmental visual features of all observation data nodes to construct an environmental cognitive network; in response to triggering of a tool search event, confirming a search starting point based on a multi-starting-point competition mechanism; performing probability diffusion calculation based on multi-source behavior evidence fusion on the environment cognitive network by taking the search starting point as an initial source, and generating probability distribution of the target tool in each position in the network; based on the probability distribution, a search instruction is generated, and the search instruction comprises visual position information and semantic information of the target tool. According to the embodiment of the invention, rapid positioning and intelligent traceability of the tool can be realized under the condition of no preset map.
Owner:QIANTANG BRANCH OF ZHEJIANG DAYOU IND CO LTD +3

Position recognition method and device based on entropy-controlled hierarchical attention, equipment and medium

The application relates to the field of artificial intelligence and discloses a position recognition method and device based on entropy-controlled hierarchical attention, equipment and a medium, which comprises the following steps: receiving an input image to be recognized and performing overlapping block embedding processing to generate an initial overlapping token sequence; performing hierarchical multi-scale feature extraction on the initial overlapping token sequence, dynamically regulating the attention distribution of each layer by information entropy during feature extraction to obtain preliminary features of each layer; performing feature refining and integrating processing on the preliminary features of all layers to output corresponding scene descriptors; performing scene matching in all preset scenes according to the scene descriptors, and outputting the position information of the input image according to the scene matching result. The application can be applied to business scenes such as financial technology and medical health, dynamically regulates the attention distribution of each layer based on information entropy during hierarchical multi-scale feature extraction, adaptively focuses on the most discriminative scene area for position recognition, and improves the accuracy of visual position recognition.
Owner:PING AN TECH (SHENZHEN) CO LTD

Method and system for out-of-domain detection

The invention relates to a method and system for out-of-domain detection. Methods, systems, and aircraft for performing image analysis to assist in refueling operations are disclosed herein. The method includes receiving a 2D image from a camera, determining a domain score for the 2D image based on previously defined training data, and transmitting the 2D image to a visual position estimation system in response to the domain score being greater than a predetermined threshold, thereby creating a transmitted 2D image.
Owner:THE BOEING CO

Unified transformer-based visual place recognition framework

A unified place recognition framework handles both retrieval and re-ranking with a unified transformer model. The re-ranking modules utilizes feature correlation, attention value, and x / y coordinates into account, and learns to determine whether an image pair is from a same location.
Owner:LEMON INC(GB)

Wafer arrangement method and system for improving wavelength uniformity of MOCVD epitaxial wafer

PendingCN122304023AWaferRobotic hand
This invention provides a wafer arrangement method and system for improving the wavelength uniformity of MOCVD epitaxial wafers. The wafer arrangement method for improving the wavelength uniformity of MOCVD epitaxial wafers includes: measuring and calibrating the wavelength offset of each groove on the carrier; calculating the predicted wavelength offset of the wafer to be processed using a pre-built wavelength test model; sorting the wafers to be processed in the current batch according to their predicted wavelength offsets; sorting the grooves on the carrier in the opposite manner to their predicted wavelength offsets; then matching the sorted wafers with the sorted grooves in sequence; mapping the matching results back to the original physical number of the groove to generate a visual placement guide, which is then used by an operator or robot to perform the corresponding placement under the guidance of the visual placement guide.
Owner:JIANGXI ZHAO CHI SEMICON CO LTD

Honeycomb core aramid paper automatic overlapping machine positioning method and system

The invention discloses a positioning method and system for an automatic overlapping machine of honeycomb core aramid paper, the positioning system comprises a visual detection system, a visual position moving system, a position compensation mechanism and a control system, and the positioning system is characterized in that the visual detection system is arranged on the visual position moving system; the control system includes a recipe management system and can drive the visual position movement system and the position compensation mechanism. According to the method, image information of overlapped honeycomb core aramid paper is collected through the visual detection system, the image information of the overlapped honeycomb core aramid paper and a pre-stored reference pattern are compared through the control system, the position of the overlapped honeycomb core aramid paper is automatically adjusted, and high-precision overlapping is achieved. Position parameters are pre-stored in a formula management system of the control system, automatic position adjustment of the visual detection system is achieved when honeycomb core aramid paper of different specifications is replaced, the debugging time is saved, and the debugging material loss is reduced. Through the improvement, the production efficiency can be improved on a large scale, and the overlapping production of multi-specification and large-scale honeycomb core aramid paper is realized.
Owner:BEIJING DONGSHENGXINRUI AUTOMATIC TECH

Method for visual-haptic fusion control of a facility agriculture picking robot

PendingCN122375369AActuatorBiology
The application provides a kind of facility agriculture picking robot vision-haptics fusion control method, the method comprises: by obtaining the RGB-D image and laser radar point cloud of facility environment, extract crop ridge center line, real-time solution chassis linear velocity vector and angular velocity vector;Obtain the three-dimensional point cloud and surface image of fruit, solve the maturity grade of fruit, and construct visual stiffness prior model, estimate the elastic modulus of fruit;According to the elastic modulus on-line mapping generation mechanical arm end impedance controller's target stiffness matrix and damping matrix;Adopt the position-based visual servo strategy, the pose error between end effector and target fruit is converted into the expected angular velocity of each joint of mechanical arm, drive end to approach target;When the contact force detected by end six-dimensional force sensor reaches the preset threshold, the system is switched from pure visual position control to adaptive impedance control mode, and the visual planning path is dynamically corrected using force feedback.
Owner:WUXI CITY COLLEGE OF VOCATIONAL TECH

Vertical turning, milling and grinding integrated combined machining center

The invention relates to the technical field of machining centers, in particular to a vertical turning, milling and grinding integrated combined machining center which comprises a base, two first electric sliding rails, a portal frame, a second electric sliding rail, a controller, a combined machining unit, an eccentric rotating workbench, a position precise control mechanism, a pose adjusting mechanism and a visual position detector. Wherein the two first electric sliding rails are fixedly installed on the left side and the right side of the top of the base respectively, the portal frame is fixedly installed at the output ends of the two first electric sliding rails, the second electric sliding rail is fixedly installed on the top of the portal frame, and the controller is fixedly installed on the front side of the base. Parallel operation of machining and pre-positioning is achieved, the defect of insufficient precision caused by coordinated control of multiple servo systems is overcome, and precision accumulative errors caused by multiple times of clamping are eliminated; the method has the outstanding advantages of being high in machining efficiency, stable in positioning precision and high in adaptability.
Owner:SHANGHAI LUOHAN OUJIE PRECISION MASCH CO LTD

Visual method-based in-vehicle sound source positioning method, controller and computer readable storage medium

Embodiments of the present application disclose a visual-based in-vehicle sound source positioning method, a controller and a computer readable storage medium. The method comprises: acquiring visual data, determining a visual position of a target person based on the visual data, determining a plurality of visual noise penalty parameters based on the visual data, generating a visual covariance matrix based on the plurality of visual noise penalty parameters, performing Kalman filtering processing on the visual position and the visual covariance matrix based on a preset Kalman filter, and obtaining a sound source position of the target person. In the embodiments of the present application, the plurality of visual noise penalty parameters are determined based on the visual data, so as to measure the interference of various noise factors on the sound source position from multiple dimensions, the visual covariance matrix is generated by comprehensively considering the plurality of visual noise penalty parameters, and finally, the Kalman filtering processing is performed on the visual position and the visual covariance matrix by using the Kalman filter, so as to obtain a sound source position with high reliability, thereby improving the positioning accuracy and robustness.
Owner:贵州华鑫信息技术有限公司

Visual position identification method based on multi-path depth separable convolution enhancement fine tuning

The invention discloses a visual position identification method based on multi-path depth separable convolution enhancement fine tuning, and relates to the field of visual basic model fine tuning and visual position identification. The method comprises the following steps of: firstly, designing an adaptive image encoder consisting of a multi-path depth separable convolution enhancement trimmer and a visual basic model, and realizing seamless field adaptation of the large-scale visual basic model only through a small number of adjustable parameters of the trimmer; respectively extracting adaptive features of the query database image and the reference database image; secondly, processing the adaptive features by using a generalized average pooling layer so as to respectively generate adaptive global features of query and reference database images; and finally, by taking the adaptive global feature as a data source, executing similarity measurement by calculating the cosine similarity between the query database image and the reference database image, thereby determining a final visual position identification result. And the model completes end-to-end training by adopting triple loss.
Owner:BEIJING UNIV OF TECH

A factory equipment operation status monitoring system and method based on graphene photoelectric sensing

This invention discloses a factory equipment operation status monitoring system and method based on graphene photoelectric sensing, relating to the field of suspended ball monitoring technology in industrial equipment. It is suitable for non-visual position detection of suspended balls inside glass containers in dusty environments. It can monitor the position of suspended balls inside glass pipes on hundreds of pieces of equipment in a factory to determine whether the equipment is operating normally. This invention uses a graphene photoelectric sensor installed on the outside of the glass pipe, detecting the suspension position of the ball through changes in the graphene photocurrent signal, achieving non-contact, non-visual monitoring. The system includes a graphene photoelectric sensor module, a laser module, and a data transmission and alarm module. It can monitor multiple devices in real time, supports wireless networking, and is suitable for industrial environments with severe dust pollution. This overcomes the shortcomings of traditional visual monitoring in dusty environments, achieving reliable position detection of suspended balls inside glass pipes, and is cost-effective and easy to operate and install.
Owner:NANJING INST OF TECH

Target screening method and device in curve scene, electronic equipment and storage medium

The invention discloses a target screening method and device in a curve scene, electronic equipment and a storage medium, and relates to the field of automatic driving, and the method comprises the steps: obtaining the detection information of a left lane line and a right lane line of a lane where a vehicle is located, and generating a vehicle center trajectory; obtaining target visual detection position information and visual detection corner point information of a vehicle in front of the vehicle; according to the target visual detection position information of the vehicle in front of the vehicle and the central trajectory of the vehicle, calculating the nearest distance between the target visual position of the vehicle and the central trajectory of the vehicle, and judging whether the nearest distance is smaller than a dangerous collision distance threshold; according to the visual detection angular point information and the central trajectory of the vehicle, calculating the nearest distance between the target visual detection angular point of the vehicle and the central trajectory of the vehicle, and judging whether the nearest distance is smaller than a dangerous collision distance threshold value or not; setting a car following target; the step of setting the vehicle following target comprises setting the vehicle in front of the vehicle determined as the dangerous target as the vehicle following target and outputting the vehicle following target to vehicle control execution.
Owner:FAW JIEFANG AUTOMOTIVE CO

A visual position recognition method, electronic device, and medium

The application discloses a visual position recognition method, an electronic device and a medium, and comprises the following steps: acquiring an input image; extracting a feature vector of the input image by using a convolutional neural network; training a principal component analysis conversion model based on unsupervised learning; reconstructing the feature vector of the input image by using the principal component analysis conversion model to generate an image description vector; acquiring an image description vector of an existing image from a database, and calculating the similarity between the image description vector of the input image and the image description vector of the existing image; when the maximum similarity is greater than or equal to a similarity threshold, regarding the existing image corresponding to the maximum similarity as a similar image of the input image to obtain a visual position recognition result.
Owner:ZHEJIANG UNIV

Methods and systems for out-of-domain detection

Disclosed herein are methods, systems, and aircraft for performing image analysis for aiding refueling operations. A method includes receiving a 2D image from a camera, determining a domain score for the 2D image based on previously defined training data, and sending the 2D image to the vision position estimation system in response to the domain score being greater than a predefined threshold, thus creating a sent 2D image.
Owner:THE BOEING CO

Methods, apparatus, and programs for fast-moving audio source rendering

The disclosure relates to methods of rendering object-based audio content, comprising: determining a rendering position at each of a plurality of time instances and maintaining a list of one or more known rendering positions including a current rendering position and a number N (N ≥ 0) of previous rendering positions at a current and previous time instances, respectively; determining a measure of a time offset based on the current rendering position and a rendering position at the preceding time instance; determining a modeled time instance corresponding to one of the known rendering positions, closest in time to a timing that precedes the current time instance by the time offset; and determining a modeled rendering position based on a rendering position among the known rendering positions that corresponds to the determined modeled time instance, for modeling an offset between auditory and visual positions of an audio object. The disclosure further relates to corresponding apparatus, programs, and computer-readable storage media.
Owner:DOLBY INTERNATIONAL AB

Image processing method, computing device, storage medium, and computer program product

Embodiments of the description in the present disclosure relate to the technical field of computers, and in particular to an image processing method, a computing device, a storage medium, and a computer program product. The image processing method comprises: acquiring an image to be processed, a prompt text associated with said image, and visual position information of said image; acquiring a target text feature corresponding to the prompt text, a target image feature corresponding to said image, and a target position feature corresponding to the visual position information; and on the basis of the target text feature, the target image feature, and the target position feature, acquiring an image processing result corresponding to said image, wherein the image processing result is a result of processing said image on the basis of text content of the prompt text.
Owner:ALIBABA (CHINA) CO LTD

System for coordinating geographic information with virtual objects

A method for visually displaying visual location indicators on a display of a device having a camera element, including determining position and orientation of the device, determining a geographic boundary based on position of the device, determining locations within the geographic boundary, determining objects proximate the device using the camera element, correlating objects proximate the device with locations within the geographic boundary, associating one visual location indicator with one location proximate the device, and displaying the one visual location indicator with the one location on the display of the device when the camera element is directed toward the one location.
Owner:COREPORA INC

Indoor positioning and tracking method and system based on fusion of uwb and vision, and storage medium

Embodiments of the present application relate to an indoor positioning and tracking method and system based on UWB and vision fusion, and a storage medium. The method first detects in real time by using a camera to obtain a plurality of visual position points, determines a target candidate point based on a position trajectory in the plurality of visual position points, determines a visual movement direction based on a visual trajectory of a target pedestrian and the target candidate point, and determines whether the visual movement direction matches a trajectory direction of a first UWB trajectory, wherein the first UWB trajectory is a UWB trajectory of the target pedestrian. If the visual movement direction matches the trajectory direction of the first UWB trajectory, the position point of the position trajectory at a current time is updated based on the target candidate point. If the visual movement direction does not match the trajectory direction of the first UWB trajectory, the position point of the position trajectory at the current time is updated based on a position point of a position trajectory at a previous time and a position point of the first UWB trajectory at the current time. The method fuses visual positioning and UWB positioning, matches trajectory directions, updates the position trajectory based on a matching result, and improves the accuracy of positioning and tracking.
Owner:DONGLAI INTELLIGENT TRANSPORTATION TECH (SHENZHEN) CO LTD