Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

213 results about "Image mode" patented technology

Image mode refers to the type of image acquisition (modality). For example, most ultrasound systems use 2D, Color Flow, M Mode and Doppler modes.

Multi-modal data glaucoma classification method and system, terminal and medium

The invention discloses a multi-modal data glaucoma classification method and system, a terminal and a medium, and the method comprises the steps: obtaining a color fundus image and an OCT image in different image modes, and carrying out the advanced feature extraction of the color fundus image and the OCT image based on a first feature extraction module and a second feature extraction module, the mode specificity characteristics of the two modes are obtained; fusing the mode specific features of the two modes based on a shallow feature fusion module to obtain shallow fusion features; inputting the mode specific features of the two modes and the shallow fusion features into a deep feature fusion module for fusion to obtain final fusion features; and based on a multi-layer perceptron classifier, performing classification prediction on the final fusion feature to obtain a classification result, the classification result including no glaucoma, early glaucoma or mid-late glaucoma. According to the method, through two-stage feature extraction, two-stage feature fusion and the classifier, accurate glaucoma classification can be obtained.
Owner:SHENZHEN UNIV

Laser SLAM mapping positioning method and system in dynamic environment

The invention discloses a laser SLAM mapping positioning method and system in a dynamic environment. The method comprises the steps that point cloud information and IMU measurement data in the dynamic environment are collected; the data preprocessing module carries out pre-integration by using IMU data so as to remove point cloud distortion; the pose estimation module realizes positioning based on error Kalman filtering by using IMU measurement information and point cloud feature matching; the dynamic point filtering module stores the three-dimensional point cloud in a two-dimensional image mode, judges whether the three-dimensional point cloud is a dynamic point or not based on the change condition of the point cloud at the same pixel, and finally carries out clustering elimination to output the point cloud only containing a static scene; and the loopback detection module extracts descriptors based on the static scene point cloud, and corrects the pose deviation through ICP (Inductively Coupled Plasma) matching. According to the method, the problem that dynamic object interference occurs in the mapping positioning scene is effectively solved, the dynamic points do not interfere with positioning, the map does not contain ghosts generated by moving objects, help is provided for follow-up navigation and the like, and the robustness and accuracy of mapping positioning are improved.
Owner:SOUTH CHINA UNIV OF TECH

Construction state monitoring and risk assessment method and device based on BIM (Building Information Modeling) multi-mode conversion

The invention provides a construction state monitoring and risk assessment method and device based on BIM multi-mode conversion, and relates to the technical field of building information models. According to the method, a standardized image mode is generated by analyzing and extracting component information of a BIM model, and a BIM text mode is generated by using natural language description; constructing a graph structure mode based on space and construction logic, and realizing unified alignment and deep fusion of multi-modal data through multi-level modal alignment and a cross-modal attention mechanism to obtain a cross-modal fusion representation which is used for inputting a state recognition model and automatically detecting an execution deviation so as to monitor a construction state; and then introducing a deviation conduction mechanism to quantitatively calculate a comprehensive risk index of the component so as to carry out risk assessment. According to the method, the fusion representation which not only keeps semantic consistency but also conforms to construction logic can be obtained, the abstract cross-modal semantic features are converted into quantifiable and interpretable construction states and risk indexes, and powerful support is provided for intelligent analysis and application in a construction scene.
Owner:XIAMEN UNIV OF TECH

Gas identification method based on multi-source information fusion and environmental perception

The invention discloses a gas recognition method based on multi-source information fusion and environmental perception, and the method comprises the steps: constructing a deep feature learning framework of multi-source fusion through combining the spatial response features, time sequence features and external environmental information of gas; the method comprises the following specific steps: preprocessing collected gas data, and respectively extracting features of an image mode, a sequence mode and an environment mode; fusing the image features and the sequence features through a cross attention fusion module, and capturing the space-time correlation of the data; a cross-modal attention compensation module is introduced, so that main-modal gas data adaptively gathers key information in an auxiliary-modal environment, and effective compensation of environmental factors on gas recognition performance is realized; and finally, gas category prediction is performed through a classification decision head. According to the method, the problems that the detection is easily interfered by environmental factors, the stability is poor or the qualification is inaccurate due to the fact that modeling depends on single modal data in the existing gas identification technology are solved.
Owner:BEIJING INFORMATION SCI & TECH UNIV

Slope stability grade identification method and system based on multi-modal deep learning

The invention belongs to the technical field of geological disaster risk assessment, and particularly relates to a slope stability grade identification method and system based on multi-modal deep learning. Comprising the following steps: acquiring a remote sensing image and corresponding parameter data, and respectively preprocessing the remote sensing image and the corresponding parameter data to obtain a high-dimensional image feature vector and a parameter feature vector; performing dynamic fusion of multi-modal information on the high-dimensional image feature vector and the parameter feature vector through a modal-level gating and fine-grained weighting mechanism to obtain a fusion feature; and inputting the fusion features into a preset classifier to obtain probability distribution of slope stability levels, and outputting the class with the maximum probability as a stability level identification result. According to the method, through combined modeling of the image modality and the parameter modality and introduction of an improved gating attention mechanism in a feature fusion stage, a reliable explanatory basis can be provided while high-precision prediction is ensured, so that the engineering applicability and the popularization value of the model are enhanced.
Owner:SHAANXI PROVINCIAL GEOLOGICAL ENVIRONMENT MONITORING STATION +1

Thermal barrier coating defect analysis method based on terahertz multi-domain imaging weighted fusion

The invention provides a thermal barrier coating defect analysis method based on terahertz multi-domain imaging weighted fusion, which comprises the following steps: according to different positions of defects in a thermal barrier coating, respectively intercepting signals of imaging modes of the surface, the interior, the lower interface and the integral structure of the thermal barrier coating from thermal barrier coating signals; calculating a time domain index value and a statistical domain index value in each signal, arranging the index values into a two-dimensional matrix, and drawing a two-dimensional imaging graph; carrying out fusion imaging on the two-dimensional image by adopting a weighted fusion method to obtain a terahertz fusion image; segmenting and identifying defect features in the terahertz fusion image to obtain defect type and size information; according to the thermal barrier coating falling failure judgment criterion, the failure state of the thermal barrier coating is evaluated according to the defect area proportion; according to the method, the terahertz multi-domain imaging graph is adopted to effectively extract the early internal defect characteristics of the thermal barrier coating, the defect characteristics are enhanced through weighted fusion imaging, and finally early internal defect detection and failure evaluation are achieved.
Owner:AVIC XIAN AIRCRAFT IND GRP CO LTD

Knowledge graph construction method and device, equipment, storage medium and computer program product

The invention provides a knowledge graph construction method and device, equipment, a storage medium and a computer program product. The method comprises the following steps: acquiring a script text and a deductive video corresponding to the script text; constructing a plurality of first nodes of a text mode based on the script text, and constructing a plurality of second nodes of an image mode and a plurality of third nodes of an audio mode based on the deductive video; based on the content information corresponding to the fourth node and the fifth node, determining a semantic consistency score between the fourth node and the fifth node; determining a time sequence synchronization score between the fourth node and the fifth node based on the time sequence information corresponding to the fourth node and the fifth node; in response to the fact that the semantic consistency score is larger than a first threshold value and the time sequence synchronism score is larger than a second threshold value, constructing a first relationship between the fourth node and the fifth node; and constructing the knowledge graph based on the plurality of first nodes, the plurality of second nodes, the plurality of third nodes and the plurality of first relationships.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Small sample classification method for multi-granularity image-text prototype matching based on task-driven structure

The invention relates to the technical field of small sample image classification, in particular to a small sample classification method for multi-granularity image-text prototype matching based on a task-driven structure. The method comprises the following steps: constructing a support set and a query set; inputting the two features into a multi-granularity feature set extraction network, and extracting low-level, middle-level and high-level image features by a multi-level image feature extraction network consisting of an image modal input improved ResNet12 and an embedded task-driven mapper selection module; the text mode calls a large language model (LLM) three times successively to progressively generate word-level, sentence-level and paragraph-level semantic descriptions, and a text editor is combined to form multi-view text features; the image and text feature set is input into a semantic graph interactive fusion network, and multi-level information fusion is realized through node and edge construction, cross-modal attention, graph convolution and a gating mechanism; the matching score of each layer is calculated through the fusion features through a multi-granularity prototype matching network, so that classification prediction is carried out, and the performance of the model is optimized by adopting a multi-granularity composite weighted loss function in a training stage.
Owner:CHANGCHUN UNIV OF SCI & TECH +1

Dual-mode forged information detection method fusing emotion features

The invention discloses an emotion feature fused bimodal forged information detection method, which comprises the following steps of: extracting emotion features, respectively extracting emotion categories of a text and an image through a large language model, respectively giving emotion intensity of the categories, and scoring the emotion; analyzing the emotion consistency of the mixed modal features of the text and the image; the hybrid expert network calculates domain probabilities to which a text mode, an image mode and a hybrid mode belong respectively through a text domain encoder, an image domain encoder and an image text hybrid feature domain encoder based on domain class processing features; the hybrid expert network learns feature information of a text mode, an image mode and a hybrid mode according to the domain information; the AdaIN network adjusts modal distribution, the AdaIN network adjusts feature distribution according to contribution of each modal, the adjusted features are fused and input into the classifier, a final classification result is obtained, and the information detection accuracy is greatly improved.
Owner:SHENZHEN KIM DAI INTELLIGENCE INNOVATION TECHNOLOGY CO LTD

Circuit board defect identification method and system based on multi-dimensional image data

The invention relates to a circuit board defect identification method and system based on multi-dimensional image data, and belongs to the technical field of data identification processing, and the method comprises the following steps: obtaining synchronous image data of a circuit board to be detected in a plurality of imaging modes; performing space-spectrum joint registration on each modal image to generate a multi-dimensional image cube with a unified coordinate system and pixel alignment; inputting the multi-dimensional image cube into a pre-trained multi-branch heterogeneous fusion neural network; generating a pixel-level defect probability graph by utilizing a defect sensing context decoder, and performing geometric constraint optimization on the probability graph by combining prior information of a circuit board design layout; outputting defect types, positions and confidence coefficients, and establishing an interpretable defect fingerprint database according to the multi-dimensional response characteristics of the defects; the method has the beneficial effects that false defect signals generated by image noise and circuit board surface texture interference can be effectively inhibited, the omission ratio and the false detection ratio are greatly reduced, and pixel-level accurate defect positioning is realized.
Owner:SICHUAN MEIJIESEN CIRCUIT TECH CO LTD

Underwater online intelligent monitoring analysis system and method

The invention belongs to the technical field of underwater sea condition data monitoring, and discloses an underwater online intelligent monitoring analysis system and method. According to the method, an acoustic Doppler current profiler (ADCP) carried by a seabed observation system is used for carrying out seabed flow measurement, the speed and direction of water flow are obtained, and a binocular camera carried by the seabed observation system is used for determining the size of underwater organisms; performing image mode recognition and classification on the underwater creatures by using an image recognition algorithm in the underwater observation system, and recognizing the types and the number of the underwater creatures; transmitting data to a communication module of the virtual anchor system unmanned platform by using the armored cable, performing error control on the data transmitted by the armored cable, and detecting and correcting error information of the data in the transmission process; and judging whether the returned information is successful or not through an information transmission control method in the underwater observation system. According to the invention, the method can achieve the image recognition of Branchiostoma under a complex sea condition, and improves the operation efficiency and safety.
Owner:QINGDAO HAIYAN ELECTRONICS CO LTD

System and method for multi-modal contrast in few-shot classification

An LMMs-boosted LMC framework system includes a user interface, a multi-modal feature generation module, a multi-modal feature contrast module, and a prediction logic module. The user interface operates in support image mode to process support set images for training, transforming them into text and visual features within a multi-modal support feature pool. In test image mode, the user interface transmits test images to the contrast module, which compares their features with the support pool using visual and textual modalities. The prediction logic module integrates these comparisons to generate a classification index with the most likely classification and confidence score.
Owner:THE HONG KONG UNIV OF SCI & TECH

Multi-mode heterogeneous information collaborative weld defect X-ray image intelligent diagnosis and credible traceability method

The invention discloses a welding seam defect X-ray image intelligent diagnosis and credible traceability method based on multi-modal heterogeneous information collaboration, which is characterized in that a defect analysis network fusing multi-domain feature modeling and graph structure expression is constructed on the basis of bimodal data formed by a welding seam X-ray image and an industry detection standard text. In the image mode, dividing the weld seam image into a plurality of local area units through superpixel segmentation, taking the areas as image nodes, respectively extracting spatial domain, frequency domain, wavelet domain and edge domain features, and constructing a weighted graph structure by combining the spatial adjacency relation and the feature similarity relation between the areas; realizing overall modeling and correlation analysis of weld defect structure information by using a graph convolutional network; in a text mode, feature coding is carried out on an industry detection standard text, and the feature coding is used as an important prior constraint for defect judgment. Collaborative modeling of an image detection result and standard semantic information is achieved through a gating fusion mechanism, a mapping relation between a detection conclusion and a standard term is established, and interpretable expression and result credible traceability of the weld defect diagnosis process are achieved. And a welding seam X-ray film automatic digital acquisition and observation device is adopted in a matched manner, so that stable transmission, positioning observation and high-resolution digital imaging of the industrial ray film are realized, and reliable and consistent image data input is provided for the intelligent diagnosis method. The method is suitable for intelligent defect detection under complex welding seam structures and multi-working-condition imaging conditions, and has high engineering application value and popularization prospect.
Owner:TAIYUAN UNIVERSITY OF SCIENCE AND TECHNOLOGY

Glaucoma multi-mode auxiliary diagnosis device and electronic equipment

According to the glaucoma multi-mode auxiliary diagnosis device and the electronic equipment, firstly, first feature extraction and second feature extraction are carried out on an eye fundus image and an OCT image respectively, then semantic spaces of an eye fundus image mode and an OCT image mode are aligned by utilizing comparison loss, and the first feature and the second feature are fused by utilizing a cross attention mechanism, so that an eye fundus image is obtained. According to the method, the two modal data are subjected to fusion to obtain fusion features, then the fusion features are subjected to feature extraction to obtain third features, and glaucoma classification judgment is performed based on the third features, so that global modeling of the two modal data can be realized, effective features are extracted to perform glaucoma classification judgment, and the accuracy and timeliness of diagnosis are ensured.
Owner:CENT SOUTH UNIV

Method for predicting service life of diamond grinding wheel for spiral groove grinding of solid carbide cutter

The invention discloses a method for predicting the service life of a diamond grinding wheel of a solid carbide cutter grinding spiral groove, which comprises the following steps that a machine learning module is adopted to estimate the abrasion loss of the grinding wheel diameter along the axial direction of the grinding wheel, and the input quantity of machine learning is cutter parameters, grinding wheel parameters, grinding parameters and other information; the output quantity of machine learning is the abrasion loss of the grinding wheel along the axial diameter; the input quantity of machine learning is substituted into a machine learning module to predict the abrasion state of the grinding wheel; if the set abrasion loss is exceeded, the grinding wheel is refinished and fed back to the machine tool, otherwise, the abrasion state, the grinding wheel information, the grinding path and the spiral groove information of the grinding wheel are substituted into a module for calculating the front angle and core thickness error, the front angle and core thickness error is obtained, and whether the tolerance requirement is met or not is judged; under the condition that the front angle error and the core thickness error meet the requirements, the spiral groove is ground; the shape of a machined spiral groove is obtained multiple times in an image mode and fused into an accurate contour, then the shape of the grinding wheel is calculated according to the enveloping principle, and grinding wheel abrasion information estimated through machine learning is updated. According to the scheme, the service life of the grinding wheel can be remarkably prolonged, the grinding time of the spiral groove of the cutter is shortened, and efficient machining of the spiral groove of the hard alloy cutter is achieved.
Owner:HARBIN UNIV OF SCI & TECH

Multi-modal rumor detection system based on multi-domain perception

The invention discloses a multi-modal rumor detection system based on multi-field perception, relates to the technical field of rumor detection, and is used for solving the technical problem of low multi-field rumor detection precision in the prior art. The rumor detection system comprises a multi-view feature extraction module, a multi-domain perception module, a multi-level knowledge fusion module and a final classifier. Wherein the multi-view feature extraction module is used for respectively extracting corresponding single-mode features from a text mode and an image mode, and acquiring aligned image and text features as multi-mode view features; the multi-domain sensing module is used for learning cross-domain shared knowledge and domain specific knowledge at a text view angle, an image view angle and a multi-mode view angle; the multi-level knowledge fusion module is used for fusing cross-domain shared knowledge and domain specific knowledge of a text view angle, an image view angle and a multi-mode view angle; and the final classifier is used for obtaining a prediction result according to the cross-domain shared knowledge and the domain specific knowledge obtained by the multi-level knowledge fusion module.
Owner:CHONGQING UNIV

Large model image segmentation method and system based on multi-scale fusion, and medium

The invention discloses a large model image segmentation method and system based on multi-scale fusion and a medium, and relates to the technical field of image processing. Obtaining an image segmentation model; the structure construction mode of the image segmentation model is that a traditional SAM is used as a network basis, an Adapter module is inserted between a multi-head attention layer and a multi-layer perceptron in each VIT block of an image encoder of the SAM, and a four-layer CBR-Net module is arranged at the end of the image encoder in parallel. The output end of the image encoder and the back of an up-sampling layer of a mask decoder of the SAM are respectively connected with an Atte-FFB module in series, and residual connection is constructed between the first three-layer output of the CBR-Net module and the Atte-FFB module of the corresponding layer in the mask decoder of the SAM network; residual connection is constructed between the fourth layer output of the CBR-Net module and an Att-FFB module behind the output end of the image encoder; and performing image segmentation on the to-be-segmented image through the image segmentation model. According to the method, the segmentation precision is improved, and the model generalization ability of a cross-image mode is enhanced.
Owner:XI AN JIAOTONG UNIV

Multi-point confocal image scanning microscope and imaging method

PCT designated stage expiredWO2025119194A1MicroscopesData setDepth imaging
The present application relates to the technical fields of optical elements, systems, instruments and imaging, and discloses a multi-point confocal image scanning microscope and an imaging method. According to the present application, an imaging light source assembly, a light source collimation assembly, a multi-focus generation assembly, a multi-focus moving assembly, a multi-focus projection assembly, and a camera detection assembly are arranged along a light path, so as to achieve complete image recording. A unique multi-focus illumination mode provided by the present application is used in combination with an optical locked-phase detection technology, so that a defocus-induced stray signal during thick sample imaging can be effectively removed, and depth imaging is performed; the multi-focus generation assembly is switched, so that switching between a wide-field mode, a confocal mode and a super-resolution mode is easily achieved. Moreover, the present application uses a pixel redistribution algorithm to perform super-resolution reconstruction, and uses a multi-image deconvolution reconstruction algorithm and redundant information in a raw data set to perform frame reduction reconstruction, thereby increasing the imaging speed. Compared with the existing methods, the present application can increase the imaging speed, improve the imaging depth, reduce phototoxicity, and widen the diversity of imaging modes.
Owner:PEKING UNIV +1

Infrared and visible light image fusion method and system based on text-guided semantic perception

The invention discloses an infrared and visible light image fusion method and system based on text-guided semantic perception, and relates to the technical field of image fusion. A text prompt is obtained by using a reference image subjected to semantic segmentation, and the features of a source image are respectively obtained by using encoders with the same structure and independent parameters. A semantic correspondence alignment module is provided between an encoder and a decoder, and rich feature representations related to text features are obtained, so that the difference between a text mode and an image mode is reduced. In a decoder stage, a semantic modulation module is provided, interaction between a text and an image is fully realized, effective combination of texture details and advanced semantic information is realized in combination with shallow layer features, and finally a fused image with rich semantics and complete details is generated. According to the method, the overall quality of the fused image is remarkably improved in the aspects of semantic consistency and perception quality.
Owner:DALIAN UNIV

Tea quality evaluation method based on spectral image bimodal fusion

The invention discloses a tea quality evaluation method based on spectral image bimodal fusion. The tea quality evaluation method comprises the following steps: acquiring hyperspectral data of different grades of tea; preprocessing the tea hyperspectral image data; acquiring spectrum and image modal information of the tea hyperspectral image; a spectrum-image dual-mode fusion neural network (SIFNet) model is established, and the SIFNet adopts a dual-branch network architecture. A spectrum feature extraction branch is combined with a one-dimensional convolutional neural network and a bidirectional long-short-term memory network module to extract time sequence dependence features of a spectrum mode; an image feature extraction branch introduces a two-dimensional convolutional neural network and a fuzzy logic processing module to extract accurate features and fuzzy features of an image modal, and advantage complementation of multi-modal information is realized through a feature depth fusion technology; the SIFNet model is trained; and evaluating the quality of the tea to-be-detected sample. Through deep coupling and complementary enhancement of the spectrum-image bimodal features, the tea quality grading accuracy is improved.
Owner:JIANGSU UNIV +1

Joint semantic segmentation method for camera and laser radar in cross-country environment

The invention discloses a joint semantic segmentation method for a camera and a laser radar in an off-road environment, and relates to the technical field of computer vision. According to the method, the 2D image and the 3D point cloud data are respectively processed by constructing the double-branch basic segmentation network, so that efficient processing and fine segmentation of the point cloud are realized; a multi-scale one-way knowledge distillation framework is designed, feature fusion of a common-view area is realized through a channel self-attention mechanism, and knowledge migration from an image mode to a point cloud mode is realized by adopting a dynamic temperature adjustment and Logit standardization technology; in the training stage, multi-modal data is utilized, and high-precision segmentation can be completed only through point cloud input in the reasoning stage. According to the method, the semantic segmentation precision and robustness in the cross-country environment are remarkably improved, the segmentation performance of the occlusion region and the sparse feature region is effectively improved, and meanwhile, the computing resource consumption is greatly reduced.
Owner:CHONGQING UNIV

Brain load assessment method and system based on hemodynamic information and electroencephalogram signals

The invention discloses a brain load assessment method and system based on hemodynamic information and electroencephalogram signals, and relates to the field of brain-computer interface classification systems. The technical problem that how to fuse electroencephalogram signals and hemodynamic information and comprehensively utilize complementary characteristics of the electroencephalogram signals and the hemodynamic information in time and space dimensions to achieve accurate recognition and dynamic evaluation of the mental load state in the prior art is urgently needed to be solved in the current brain-computer interface field is solved. The method comprises the following steps: acquiring EEG data and fNIRS data of a testee; performing preprocessing, selecting a suitable frequency band, removing interference such as high-frequency noise and low-frequency baseline drift, and constructing a corresponding training data set; respectively constructing an EEG single-mode network and an fNIRS single-mode network which are used for extracting related characteristics; and constructing an EEG-fNIRS multi-mode classification network for fusing data of brain imaging modes in the EEG single-mode network and the fNIRS single-mode network, and performing classification evaluation on the hemodynamic information of the subject and the brain power load data of the electroencephalogram by using the EEG-fNIRS multi-mode classification network.
Owner:HARBIN INST OF TECH

Question and answer data construction method and device

A question and answer data construction method comprises the following steps: 1) reading samples containing image and text data in batches, constructing a unified instruction for each batch of samples, and calling a large language model to extract candidate questions which have discriminability and teaching value and can be answered from images to form multiple groups of candidate questions; (2) carrying out structured deduplication merging on the cross-batch candidate problems obtained in the step (1) to obtain a deduplication problem pool; 3) for each sample, identifying an available image mode of the sample, screening candidate problems matched with the mode from the problem pool, and randomly extracting candidates of which the number is greater than a target value q; 4) constructing a request only based on image answering, calling the multi-modal large language model by taking the sample image as input, and generating a question-answer pair; if reliable answers of the candidate questions cannot be obtained only through the images, replacing the candidate questions with new candidates and retrying until q question and answer pairs capable of being answered are accumulated and obtained; 5, the generated question and answer pairs are subjected to structured verification and safe disking.The method supports parallel processing and breakpoint continuous running so as to meet the large-scale data construction.The method has the advantages of being good in universality and expandability and standardized in result.
Owner:ZHEJIANG UNIV

User guidance in ultrasound imaging

The invention provides an apparatus for providing guidance to a user of an ultrasound acquisition system to acquire standardized images of a target anatomical structure of a patient. A processing unit is configured to toggle a user interface between two modes: a guidance mode in which only guidance information is presented on the user interface, and an imaging mode in which a received ultrasound image is presented on the user interface. The processing unit toggles the user interface between these two modes based on a spatial relationship between a current probe position and a target view. When the ultrasound acquisition system is within a threshold distance of the target view then the processing unit toggles the user interface to display the imaging mode, otherwise the guidance mode is displayed.
Owner:KONINKLIJKE PHILIPS NV

Private domain community-oriented text and image multi-mode user interest identification method and system

The invention discloses a text and image multi-mode user interest identification method and system oriented to private domain communities. The system comprises a text feature extraction module, an image feature extraction module, a multi-modal fusion module, an interest probability aggregation module and a dynamic threshold judgment module. The multi-modal fusion module adopts a gating mechanism to realize self-adaptive fusion, a gating factor is automatically adjusted according to the characteristics of an input sample, and weighting is carried out between a text mode and an image mode, so that the robustness of a fusion result can still be ensured when the quality of information in different modes is unbalanced. And the interest probability aggregation module performs time sequence smoothing on the instant interest probability through an exponential weighted moving average method, and inhibits short-time noise interference. And the dynamic threshold value judgment module is used for calculating a dynamic threshold value according to the mean value and the standard deviation of the recent window, comparing the smoothed interest probability with the dynamic threshold value, and judging that the corresponding interest label is activated when the smoothed interest probability is greater than or equal to the dynamic threshold value. According to the method, the user interests can be accurately recognized in a complex and changeable private domain community environment, and the adaptability, stability and real-time performance of interest recognition are achieved.
Owner:北京娱广科技有限公司

Systems and methods for inspecting a worksurface

A method of evaluating a surface is presented that includes imaging the surface, with an imaging system. Imaging includes providing a camera of the imaging system proximate the surface. Imaging also includes causing the imaging system and the surface to move relative to each other, such that a distance between the imaging system and the surface is substantially maintained. Imaging also includes capturing image data of the surface. The image data is captured in a near dark field mode or a dark field image mode. The method also includes analyzing the image data and detecting a topography and / or appearance of the surface. The method also includes generating an evaluation regarding the surface based on the detected topography and / or surface appearance.
Owner:3M INNOVATIVE PROPERTIES CO

Light field time domain super-resolution method based on event camera

The invention discloses a light field time domain super-resolution method based on an event camera, and belongs to the field of optical engineering. A light field camera and an event camera are used for collecting light field event data, decoding the light field data, denoising the event data and generating an event image, the light field data and an event grayscale image are subjected to joint calibration in a sub-aperture image mode, joint optimization is carried out through calibration results of multiple viewpoints, a checkerboard is shot, and a target object is obtained. And obtaining an affine transformation matrix and a perspective transformation matrix, and obtaining a super-resolution result of projecting the event image onto a light field image plane. Starting from the expression mode of the light field macro pixel, the method has applicability to light field data of any macro pixel size; compared with a traditional time domain super-resolution method, the time domain super-resolution of the light field camera at any frame rate can be realized by segmenting the event stream at different scales.
Owner:STATE GRID ZHEJIANG ELECTRIC POWER CO LTD HANGZHOU POWER SUPPLY CO

Picture mode resolution enhancement for e-beam detector

A charged particle detector includes a plurality of sensing elements, with each sensing element being further divided into sub-sensing elements. The sub-sensing elements may be individually addressed during high-resolution image acquisition in a picture mode, and may be grouped together during high speed detection in a beam mode. The arrangement allows a selectable tradeoff between speed and resolution without introducing significant parasitic parameters.
Owner:ASML NETHERLANDS BV

Game data processing method and device, electronic equipment and storage medium

Embodiments of the present application provide a game data processing method and device, electronic equipment and readable storage medium, the method comprises: in the game process, the terminal can obtain game resources, wherein the game resources can include scene resources, model resources and obstacle information for scene resources, model display parameters for model resources, then the corresponding game scene of the scene resources can be displayed in the form of two-dimensional image mode, and the corresponding game model of the model resources can be displayed in the form of three-dimensional image mode in the game scene according to the model display parameters, then the game scene can be processed according to the obstacle information, the corresponding obstacle map is generated, and the moving information of the game model is obtained, and then the target moving path of the game model in the obstacle map is determined according to the moving information, so as to control the game model to move in the game scene.
Owner:BEIJING 58 INFORMATION TTECH CO LTD

automobile

An automobile includes an entertainment execution device, a display device and a mode select controller. The entertainment execution device executes entertainment software, communicates with radio control equipment, or plays back a video content. The display device displays an image output from the entertainment execution device on a windshield or on a flexible display capable of being raised or lowered between a front windshield and a steering wheel. The mode select controller enables selection between an entertainment mode for operating the entertainment execution device and a driving mode for driving the automobile and executes control according to each mode. The mode select controller controls the display device to display the image output in a first region of the windshield in the driving mode, and to display the image output in a second region of the windshield or on the flexible display in a raised state in the entertainment mode.
Owner:NISSAN MOTOR CO LTD