Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

125 results about "Supervised clustering" patented technology

Further quoting from the article: Supervised clustering is the task of automatically adapting a clustering algorithm with the aid of a training set consisting of item sets and complete partitionings of these item sets..

Network security big data state evaluation method based on pattern recognition

The invention relates to the technical field of network security, in particular to a network security big data state evaluation method based on pattern recognition, which comprises the following steps of: extracting multi-modal features from a network flow log, a system event log, a host behavior log and threat intelligence data, generating a feature matrix, performing feature dimensionality reduction by adopting an auto-encoding network, and obtaining a network security big data state evaluation result; carrying out attack behavior classification and abnormal mode identification in combination with unsupervised clustering and a graph neural network; constructing an attack transition probability matrix based on a Markov model; forming a time sequence attack chain; predicting an attack development trend; and a dynamic protection instruction is issued to the safety equipment. According to the method, the unknown attack detection capability can be improved, the time sequence attack traceability is enhanced, the security situation assessment is optimized, and the method is suitable for security situation awareness in cloud computing, industrial internet and large-scale network environments.
Owner:SHANDONG ENERGY GRP CO LTD +1

Cluster-based histopathology phenotype representation learning by self-supervised multi-class token hierarchical vision transformer

The system and method for processing a digital pathology image using a machine learning model that includes a self-supervised hierarchical Vision Transformer (ViT) configured to perform unsupervised clustering with multiple classification tokens. The method includes receiving a digital pathology image that depicts a tissue slice stained with histological dyes. The digital pathology image may be processed to generate a result comprising multiple predicted classifications of individual patches of the digital pathology image. The result is generated by a machine-learning model using a self-supervised hierarchical Vision Transformer (ViT) that may further comprise a multi-head self-attention module configured to predict a crosspatch relevance metric using an attention mechanism for each individual patch in the digital pathology image thereby assigning the individual patches to a cluster based on the crosspatch relevance metrics.
Owner:VENTANA MEDICAL SYSTEMS INC

Data annotation method and system based on user behavior and attention tracking

The invention discloses a data labeling method and system based on user behaviors and attention tracking, and the method comprises the steps: synchronously collecting multi-source behavior signals of a mouse, a keyboard, eye movement and the like of a doctor in real time, combining identity and interface metadata, and carrying out the standardized normalization, abnormality elimination and short time sequence behavior unit division. And extracting individual behavior micro-modes by using unsupervised clustering, and constructing a behavior portrait library. Through multi-modal time sequence modeling and a self-adaptive space-time attention mechanism, behavior characteristics, an interface area and a report text are deeply fused, a multi-level correlation probability is output, and high-precision automatic tagging of content and an image area is realized.
Owner:GUANGZHOU FANGXIN MEDICAL TECH CO LTD

Flue gas treatment intelligent regulation and control method and system based on artificial intelligence and medium

The invention relates to the technical field of data processing, and discloses a flue gas treatment intelligent regulation and control method and system based on artificial intelligence and a medium. The method comprises the following steps: collecting parameter data of a flue gas system through a sensor network and processing to obtain preprocessed data; using unsupervised clustering to identify working condition features and categories; constructing an emission concentration error active switching control model according to the working condition information; different working condition control parameters are optimized by adopting a differential evolution algorithm; optimized parameters are input into a controller, and modes are dynamically switched according to pollutant errors. The control strategy can be automatically switched according to the working condition characteristics of the flue gas treatment system, the complex and changeable industrial production environment can be effectively dealt with, and the system operation cost is optimized while stable and up-to-standard emission of pollutants is guaranteed.
Owner:BEIJING HANHAI QINGTIAN ENVIRONMENTAL PROTECTION TECH CO LTD

Assurance of user behavioral patterns in software applications with quasi-supervised clustering

Systems, methods, and other embodiments associated with quasi-supervised clustering for activity pattern characterization and anomalous activity detection are described. In one embodiment, a method generates a first sparse similarity matrix for nearest neighbors of a plurality of data points. The data points each characterize a pattern of activity associated with an account. The method generates a second sparse similarity matrix for random neighbors of the plurality of data points. The method recursively clusters the plurality of data points based on the first sparse similarity matrix. The method quasi-supervises the recursive clustering based on the second sparse similarity matrix to stop the iterative clustering when the data points are split into N clusters. The value of N is not pre-determined. The method detects that the individual data point has changed clusters, indicating anomalous activity. And, the method generates an electronic alert that the anomalous activity is associated with the account.
Owner:ORACLE INT CORP

Radiation source new individual identification method based on signal enhancement double contrast learning

The invention discloses a radiation source new individual identification method based on signal enhancement double contrast learning, and belongs to the field of software radio. According to the method, a multi-physical domain enhancement operator sequence is adopted, so that the risk that feature distortion is possibly introduced in traditional data enhancement is overcome, and a high-quality feature discrimination basis is provided for subsequent comparative learning by constructing positive and negative sample pairs with strong correlation. A feature decoupling architecture of double comparative learning is adopted, and the category discrimination advantage of supervised comparative learning and the feature decoupling capability of unsupervised comparative learning are organically fused. Dynamic semi-supervised clustering is realized by adopting a dual decision clustering mechanism, firstly, the number of potential categories is determined by utilizing density peak clustering DPC, secondly, a semi-supervised decision module is introduced to realize known category accurate identification and unknown category adaptive discovery of radiation source individuals, and then new radiation source individual identification is realized. The method is suitable for the field of software radio, and high-precision radiation source new individual identification is realized.
Owner:BEIJING INST OF TECH

Safety management method and system based on artificial intelligence and big data

The invention relates to the technical field of safety monitoring artificial intelligence, and discloses a safety management method and system based on artificial intelligence and big data. The method comprises the steps that multi-modal real-time monitoring data is collected through a distributed sensor array, and a standardized sequence is generated after preprocessing; the system accesses a historical security event library, adaptively divides risk category clusters by using an unsupervised clustering algorithm, and extracts key risk features and influence domains for each cluster. On the basis, a plurality of special self-adaptive security analysis engines are configured to perform parallel analysis on real-time data to generate a plurality of security analysis feature maps. And carrying out weighted fusion and enhancement on the feature maps through a feature fusion network to obtain a unified security situation feature vector, and finally outputting a security level and a management and control instruction through a decision model. According to the method, autonomous identification, dynamic evaluation and accurate response of complex risks are realized, and the intelligent level of safety management and the decision accuracy are improved.
Owner:SHENYANG UNIV

Assurance of user behavioral patterns in software applications with quasi-supervised clustering

Systems, methods, and other embodiments associated with quasi-supervised clustering for activity pattern characterization and anomalous activity detection are described. In one embodiment, a method generates a first sparse similarity matrix for nearest neighbors of a plurality of data points. The data points each characterize a pattern of activity associated with an account. The method generates a second sparse similarity matrix for random neighbors of the plurality of data points. The method recursively clusters the plurality of data points based on the first sparse similarity matrix. The method quasi-supervises the recursive clustering based on the second sparse similarity matrix to stop the iterative clustering when the data points are split into N clusters. The value of N is not pre-determined. The method detects that the individual data point has changed clusters, indicating anomalous activity. And, the method generates an electronic alert that the anomalous activity is associated with the account.
Owner:ORACLE INT CORP

Optical module fine granularity health degree assessment method, system and device and medium

The invention discloses an optical module fine-grained health degree assessment method, system and device and a medium, and the method comprises the steps: data collection and preprocessing: carrying out the semi-supervised clustering, and preliminarily dividing the data into fault samples and normal samples; the method comprises the following steps: carrying out fine granularity division on a normal sample through a Transform auto-encoder; performing fine-grained division on the'fault 'samples through respective Gaussian mixture models of the'fault' samples and the completely healthy samples; judging whether the port state triggers a preset event, and if not, directly outputting an evaluation result; and if at least one preset event is triggered, updating the model for classification in each stage, carrying out health degree grade re-evaluation, and outputting an evaluation result. According to the method, a five-level health degree quantitative standard is established, the full life cycle state of the optical module from complete health to irreversible serious fault is covered, and the method can adapt to partial data drift caused by aging of the optical module.
Owner:ZHEJIANG LAB

Double-current comparison and DHI combined equipment fault detection method and system

The invention provides a double-current comparison and DHI combined equipment fault detection method and system, and belongs to the technical field of power equipment state monitoring and fault diagnosis. The method comprises a time sequence perception adversarial enhancement module, a generative adversarial network (GAN) data expansion module, a double-flow contrast attention network (D-CAN) feature extraction module and a dynamic health index (DHI) calculation module. The core lies in that a fault sample with time sequence correlation is supplemented through a physically constrained GAN, robust features are extracted by using a ResNet1D and Transform fused double-flow network, and a health state is quantified in combination with unsupervised clustering and mahalanobis distance. Through simulation and experimental verification, zero-delay detection of the early fault of the transformer can be realized, the output health index and the 3D visualization result can provide an accurate basis for operation and maintenance of the transformer, and the diagnosis accuracy, the data utilization rate and the dynamic adaptability of state evaluation are remarkably improved.
Owner:NANJING SAC RAIL TRAFFIC ENG CO LTD +1

Processing multiplex images and analysis of immune enriched spatial proteomic data

Techniques are disclosed herein that encompass image pre-processing and a semi-supervised clustering for optimization and analysis of immune-enriched single-cell proteomics data generated via multiplexed imaging technologies. This is achieved through an image pre-processing pipeline, which converts image data contained in one type of file (e.g., .mcd) into another type of file (e.g., .tiff) and removes artifact signals from the image data using various algorithms to generate improved image data. Thereafter, a semi-supervised clustering pipeline analyzes the improved image data using various techniques, including implementing a supervised algorithm to identify metaclusters such as general immune phenotypes (e.g., CD4−T-cells, Macrophages, Neutrophils, etc.) as well as non-immune phenotypes while implementing an unsupervised algorithm that enables the identification of specific subclusters and a more in-depth cellular status characterization.
Owner:UNIV OF SOUTHERN CALIFORNIA

Hierarchical clustering-based unsupervised silicon wafer surface defect detection method and device

The invention discloses an unsupervised silicon wafer surface defect detection method and device based on hierarchical clustering, relates to the field of image processing and computer vision, and utilizes an unsupervised clustering technology to detect silicon wafer surface defects. The method comprises the following steps: extracting multi-scale features from a plurality of normal silicon wafer images through a pre-trained convolutional neural network; integrating the multi-scale features of the plurality of pictures by adopting a hierarchical fusion strategy; constructing a feature memory library by using a clustering algorithm, and generating a clustering center; generating patch features for the test image through a consistent fusion process; through comparative supervision, normal patches are close to a clustering center, and abnormal patches are pushed away; calculating the distance between the patch and the clustering center, and judging abnormity; and mapping the position of the abnormal patch, and positioning scratches, cracks, pollution or recesses. According to the method, through an unsupervised mode, a large amount of defect labeling data is not needed, the data preparation cost is effectively reduced, and the accuracy and robustness of silicon wafer surface defect detection are remarkably improved.
Owner:TIANJIN UNIV

Instrument electric quantity intelligent calibration method and system

The invention discloses an intelligent calibration method and system for the electric quantity of an instrument, and relates to the technical field of electric energy metering of an electric power system.The intelligent calibration method comprises the steps that an electric meter network diagram with a main meter and sub-meters as nodes is constructed, and robust composite correlation indexes are calculated on the basis of reading increments in a sliding time window; the interference of common background load and instantaneous noise is suppressed by using a threshold clipping and main table purification mechanism, so that the depiction capability of the correlation graph on a physical connection relationship is improved fundamentally, and the phase to which the sub-table belongs can be identified more accurately in a stock area with unknown phase or incomplete topology; a graph neural network model is introduced based on the network graph, electric meter embedding representation with physical meaning is obtained through neighborhood information aggregation, phase grouping is achieved in combination with unsupervised clustering, split-phase energy balance loss is designed in the training stage, and the difference between the reading of a main meter and the sum of the reading of sub-meters in each phase group is used as a constraint. Alignment of the embedding space and the energy conservation relation is pushed in a self-supervised mode.
Owner:BEIJING ZHONGBAO HUATONG TECH CO LTD

Academic community discovery and analysis method driven by citation network

The invention discloses an academic community discovery and analysis method driven by a citation network, and the method comprises the steps: constructing the citation network, constructing an adjacent matrix A and a node attribute matrix X according to the citation network, and extracting a core sub-network information matrix S based on a k-core algorithm; constructing a dual-channel sparse graph attention auto-encoder, respectively taking (X, A) and (S, A) as input of the two channels to learn low-dimensional embedding of the two channels, and obtaining joint embedding Z through a dynamic weighted fusion strategy; reconstructing adjacent matrix information by adopting an inner product decoder, reconstructing node attribute characteristics and core sub-network information, and calculating reconstruction loss; and inputting the joint embedded Z into a self-supervised clustering module, performing joint optimization on the information embedding and reconstruction module and the self-supervised clustering module, and finally outputting a group division result of the papers / authors. According to the method, complementary fusion of attributes and a core structure is realized, robustness in noise and sparse scenes is enhanced, and discriminability and stability of group division are improved.
Owner:XIAN UNIV OF TECH

Automatic driving test scene generation method and device based on text collision report, equipment and medium

The invention discloses an automatic driving test scene generation method and device based on a text collision report, equipment and a medium, and relates to the technical field of automatic driving, and the method comprises the steps: obtaining an unstructured text collision report, converting the unstructured text collision report into a high-dimensional numerical vector through preprocessing and feature extraction, and obtaining a high-dimensional numerical vector; a typical collision mode is identified and classified by using an unsupervised clustering algorithm, a logic scene is further constructed in a parameterized manner, physical feasibility is verified through a kinematics verification module, risk levels are divided, and a risk scene library is formed. And finally, generating an automatic driving test case set by adopting a stratified sampling strategy. By using the collision text data, real, diversified and physically feasible test scenes are generated, and the efficiency and accuracy of automatic driving safety assessment are improved.
Owner:CENT SOUTH UNIV

Edge missing network community detection method based on dual-channel deep embedding clustering

The invention provides an edge deletion network community detection method based on dual-channel deep embedding clustering, and the method comprises the steps: firstly proposing three edge deletion strategies which are used for simulating an edge deletion condition possibly occurring in an actual network; then, an edge enhancement algorithm based on random walk is put forward to recover missing edge information under different edge deletion strategies; a clustering model based on double-channel depth embedding is introduced and comprises an information embedding core module and a self-supervised clustering core module. According to the method, the missing edge condition in a real network is simulated by providing three different edge deletion strategies, and the missing edge is effectively recovered in combination with an edge enhancement algorithm based on random walk, so that the robustness and accuracy of community detection in an edge missing scene are greatly improved. Network core node information and high-order neighbor information are respectively subjected to embedded representation, node features with higher distinction degree are obtained through a dynamic weighted fusion strategy, and more sufficient information support is provided for subsequent community division.
Owner:XIAN UNIV OF TECH

Geothermal fluid source discrimination method based on geochemical index and machine learning fusion

The invention discloses a geothermal fluid source discrimination method based on geochemical index and machine learning fusion, and relates to the technical field of geothermal resource intelligent exploration, and the method comprises the steps: firstly collecting and calculating multi-dimensional geochemical indexes, and constructing a standardized feature sequence; establishing an end member feature library by using unsupervised clustering; identifying the source type of the fluid through a discrimination model fused with an attention mechanism; aiming at the mixed source fluid, constructing an optimization model embedded with physical and chemical constraints, and quantitatively inverting the contribution proportion of each end member; and finally, on the basis of the time sequence proportion data obtained by inversion, predicting a future evolution trend by adopting a time convolutional network-long and short-term memory network model fused with an attention mechanism. According to the geothermal fluid source discrimination method based on geochemical index and machine learning fusion provided by the invention, the whole-process intelligent analysis of the geothermal fluid source from qualitative identification to quantitative prediction is realized.
Owner:CHINA UNIV OF GEOSCIENCES (BEIJING)

Low-temperature economizer digital twinborn body construction method

The invention relates to the technical field of computers, in particular to a low-temperature economizer digital twinborn body construction method, and aims to solve the problems that an existing model statically solidifies, multi-source data fusion is difficult, and degradation recognition lags. The method comprises the following steps: constructing a multi-physical field reference model covering fluid, heat transfer and corrosion mechanisms; collecting and carrying out time-space alignment on temperature, pressure, flue gas components and ash deposition data, and combining wavelet and median filtering to carry out de-noising; introducing extended Kalman filtering to correct model state variables on line, and realizing dynamic updating of parameters; a CNN-GRU hybrid network is embedded to extract time sequence degradation features, and unsupervised clustering is combined to identify working condition states; and when the detection is abnormal, triggering local high-fidelity CFD re-simulation, and completing closed-loop optimization and state synchronous mapping. Through the mechanism, the internal physical field error lt of the equipment is realized; 5% high-precision dynamic mapping, more than 95% of anomaly detection rate and rapid deployment of a new unit within 72 hours are realized, the prediction accuracy and preventive maintenance capability are remarkably improved, and the service life of equipment is prolonged by 15%-20%.
Owner:HUANENG SHANTOU HAIMEN POWER GENERATION CO LTD

Intelligent color box printing color analysis control method

The invention relates to an intelligent color box printing color analysis control method, and provides a reinforcement learning decision closed-loop control method integrating distributed color detection, space-time multi-modal data processing, feature extraction and self-supervised clustering and multi-scale reward modeling. Multi-variable normalization and synchronization are completed through multi-point real-time collection of color, process and environment data, color change meta-state features are extracted based on a self-encoding neural network, and a meta-state distribution model related to space and working conditions is established through self-supervised clustering. And the color risk is evaluated in combination with a multi-scale reward function, so that color parameter optimal adaptive adjustment and cooperative control driven by reinforcement learning are realized. In an actual control link, the system can dynamically optimize a meta-state aggregation and strategy network structure based on a feedback signal, a stable and self-adaptive color consistency closed-loop control system is constructed, and the color consistency and process intelligence level of a printed product under a complex working condition are effectively improved.
Owner:ZHUHAI DINGHUI PRINTING CO LTD

Scene complexity quantitative characterization method for testing automatic driving automobile

The invention provides a scene complexity quantitative characterization method for testing an automatic driving automobile, and belongs to the technical field of automatic driving testing. The method comprises the following steps: establishing a scene model and a rare target object model based on data-mechanism hybrid modeling; controllable reproduction of a complex scene is realized through physical simulation technologies such as video injection and echo simulation; constructing a causal graph structure to extract scene high-dimensional features; self-adaptive extraction, derivation and generalization of a scene are realized by using unsupervised clustering, a generative adversarial network and an optimization search algorithm; developing a data conversion algorithm to generate a standardized test scene library; analyzing complexity mapping relations between scene elements and perception, decision and execution systems, and establishing scene complexity quantitative representation; and finally, constructing a multi-dimensional evaluation system fusing safety and anthropomorphism, and realizing integrated performance evaluation of the automatic driving automobile fusing scene complexity. The scene complexity is converted into a quantifiable index from a qualitative concept, and the scientificity of testing and the fairness of evaluation are improved.
Owner:CHANGCHUN AUTOMOTIVE TEST CENT

Intelligent automatic drawing system and method

The invention discloses an intelligent automatic drawing system and method, and relates to the technical field of agricultural remote sensing monitoring. The method comprises the following steps: acquiring a multi-temporal remote sensing image of a farmland area, performing preprocessing and spatial alignment, dividing the farmland into spatial units, extracting time sequence spectrum and vegetation index characteristics of the spatial units, and forming a visual characteristic vector sequence; analyzing the sequence by using unsupervised clustering, and automatically identifying the time sequence evolution trajectory categories representing different growth states; and finally, the system automatically marks the space units which are judged to be abnormal categories on a map, generates early warning and outputs a farmland abnormity thematic map. The system comprises a data preprocessing module, a data preprocessing module, a track classification module and an abnormal graph generation module. According to the invention, full-automatic and intelligent dynamic monitoring and abnormity identification of the farmland growth state are realized, the limitation of traditional single-temporal analysis or artificial visual interpretation is overcome, and the monitoring efficiency and objectivity are significantly improved.
Owner:CHANGCHUN PLANNING PREPARATION RES CENT (CHANGCHUN URBAN & RURAL PLANNING & DESIGN INST)

Automatic identification, classification and development trend analysis method of net red villages based on multi-source data fusion and natural language processing

The method for automatic identification, classification and development trend analysis of net red villages based on multi-source data fusion and natural language processing comprises the following steps: UGC data is crawled from Xiaohongshu and Douyin through a distributed master-slave architecture, de-duplicated based on SimHash, and normalized in time and coding format; a text semantic fingerprint is generated, and multi-level semantic cache fingerprint matching is performed; for unassigned text, its complexity is calculated, and a large language model API is adaptively called to automatically complete and extract five-level administrative divisions; weights are determined based on the analytic hierarchy process, interaction indicators such as likes, comments, collections and forwards are integrated, and a comprehensive network heat index of the village is obtained; an external text mining tool is connected, and batch word frequency analysis, semantic network analysis and sentiment tendency evaluation are performed; a document-term matrix is constructed, TF-IDF weighting is performed, and unsupervised clustering algorithm is used for clustering analysis of village characteristics; cross-dimension analysis is performed on the clustering results, and a development portrait, advantage mining and operation suggestion warning are automatically generated in combination with the SWOT model.
Owner:ZHEJIANG UNIV OF TECH

A charging facility fault early warning model construction method based on multi-source data perception

The application provides a charging facility fault early warning model construction method based on multi-source data perception, and belongs to the technical field of charging facility fault early warning.The application forms a multi-source time series data set by collecting charging pile operation state data and environmental data and performing time series alignment, expands fault samples by using an adversarial generative network and a semi-supervised clustering algorithm to form a balanced sample data set, constructs a fault early warning model containing a time series memory pool encoding layer and a frequency domain convolution feature extraction layer, cooperatively optimizes early warning accuracy and response time delay by using a double-layer game optimization framework to obtain an optimal parameter combination, deploys the optimized model to an edge computing module to realize real-time fault early warning, establishes an online incremental learning mechanism based on sliding time window detection data distribution offset, and uses an elastic weight consolidation technology to update model parameters and retain historical knowledge, thereby solving the problem of early warning performance degradation of charging facility fault early warning in a data distribution evolution environment.
Owner:CHINA CONSTR EIGHTH BUREAU DEV & CONSTR CO LTD

A medical term normalization method

The application relates to a medical term normalization method, and belongs to the technical field of data processing. The method solves the problems that coverage cannot be guaranteed and new variations cannot be dynamically adapted in the prior art. The method comprises the following steps: acquiring historical unmatched original words to construct an original word set; acquiring terms labeled for part of the original words in the original word set; adopting a supervised clustering algorithm to cluster the original word set to obtain original word clusters; for each original word cluster, an improved genetic algorithm is adopted to obtain a regular expression for mapping the original word cluster to a term corresponding to the original word cluster. The method realizes the improvement of coverage, can automatically discover new variations, and dynamically adapts new variations.
Owner:BEIJING YIYONG TECH CO LTD

Open set fault diagnosis method based on two-stage entropy perception consensus domain confrontation framework

The invention relates to an open set fault diagnosis method based on a two-stage entropy perception consensus domain confrontation framework, and the method comprises the following steps: S1, obtaining a training data set of a rotating machine, the training data set comprising a source domain data set and a target domain data set; s2, constructing a model framework, wherein the model framework comprises a feature extractor and loss functions of five modules; the feature extractor comprises a convolutional neural network, a bidirectional long-short-term memory network and a multi-head attention mechanism which are arranged in sequence; the five modules comprise an original classifier module, a two-stage entropy sensing module, an auxiliary classifier module, a self-supervised clustering module and a consensus domain confrontation module; inputting the training data set into a feature extractor, and adjusting parameters of a model framework by using an optimization algorithm to complete training of the whole framework; and S3, obtaining a test data set, inputting the test data set into the trained model framework, and outputting a fault diagnosis result. The problem that an existing domain self-adaption method cannot effectively process unknown faults is solved, and the open set fault diagnosis method is achieved.
Owner:TAIYUAN UNIVERSITY OF SCIENCE AND TECHNOLOGY

Pineapple quality grading method and system based on image processing and medium

The invention provides a pineapple quality grading method and system based on image processing and a medium, and relates to the technical field of pineapple quality grading. The method comprises the following steps: firstly, constructing a pineapple sample data set containing different quality gradients, collecting image group data of each pineapple sample, and extracting the color, size, defect, sugar and other characteristics of each pineapple sample; secondly, performing feature clustering on the samples by using an unsupervised clustering algorithm, and determining the quality gradient of the samples; and constructing a convolutional neural network model, taking the image data as input for training, and taking the corresponding defect features as labels to construct a defect detection model. Thirdly, collecting image data of pineapples to be graded, obtaining defect features by applying the trained model, and analyzing related features of the defect features; and finally, calculating a characteristic offset index between the to-be-graded pineapple and the sample pineapple, and determining the quality gradient of the to-be-graded pineapple according to the quality of the sample pineapple corresponding to the minimum offset.
Owner:SOUTH SUBTROPICAL CROP RES INST CHINA ACAD OF TROPICAL AGRI SCI +1

User behavior prediction method and device based on multi-source feature fusion, equipment and storage medium

PendingCN121234172AData sourceEngineering
The invention discloses a user behavior prediction method and device based on multi-source feature fusion, equipment and a storage medium, and the method comprises the steps: collecting user behavior data from different data sources, carrying out the feature extraction of the user behavior data, and obtaining the user behavior features corresponding to each data source; determining an importance weight corresponding to each data source according to a current prediction demand, and performing weighted fusion on each user behavior feature based on the importance weight to obtain a fused user feature; a preset unsupervised clustering algorithm is adopted, dynamic grouping processing is carried out according to the fused user features, and a plurality of user groups are generated; and constructing a corresponding behavior prediction model based on each user group, wherein the behavior prediction model is used for performing behavior prediction on the new user based on the current prediction demand. By dynamically adjusting the weight and fusing the multi-source features, the accuracy of dynamic grouping of the users can be improved, and the prediction model is specifically constructed based on different groups, so that high-precision user behavior prediction is realized.
Owner:SHENZHEN AOTIAN COMM CO LTD

A semi-supervised clustering method and its open-ended question answering text encoding method

This invention relates to the field of data representation technology, and discloses a semi-supervised clustering method and its open-ended question answer text encoding method. The semi-supervised clustering method includes: acquiring a dataset to be clustered, its labeled dataset, and its unlabeled dataset; mapping the dataset to be clustered to a spatial density map and / or a topological density map; and using the labeled and unlabeled datasets to cluster the data in the dataset to be clustered into several clusters, wherein each data point in each cluster has a cluster label, which can be an existing label or a new label. This invention uses a semi-supervised clustering method to efficiently and accurately cluster open-ended question answer text data, and can discover new classes with limited prior knowledge. After clustering, keywords can be extracted and encoded from the open-ended question answer text of each class, facilitating a quick and accurate understanding of the situation of the interviewed group and improving the efficiency and quality of diagnosis and treatment.
Owner:RENMIN UNIVERSITY OF CHINA

Oral lichen planus molecular subtype based on serum protein marker

The invention discloses an oral lichen planus molecular subtype based on a serum protein marker. The molecular typing is based on a multi-center clinical queue, a new molecular typing is obtained by using transcriptome sequencing data through a Seed-Kmeans semi-supervised clustering algorithm, and a serum marker is further screened out through multiple omics data including the transcriptome sequencing data, proteome sequencing data and serum enzyme-linked immunosorbent assay experimental data. The method can be used for further judging molecular typing. The molecular typing can obviously improve the curative effect of the oral lichen planus medicine, and is beneficial to the realization of OLP individualized precision medical treatment.
Owner:THE STOMATOLOGIAL HOSPITAL OF ZHEJIANG UNIV SCHOOL OF MEDICINE +1

Automated meta-learning in clustering using a machine learning clustering meta learning model

A method of creating a machine learning clustering meta learning model for use in solving a machine learning clustering problem includes obtaining a plurality of information related to the machine learning clustering problem, wherein the plurality of information includes classification datasets, machine learning transformers and clustering estimators, creating a set of clustering datasets using the classification datasets, generating trained clustering pipelines by training the set of unsupervised clustering pipelines responsive to the clustering datasets, processing the trained clustering pipelines to generate internal scores and external scores for the set of clustering datasets, creating an encoded clustering pipeline by encoding the trained clustering pipeline using the external score as a label and generating a trained supervised machine learning model by combining the internal scores and the encoded clustering pipelines.
Owner:INTERNATIONAL BUSINESS MACHINE CORPORATION