Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

754 results about "Forward propagation" patented technology

Engineering safety progress intelligent monitoring method based on multi-source data collaboration

The invention discloses an engineering safety progress intelligent monitoring method based on multi-source data collaboration, which relates to the technical field of intelligent engineering monitoring, and comprises the following steps of: mapping multi-source engineering monitoring data into nodes and edges of a graph in real time by utilizing an incremental graph updating algorithm, generating a dynamic knowledge graph, and generating a dynamic mapping result; performing graph traversal on the dynamic knowledge graph through an association subgraph extraction algorithm, extracting a security event matrix and a project progress matrix, calculating an SPI index value by using a dynamic weighted fusion algorithm, synchronously performing multi-threshold interval grading on the SPI index value, generating an SPI early warning level, performing state coding on the SPI early warning level, and generating a comprehensive feature vector; according to the method, data of different types can be effectively integrated and a basis is provided for formulating a targeted engineering safety progress solution through an incremental graph updating algorithm and a Bayesian causal graph model, and node probability distribution characteristics are extracted by using a forward propagation layer. The method is advantaged in that the incremental graph updating algorithm and the Bayesian causal graph model are utilized to effectively integrate data of different types and provide a basis for formulating a targeted engineering safety progress solution.
Owner:SHAANXI HUISHENG SPACE-TIME INFORMATION TECH CO LTD

Training method, data processing method, electronic equipment and computer readable storage medium

The invention relates to a training method, a data processing method, electronic equipment and a computer readable storage medium, and relates to the technical field of computers. The training method comprises the steps that a training sample is input into a to-be-trained model, forward propagation is executed, output of the to-be-trained model is obtained, the to-be-trained model comprises a plurality of expert layers, and a specific activation value is not stored in the forward propagation process; determining the value of a loss function according to the output of the to-be-trained model; according to the value of the loss function and the loss function, back propagation is executed, in the back propagation process of each expert layer, a recalculation task of a specific activation value and a first communication task are executed in an overlapping mode, and for each expert layer, the first communication task comprises a communication task of data interaction among multiple devices corresponding to the expert layer.
Owner:BYTEDANCE TECHNOLOGY CO LTD +1

Semantic alignment-based large language model equipment life prediction method and system

The invention belongs to the technical field of industrial equipment life prediction, and discloses an equipment life prediction method and system of a large language model based on semantic alignment, and the method comprises the steps: obtaining original multi-dimensional sensor time sequence data, and obtaining an embedded matrix after preprocessing; constructing a prompt text with domain semantics to obtain a natural language embedded representation; constructing a semantic text prototype, and realizing alignment of the embedding matrix and the semantic text prototype to obtain a patch embedding sequence; and splicing the patch embedding sequence and the natural language embedding representation, inputting the spliced patch embedding sequence and the natural language embedding representation into a pre-trained large language model for forward propagation, extracting hidden vectors output corresponding to the patch embedding sequence, splicing and flattening the hidden vectors into a single vector, and outputting to obtain an equipment life prediction result. According to the method, cross-modal knowledge learned by the LLM in large-scale pre-training and the powerful reasoning ability are fully utilized, and accurate prediction of the residual life of the equipment is achieved.
Owner:QILU UNIVERSITY OF TECHNOLOGY (SHANDONG ACADEMY OF SCIENCES) +1

Method for predicting workability of crane ship based on BP (Back Propagation) neural network

The invention relates to the technical field of crane ship operation prediction, and discloses a BP neural network-based crane ship operability prediction method. The method comprises the following steps: collecting and preprocessing environmental data of a crane ship operation sea area, classifying and extracting features according to a preset operation state, and dividing a training test set; detecting sample data quality, and marking abnormal samples; using an input feature set of a non-abnormal sample and a historical operation state label matching result to train a model, and generating a BP neural network prediction model; the model parameters are optimized and corrected based on the incidence relation between the weight offset parameters and the prediction errors; and inputting real-time environment data, calculating an output probability along a network forward propagation path, determining the operability of the crane ship, and generating a prediction result. The method can effectively process the non-linear relationship between environmental factors, reduce abnormal data interference, improve the accuracy, objectivity and consistency of prediction, and provide a reliable basis for operation decision making of the crane ship.
Owner:CCCC THIRD HARBOR ENGINEERING CO LTD

Marine main engine power real-time optimization method based on hybrid network model

The invention provides a ship main engine power real-time optimization method based on a hybrid network model, and belongs to the technical field of ship energy saving.A hybrid neural network model is constructed, the hybrid neural network model is based on a physical information neural network, a KAN network is introduced to serve as a front-end network structure, and the power of a ship main engine is optimized in real time; the high-dimensional nonlinear mapping module is used for establishing high-dimensional nonlinear mapping from navigational speed, a ship type parameter matrix, propulsive efficiency, fuel conversion efficiency and environmental factors to the minimum power of a main engine; a composite loss function is adopted to train the hybrid neural network model, wherein the composite loss function is formed by weighted summation of mean square error loss, dynamics constraint loss, propulsive efficiency constraint loss and fuel consumption constraint loss; using the trained hybrid neural network model to receive ship operation parameters and environment parameters collected in real time, and outputting a minimum power prediction value of the ship main engine through one-time forward propagation calculation to realize real-time optimization of the main engine power.
Owner:QINGDAO INNOVATION & DEV CENT OF HARBIN ENG UNIV +1

Assembly line optimization method and device for multi-modal large model training

The embodiment of the invention provides an assembly line optimization method and device for multi-modal large model training, and relates to the technical field of artificial intelligence. According to the method, each model slice of a to-be-trained multi-modal large model is deployed to different training devices, each training batch is divided into a plurality of micro-batches, for each training batch, training sequences corresponding to different training devices are determined according to a parallel scheduling strategy, and the training sequences are used for training the multi-modal large model. The training sequence comprises a forward propagation position, an input back propagation position and a weight back propagation position of each micro-batch, and finally, on each training device, a training process of a corresponding training batch is executed based on the training sequence until training is completed. A back propagation process is divided into back propagation of an input matrix and back propagation of a weight matrix, three calculation stages are jointly formed by the back propagation process and forward propagation, peak shifting calculation can be carried out, bubble time can be effectively filled, idle duration can be reduced, the bubble proportion of an assembly line can be remarkably reduced, and training throughput can be greatly improved.
Owner:PENG CHENG LAB

Digital twinborn calibration framework construction method based on Lyapunov strategy

The invention discloses a digital twin calibration framework construction method based on a Lyapunov strategy, and the method comprises the steps: constructing a digital twin model containing a proxy network to simulate the dynamic state of a system, and defining a state, an action space and a reward function through a Markov decision process. A constrained Lyapunov action-commentator (CLAC) algorithm is introduced, a strategy network and a Lyapunov network are optimized, and the algorithm can stably act under high noise and deviation. Real-time parameter optimization is realized by means of single-time neural network forward propagation, an experience playback pool and the like. Each network structure is clear, and a specific initialization and optimization method is adopted. According to the method, a calibration problem can be converted into a parameter tracking task, efficient training is performed under unmarked data, constraint requirements such as stability can be met, and the method is suitable for scenes such as industrial robot joint control.
Owner:CHINA YANGTZE POWER

Underwater acoustic target recognition method based on effective receptive field regulation

Disclosed in the present invention is an underwater acoustic target recognition method based on effective receptive field regulation. A proposed AEU-Net model has a plurality of resolution branches, with each resolution branch having independent convolution kernels, which adaptively adjust their sizes during training while aligning with their respective resolutions. The AEU-Net model comprises an ERF-Server, which can regulate an effective receptive field during forward propagation, and has two operations of increasing or decreasing the effective receptive field of a feature map of any designated module. In the present invention, acoustic physical information of an underwater acoustic target sample can be captured in a plurality of interaction dimensions, and feature fusion of information from effective receptive fields across a plurality of scales is performed, so as to adapt to sonar image targets of different sizes, thereby improving the accuracy and recognition speed in the recognition of an underwater acoustic target.
Owner:ZHEJIANG UNIV

STFT dimension transformation-based spiking neural network mechanical fault diagnosis method

The invention is applied to the field of mechanical fault diagnosis signal processing, and particularly provides a pulse neural network mechanical fault diagnosis method based on STFT dimension transformation, and the method comprises the steps: collecting a one-dimensional mechanical vibration signal, carrying out the wavelet decomposition, carrying out the wavelet reconstruction of a low-frequency component and a denoised high-frequency component, and carrying out the wavelet reconstruction of the low-frequency component and the denoised high-frequency component; obtaining a denoised one-dimensional vibration signal; performing short-time Fourier transform, and converting the time-frequency two-dimensional matrix into a time-frequency two-dimensional matrix; inputting the time-frequency two-dimensional matrix into an improved HH threshold neuron model, carrying out Poisson sparse coding on the time-frequency two-dimensional matrix, and only carrying out pulse response on signal significant features; constructing a suprathreshold coding convolutional network with residual connection, inputting a sparse coding matrix, training by adopting an unsupervised learning rule based on STDP, and adaptively adjusting a network synaptic weight; and inputting to a trained above-threshold coding convolutional network, and obtaining pulse emission activity of neurons of an output layer through network forward propagation to determine a fault diagnosis result.
Owner:WESTLAKE INSTITUTE FOR OPTOELECTRONICS

Intelligent management platform system for whole-process data integration of seawall engineering construction

The invention relates to the technical field of intelligent construction and digital twinning, and discloses an intelligent management platform system for seawall engineering construction whole-process data integration, and the system comprises a multi-source heterogeneous data collection module which is used for obtaining and standardizing construction data; the core data fusion and reasoning module is used for constructing a construction process causal atlas representing the relationship between elements and calculating the credibility of each state node through a multi-source evidence theory fusion engine; the intelligent application and evolution module is used for carrying out risk forward propagation early warning and root cause backward tracing diagnosis based on the construction process causal atlas; and the comprehensive visualization and decision support module is used for performing fusion display on the analysis result and the BIM / GIS model. According to the method, the causal knowledge graph and evidence fusion reasoning mechanism is constructed, the problems of data islands and information uncertainty are solved, active early warning and accurate traceability of construction risks and dynamic self-adaption of the system are achieved, and the intelligent management level of seawall engineering is remarkably improved.
Owner:ZHEJIANG HYDROPOWER CONSTR & INSTALLATIONCO

Edge end spiking neural network compression and deployment method and system

The invention relates to an edge end spiking neural network compression and deployment method and system, and belongs to the technical field of machine learning, the edge end spiking neural network compression and deployment method performs forward propagation on an initial spiking neural network model based on acquired multi-modal data, obtaining a memory use state in a forward propagation process based on hardware sensing initialization, and performing dynamic sparsification on the initial pulse neural network model based on the memory use state to construct a sparsified pulse neural network; based on the acquired activation statistical information of the sparse spiking neural network, performing dynamic pruning on the sparse spiking neural network by adopting a pruning strategy based on neuron activeness so as to realize compression of a spiking neural network model; the compressed spiking neural network model is optimized according to the edge end configuration, and the optimized spiking neural network model is deployed to the edge end to execute the reasoning task, so that the storage requirement and the calculation complexity are reduced.
Owner:HUBEI ENG UNIV

Neural network global one-time structured pruning method, system and device and medium

The invention relates to a neural network global one-time structured pruning method, system and device and a medium, and the method comprises the steps: carrying out the parallel capturing of the input activation tensors of all target layers in a to-be-pruned neural network through a calibration data set through single-time forward propagation; according to an input activation tensor, synchronously calculating a difference entropy index and amplitude response intensity for an intermediate neuron weight group of each target layer, and performing normalization and fusion to form a static global importance map; determining an importance threshold according to a preset pruning rate, and generating a global to-be-pruned index set of which the mixed importance score is lower than the importance threshold at one time based on the global importance map; and on the basis of the index set, performing one-time physical structured pruning on the weight matrix of each target layer. Therefore, the static global importance map is generated through single forward propagation and parallel capture of the activation tensor, and maskless one-time pruning is completed through physical structured pruning.
Owner:SHANGHAI BANGTU INFORMATION TECH CO LTD

Out-of-distribution prediction

A set of features of a training document are identified in a training document for training a machine learning model. A subset of the features is selected to be omitted from a training forward propagation. As a result of omitting the subset of the set of features, a different subset of the set of features is used to train the machine learning model to classify documents and distinguish between an out-of-domain document and in-domain document.
Owner:CITIGROUP

Model training method and device, equipment and storage medium

The invention provides a model training method and device, equipment and a storage medium, and relates to the technical field of computers, in particular to the technical field of neural network models and model training. The specific implementation scheme is as follows: a calculation unit executes quantization matrix multiplication based on Hadamard pre-transformation on an activation tensor and a weight tensor of a target model stored in a memory so as to generate an output tensor of a linear layer based on a low-precision tensor with smaller data bit width; using the output tensor and a subsequent network layer of the target model to complete forward propagation so as to obtain a loss value; and according to the loss value, updating model parameters of the target model stored in a memory through a back propagation algorithm. By means of the technical scheme, on the premise that the model training precision is guaranteed, memory resource occupation and the calculation amount in the calculation process can be remarkably reduced, and the training cost is reduced.
Owner:BEIJING BAIDU NETCOM SCI & TECH CO LTD

Elevator fault diagnosis method and system

The invention relates to the technical field of elevator fault diagnosis, and discloses an elevator fault diagnosis method and system.The elevator fault diagnosis method comprises the steps that elevator operation data are collected, operation stages are divided for the elevator operation process according to the elevator operation data, and the residual error of the elevator operation data is calculated in the operation stages; based on an elevator structure, a causal chain graph reflecting the relation between components is established through fault mode analysis, and a direct upstream node set of downstream nodes in the causal chain graph serves as a parent set of the downstream nodes; in the operation stage, performing block replacement processing for keeping time sequence characteristics on an upstream node residual sequence of each edge in the causal chain graph to generate a contrast residual sequence, and calculating the propagation intensity of the edge according to the contrast residual sequence; effective propagation paths are screened according to the propagation intensity of the edges, the effective propagation paths are sorted to obtain candidate nodes, residual errors of the candidate nodes are set to be zero through virtual pinch-off, forward propagation is carried out through a semi-physical model, and fault nodes are determined in the candidate nodes based on forward propagation results.
Owner:HUNAN ELECTRICAL COLLEGE OF TECH

Medical image segmentation method based on multi-branch distillation

The invention belongs to the field of image processing, and discloses a medical image segmentation method based on multi-branch distillation, which comprises the following steps: constructing and training a multi-branch collaborative distillation image segmentation model based on uncertainty perception, and establishing a network architecture comprising a public encoder, a first decoder and a second decoder, including a guiding decoder and a guided decoder; dropout disturbance is differentiated on the output of the public encoder, multiple features are generated, multi-branch forward propagation is executed, and a virtual label is generated through fusion of a CTF module; calculating supervision loss and distribution alignment loss of the original feature map through a guided branch; performing multi-level consistency constraint on the output of the disturbance characteristic graph by the guided model to realize multi-view structure consistency, automatically balancing parameter weights through an adaptive task balancing mechanism, and constructing a total loss function; and obtaining a to-be-segmented medical image, and inputting the to-be-segmented medical image into the trained image segmentation model to obtain a segmentation result. According to the invention, high-quality medical image segmentation under a low marking rate can be realized.
Owner:SOUTHWEAT UNIV OF SCI & TECH

Tensor segmentation and mapping method of neural network operator on wafer chip

The invention provides a tensor segmentation and mapping method of a neural network operator on a wafer chip, and the method comprises the steps: expanding a forward calculation graph into a training calculation graph containing forward propagation, back propagation and gradient updating, and employing different operators in the calculation of these stages; carrying out dimension segmentation on the input tensor of each operator in the training calculation graph to generate a plurality of global candidate segmentation strategies; aiming at each candidate segmentation strategy, enumerating feasible resource mapping strategies of all operators, and combining to generate candidate design points; and screening an optimal design point through cost evaluation, and outputting a corresponding candidate segmentation strategy and a resource mapping strategy. According to the method, forward and backward operators in a neural network training process are decoupled through expansion of a calculation graph, and a segmentation scheme for avoiding tensor copying in forward and backward calculation in the training process and a resource mapping strategy corresponding to the segmentation scheme are explored, so that tensor copying is avoided, global resource occupation is reduced, and the training efficiency is improved. Therefore, the same wafer-level chip can support larger model training.
Owner:TSINGHUA UNIVERSITY

Method and device for predicting multi-working-condition flow field of wind driven generator based on neural network

The invention relates to the technical field of wind power generation, artificial intelligence and fluid mechanics, and discloses a prediction method and device for a multi-working-condition flow field of a wind driven generator based on a neural network, and the prediction method comprises the steps: obtaining a sample data set which comprises multiple groups of multi-working-condition data; and inputting the sample data set into a physical information field adversarial neural network model for training to obtain the total loss of forward propagation of the sample data in the field adversarial neural network model. According to the total loss, determining whether training of the wake flow field prediction model is completed; and under the condition that a new round of training is carried out on the wake flow field prediction model, back propagation is carried out on the total loss, and neural network parameters are optimized. In this way, a dual-branch loss collaborative optimization mechanism is formed. A physical information neural network and a domain adversarial neural network are combined, and complementary advantages of the two are fully exerted. When the data is limited or the distribution difference is large, high-precision and physically consistent wake flow field prediction can be realized.
Owner:OCEAN UNIV OF CHINA

Model reinforcement learning method, device and equipment

The embodiment of the invention provides a model reinforcement learning method, device and equipment. The scheme comprises the following steps: in a sampling stage, using an inference engine and adopting a to-be-trained target model to generate an output sequence for an input sequence under a first strategy parameter, recording a first probability value of each lexical element in the output sequence, and calculating a dominant value of each lexical element, in a training stage, after a training engine is used for forward propagation to obtain a second probability value of each lexical element generated by a target model under a first strategy parameter, a first ratio of the second probability value to the first probability value of each lexical element can be calculated, and then the lexical elements with the first ratios within a preset numerical range are screened out to participate in calculation of a target function; and optimizing the target function to update the parameters of the target model.
Owner:ALIPAY (HANGZHOU) INFORMATION TECH CO LTD

Construction method of coefficient prediction model, coefficient prediction method and storage medium

The invention relates to a construction method of a coefficient prediction model, a coefficient prediction method and a storage medium, and the construction method comprises the steps: carrying out the modeling of a bridge prestressed beam through finite element numerical software, simulating a tensioning process, constructing a bridge parameter coefficient array, dividing the bridge parameter coefficient array into a training data set and a verification data set, performing data standardization and tensor data conversion, and constructing a data loader; an AND-based full-connection multi-layer multi-task neural network is constructed, and a standardizing device, a forward propagation function, a loss function, a learning rate scheduler and an optimizer are arranged; and carrying out model training, packaging a prediction function, and integrating a standardizing device and the trained prediction model. The prediction method comprises the step of predicting the friction resistance loss coefficient and the pipeline deviation coefficient according to the prediction model. According to the method and the device, coefficient prediction can be quickly, efficiently and accurately realized by only needing several on-site easy-to-measure parameters.
Owner:JILIN JIANZHU UNIVERSITY

Segmentation federal learning method and system based on aggregation gradient broadcast

The invention provides a segmentation federal learning method and system based on aggregation gradient broadcast, and the method comprises the steps: initializing a global model, and segmenting the global model into a server model and a client model according to the capability limitation of a local device; issuing the client model to all local equipment terminals; the multiple local devices execute forward propagation at the same time, and shredded data are obtained through calculation; transmitting the shredded data and the training data label to an edge server; performing forward propagation in parallel according to the shredded data, calculating the loss of each local equipment end in combination with a corresponding training data label, and performing back propagation to obtain a loss function and a gradient of the shredded data; aggregating the gradient of the shredded data and broadcasting to all local equipment; updating the server side model and the client side model, and aggregating the server side model; and the iterative learning loop is repeatedly executed until convergence or the maximum communication round is reached. According to the invention, the communication overhead in segmentation federal learning is saved, and the model training efficiency is effectively improved.
Owner:WUHAN UNIV

Neural network training method based on back propagation algorithm

The invention discloses a neural network training method based on a back propagation algorithm, and relates to the technical field of network training, and the method comprises the steps: obtaining the neuron output data of each neural network layer through a forward propagation process, and constructing the output distribution data reflecting the activation features of each layer of neuron; obtaining error data based on a back propagation process, establishing error term distribution of each neural network layer, and constructing an entropy change estimation factor for measuring network information state change; performing disturbance control on the error data according to the value trend of the entropy change estimation factor, and dynamically adjusting the parameter updating amplitude and direction of the neural network by the processed error data and the output distribution data; and weight iteration of the neural network is guided through the parameter updating behavior. According to the method, error disturbance control and dynamic parameter adjustment are combined, a whole set of mechanism of information perception-disturbance control-dynamic optimization is formed, and dynamic perception and accurate guidance of information state evolution in the neural network training process are achieved.
Owner:NANJING DANIU INFORMATION TECH CO LTD

Quantitative perception training method and device of neural network model, electronic equipment and storage medium

The invention relates to a quantitative perception training method and device of a neural network model, electronic equipment and a storage medium. The method comprises the steps of obtaining a to-be-trained first neural network model; an operator pair and a non-module operator in the first neural network model are identified, the non-module operator is an operation which is not realized based on a module class, and the operator pair comprises a convolution operator and a batch normalization operator which are connected; the non-module operators are packaged into module operators, the operator pairs are packaged into new convolution operators, a second neural network model is obtained, the new convolution operators run the fusion process of the convolution operators and batch normalization operators during forward propagation, and the parameters of the convolution operators and the batch normalization operators are updated during back propagation; and performing quantitative perception training on the second neural network model. By adopting the method, the reasoning precision of the quantitative model can be guaranteed, and the deployment efficiency of the quantitative model is improved.
Owner:GUANGZHOU XIAOMA HUIXING TECH CO LTD

Large model deployment method based on heterogeneous resource scheduling and multi-model collaborative reasoning

The invention discloses a large model deployment method based on heterogeneous resource scheduling and multi-model collaborative reasoning, and relates to the field of artificial intelligence and server system deployment. Comprising the following steps of: 1, calculating a deployment score of a target model by utilizing a deployment scoring function according to a GPU video memory residual rate, a CPU load, network delay and a current task quantity of each node through a heterogeneous resource sensing scheduler, and selecting a node with the highest score of the target model to load or reuse the target model; a cache index key is generated, and whether key value KV pairs supporting multiplexing of the target model exist or not is inquired; step 3, if the cache is hit, skipping a forward propagation stage, and directly decoding based on a cache result, otherwise, executing complete forward reasoning and writing the result into the cache; 4, based on a plurality of user requests, performing fusion according to user priorities, model weights, reasoning costs and cue word lengths to generate execution sorting weights, and scheduling execution models in batches according to the weights; and 5, returning the model response to the user, and releasing or updating node state information.
Owner:INSPUR ENTERPRISE CLOUD TECHNOLOGY (SHANDONG) CO LTD

Large model training method and device, electronic equipment, storage medium and program product

The invention relates to a large model training method and device, electronic equipment, a storage medium and a program product. The method comprises the following steps: for any item of training data of a target large model, segmenting the training data into a plurality of parts of segmented data, storing the plurality of parts of segmented data in a nonvolatile memory, and sequentially performing forward propagation calculation and back propagation calculation on the plurality of parts of segmented data; for any part of segmented data, reading the segmented data from a nonvolatile memory to a video memory, and executing forward propagation calculation on the segmented data through a GPU (Graphics Processing Unit) to obtain an activation value corresponding to the segmented data; and for any part of segmented data, executing back propagation calculation based on the activation value corresponding to the segmented data through the GPU to obtain gradient data corresponding to the segmented data, and moving the gradient data corresponding to the segmented data from the video memory to a nonvolatile memory or a CPU memory. The display memory occupation of the activation value can be reduced.
Owner:MOORE THREADS TECH CO LTD

Autoregression text generation acceleration method and system based on dynamic token tree

The invention relates to the technical field of large model reasoning acceleration, in particular to an autoregressive text generation acceleration method based on a dynamic token tree, which comprises the following steps of: S1, performing path expansion on a current context by using a draft model, generating a candidate token tree, calculating the path probability from each leaf node to a root node, and reserving the first N leaf nodes with the highest path probability, deleting other paths; s2, combining all the candidate sequences into a combined matrix filled to be equal in length, generating a self-defined attention mask, enabling each token to only pay attention to an ancestor token of a path where the token is located, inputting the combined matrix and the self-defined attention mask into a target model, and verifying all the sequences in parallel through single-time forward propagation; and S3, mapping the position of the accepted token in the candidate token tree to the continuous logic position of the main sequence, and copying the key value cache of the target model to the corresponding position of the main sequence in batches according to the mapping relationship. According to the method, the token acceptance rate in speculation sampling of the large model acceleration technology is improved, so that the throughput is greatly improved.
Owner:PIO CLOUD COMPUTING (SHANGHAI) CO LTD

Power network anomaly detection method and system

The invention discloses a power network anomaly detection method and system, and relates to the technical field of intelligent power grid monitoring, and the method comprises the steps: constructing a neuron-like power network graph model, carrying out the node processing of equipment, introducing a multi-factor anomaly score function, and dynamically activating an abnormal node through combining with an improved Winner-Take-All mechanism; the connection weight is dynamically optimized through a topology-aware composite gradient descent method, and a multi-layer neural network structure is constructed to realize cross-layer forward propagation and feedback adjustment; and establishing an abnormal path model in combination with the propagation phase difference and the information entropy, identifying an abnormal path by using a propagation score function and a weighted shortest path algorithm, and carrying out fault tracing. According to the power network anomaly detection method provided by the invention, the neuron-like graph model is constructed, and a multi-factor anomaly activation mechanism is introduced, so that comprehensive perception of a power grid node state under a multi-dimensional time characteristic can be realized, and the anomaly recognition sensitivity and the multi-point simultaneous activation capability are effectively improved.
Owner:GUIZHOU POWER GRID CO LTD

Three-dimensional head model generation method for Gaussian point cloud reconstruction based on forward propagation

The invention discloses a three-dimensional head model generation method based on Gaussian point cloud reconstruction of forward propagation, and the method comprises the steps: compressing an input image to a potential space, obtaining a potential state, and extracting the identity information of the input image at the same time in a stage of generating a multi-view image; de-noising is carried out through a de-noising U-Net with a space and time attention module, and a multi-view image is generated through a VAE decoder and is used for training a weight fine tuning network; in the Gaussian reconstruction stage, the world coordinate origin and the light direction of the camera where each pixel in the multi-view images generated by the video diffusion model is located are calculated, the multi-view images are spliced in the channel dimension through the corresponding light embedding and light direction obtained through calculation, network input features are formed, and the network input features are used as network input features. An asymmetric Gaussian generation UNet based on LGM network improvement is utilized, a feature map is generated according to a multi-view image generated by a video diffusion model, and channel flattening processing is performed on the feature map to obtain a Gaussian point cloud. According to the invention, a more vivid and lifelike 3D head model with high quality can be generated.
Owner:SHANGHAI JIAOTONG UNIV

Private data security sharing method based on federal learning

The invention discloses a private data security sharing method based on federal learning, particularly relates to the field of private data security protection and sharing, and is used for solving the problems of insufficient model credibility and lack of precision control of data exchange in the existing cross-mechanism data collaboration process. According to the method, homomorphic encryption processing is carried out on local data of all participants, a ciphertext verifiable calculation task is constructed in combination with federated learning, forward propagation and back propagation calculation of a ciphertext are carried out by outsourcing calculation nodes, credible verification is carried out on a calculation process by utilizing zero-knowledge proof, and gradient aggregation and model updating are completed in a ciphertext domain; after model training is completed, based on federal learning model output, intelligent judgment is conducted on the sharing value of local data samples, ciphertext re-encryption and data exchange control are driven, multi-party safe and controllable data circulation is achieved, and therefore on the premise that privacy safety and compliance requirements are guaranteed, cross-mechanism data collaboration efficiency and decision-making precision are improved.
Owner:XIAMEN UNIV OF TECH

Dynamic multifunctional optical metasurface design method and system

The embodiment of the invention provides a dynamic multifunctional optical metasurface design method and system, and the method comprises the steps: obtaining function demands in different application scenes, and setting a target reflection spectrum according to the function demands; the target reflection spectrum is input into a reverse retrieval network, and super-unit structure prediction design parameters are obtained through encoder compression and decoder parameter prediction; the super-unit structure prediction design parameters are input into a forward prediction network, and a prediction reflection spectrum is obtained through network forward propagation calculation; based on the matching degree between the predicted reflection spectrum and the target reflection spectrum, establishing a bidirectional relation between a design domain and a physical domain by using a phase recovery algorithm, and predicting design parameters by continuously adjusting the super-unit structure, so that the predicted reflection spectrum gradually approaches the target reflection spectrum, and obtaining ideal design parameters of the super-unit structure; and according to ideal design parameters of the super-unit structure, super-units are arranged in an operation space, and the super-surface structure capable of realizing dynamic multifunctional regulation and control is formed.
Owner:WUHAN YILUT TECH CO LTD