An intelligent decision-making analysis method and system for a digital-twin-based manufacturing system

Through the deep integration of meta-learning and adversarial training and generative enhancement technology, a basic model with rapid adaptability is built, which solves the data scarcity and distribution differences of traditional production scheduling systems in emerging markets, realizes dynamic optimization and feasibility verification of production strategies in digital twin environments, and improves the rapid fulfillment and precise supply capabilities of multinational manufacturing enterprises.

CN120218679BActive Publication Date: 2025-07-18FUJIAN YANGTENG INNOVATION INFORMATION TECHNOLOGY CO LTD
View PDF 2 Cites 0 Cited by

Patent Information

Application Number
CN202510687618.8
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2025-05-27
Publication Date
2025-07-18
Estimated Expiration
2045-05-27

AI Technical Summary

Technical Problem

Due to defects such as data silos, model staticization, and decision-making delay, traditional production scheduling systems are difficult to meet the rapid fulfillment and precise supply needs of multinational manufacturing companies in emerging markets. Especially when the number of data in the target domain is small and the distribution of the source domain is greatly different, resulting in negative migration and increased prediction errors on the test set of the digital twin model.

Method used

Through the deep integration of meta-learning and adversarial training, a basic model with fast adaptability is built, high-quality training samples are expanded with generative enhancement technology, and multi-objective deep reinforcement learning algorithms are run in a digital twin environment, dynamically optimize production strategies, and scheduling strategies are adjusted in combination with real-time feedback to achieve closed-loop optimization.

Benefits of technology

It has improved the generalization ability and adaptability of the digital twin model in small sample scenarios in emerging markets, alleviated the prediction distortion problem caused by data scarcity, realized dynamic optimization and feasibility verification of production strategies, and formed a closed-loop optimization intelligent decision-making system.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120218679B_ABST
    Figure CN120218679B_ABST
Patent Text Reader

Abstract

The present invention relates to the field of digital twins, and particularly to an intelligent decision-making analysis method and system for a manufacturing system based on digital twins. Through the deep integration of meta-learning and adversarial training, the present invention improves the generalization ability and adaptability of the digital twin model in the small-sample scenario of emerging markets, alleviates the problem of prediction distortion caused by data scarcity, and combines generative enhancement technology to expand high-quality training samples. The simultaneously constructed multi-objective flexible scheduling mechanism realizes the dynamic optimization and feasibility verification of production strategies in the digital twin environment, and finally forms a closed-loop optimized intelligent decision-making system. Through the deep integration of meta-learning and adversarial training, and by combining generative enhancement technology to expand high-quality training samples, the simultaneously constructed multi-objective flexible scheduling mechanism realizes the dynamic optimization and feasibility verification of production strategies in the digital twin environment, and finally forms a closed-loop optimized intelligent decision-making system.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present invention relates to the field of digital twins, and more specifically, to an intelligent decision-making analysis method and system for a manufacturing system based on digital twins. Background Art

[0002] With the exponential growth of the global e-commerce market scale, multinational manufacturing enterprises are facing many challenges. Due to defects such as data islands, static models, and decision-making delays in traditional production scheduling systems, it is difficult to meet the cross-border e-commerce requirements of "fast fulfillment and accurate supply".

[0003] Especially when exploring emerging markets, due to the small number of target domain data and the large distribution difference from the source domain, the domain adversarial adaptive loss in the digital twin model will cause gradient conflicts due to insufficient samples, and the model after migration will show negative transfer on the test set, and the prediction error will increase by 25% instead. Summary of the Invention

[0004] The present invention provides an intelligent decision-making analysis method and system for a manufacturing system based on digital twins, which solves the technical problems in the related art that the number of target domain data in emerging markets is small and the distribution difference from the source domain is large.

[0005] The present invention provides an intelligent decision-making analysis method for a manufacturing system based on digital twins, including the following steps:

[0006] S100. By extracting multiple groups of simulation tasks from source domain data, training a base model with fast adaptation ability. The base model first fine-tunes the parameters on the support set of each small task, and then reversely updates the initial parameters according to the performance of the query set of all tasks, and finally obtains a meta-knowledge base that can quickly adapt to the new market;

[0007] S200. On the target domain small samples, adopt an adversarial training strategy to align the source domain and target domain features. The feature extractor is forced to learn domain-invariant representations, while the domain classifier tries to distinguish the data sources, and at the same time dynamically adjusts the weights of the prediction loss and the adversarial loss to prevent negative transfer caused by overfitting of small samples;

[0008] S300. Based on the target domain scarce data and the domain knowledge graph, use a conditional diffusion model to generate a large number of synthetic data, and construct an enhanced data set through distribution consistency test and pre-training model screening;

[0009] S400. Load the enhanced data and the adapted model in the digital twin, run a multi-objective deep reinforcement learning algorithm, dynamically adjust the weights of the order fulfillment rate and the inventory cost in combination with real-time feedback, and finally select the optimal scheduling strategy from the Pareto optimal solution set.

[0010] Further, in S100, it includes the following steps:

[0011] S110, Meta-task construction: Sample meta-tasks from the source domain data, where each task contains a support set and a query set;

[0012] S120, Internal parameter adaptation: For each meta-task, calculate the support set loss and update it based on the current meta-parameters;

[0013] S130, External meta-parameter optimization: Evaluate the performance of the adapted model on the query sets of all tasks, and update the meta-parameters by backpropagation;

[0014] S140, Convergence determination: Repeat S110 - S130 until the stopping condition is met.

[0015] Furthermore, in S140, the stopping condition is as follows:

[0016] ;

[0017] where is the convergence threshold, is the maximum number of iterations, , is the current iteration number, is the meta-loss value at the th iteration, is the meta-loss value of the previous iteration.

[0018] Furthermore, in S200, it includes the following steps:

[0019] S210, Model initialization: Load the meta-pretrained parameters as the initial parameters;

[0020] S220, Domain adversarial feature alignment: Achieve feature space domain invariance through the gradient reversal layer;

[0021] S230, Joint optimization: Synchronously optimize the prediction task loss and the domain adversarial loss;

[0022] S240, Parameter update: Alternately update the parameters of the feature extractor and the domain classifier:

[0023] S250, Few-shot early stopping strategy: Monitor the loss of the target domain validation set and terminate the training when there is no decrease for 5 consecutive iterations.

[0024] Furthermore, in S240, the calculation formula for parameter update is as follows:

[0025] Fix while updating :

[0026] ;

[0027] Fixed Update simultaneously :

[0028] ;

[0029] Among them, is the parameter of the feature extractor, is the parameter of the domain classifier, is the learning rate of adversarial learning, , is the gradient with respect to the parameter , is the gradient with respect to the parameter , is the loss function of domain adversarial, is the total loss function.

[0030] Furthermore, the calculation formula of the few-shot early stopping strategy is as follows:

[0031] ;

[0032] Among them, is the loss value on the target domain validation set, is the current training round, represents the training termination signal, which is sent when the early stopping condition is triggered, represents taking the minimum value among all historical rounds less than the current round t, represents the historical training round index, which is used to compare the current loss with the historical optimal loss.

[0033] Furthermore, in S300, the following steps are included:

[0034] S310, Domain knowledge conditional encoding: Encode the target domain prior knowledge into a conditional vector;

[0035] S320, Conditional diffusion model training: Train the denoising diffusion probability model with the target domain samples and conditions as inputs;

[0036] S330, Diversity controllable sample generation: Sample and generate augmented data from the diffusion model, and constrain it to be consistent with the target domain distribution;

[0037] S340, Generated sample quality verification: Screen valid samples through a pre-trained validator;

[0038] S350, Augmented dataset construction: Merge real and synthetic data to construct the final training set.

[0039] Furthermore, in S400, the following steps are included:

[0040] S410, Environment Initialization: Load the virtual factory model and inject enhanced data ;

[0041] S420, Multi-objective Deep Reinforcement Learning Policy Training: Construct an Actor-Critic architecture and optimize three objectives:

[0042] Actor Network : Input the state and output the action ;

[0043] Critic Network : Evaluate the action value, including three output heads:

[0044] ;

[0045] Objective Function:

[0046] ;

[0047] Among them, is the Actr network, and the parameter is , is the Critic network, and the parameter is , is the state vector, is the action vector, is the state distribution generated by the behavioral policy, is the variance penalty coefficient, is the variance of the action, is the cost evaluation value, is the quality of service evaluation value, is the feasibility evaluation value, represents the objective function, is the expectation operator, represents the variance of the action;

[0048] S430, Dynamic Weight Adjustment Mechanism: Adjust the objective weights according to the real-time feedback of the digital twin;

[0049] S440, Virtual-Physical Closed-loop Verification: Execute the policy in the digital twin to verify the feasibility;

[0050] S450, Pareto Front Policy Selection: Select the optimal policy from the non-dominated solution set.

[0051] Furthermore, the calculation formula for the optimal policy is as follows:

[0052] ;

[0053] Among them, is the optimal policy, is the Pareto optimal strategy set, is the weight of the th target, is the th target's Q value, is the feasibility score of strategy , is the system state, is the action output of strategy in state ;

[0054] The output of the optimal strategy is as follows:

[0055] ;

[0056] where is the action of the production line start / stop decision, is the resource allocation ratio, is the production capacity allocation weight.

[0057] The present invention also proposes an intelligent decision-making analysis system for a digital twin-based manufacturing system, which executes the steps in the intelligent decision-making analysis method for a digital twin-based manufacturing system as described above, including:

[0058] Meta-learning base model construction module: Automatically construct multiple groups of simulation tasks from the historical data of the mature market, and obtain a meta-learning model with fast adaptation ability through collaborative training of internal and external loops, so that it can quickly capture the common laws across markets with a small amount of target data;

[0059] Adversarial domain adaptation module: Introduce an adversarial training mechanism on a very small amount of data in the target domain, dynamically balance the prediction accuracy and domain-invariant feature learning, and eliminate the distribution differences between the source domain and the target domain through the gradient reversal and weight adaptation strategies;

[0060] Data augmentation generation module: Generate synthetic data that conforms to the characteristics of the target domain based on the conditional diffusion model and the domain knowledge graph, and expand high-quality training samples by combining distribution consistency verification and pre-trained model screening;

[0061] Scheduling optimization decision module: Simulate multi-objective scheduling strategies in the digital twin environment, dynamically optimize the order fulfillment rate, inventory cost and feasibility score through reinforcement learning, and finally output a flexible production plan verified by virtualization.

[0062] The beneficial effects of the present invention are as follows:

[0063] Through the deep integration of meta - learning and adversarial training, the present invention enhances the generalization ability and adaptability of the digital twin model in the small - sample scenario of emerging markets, alleviates the problem of prediction distortion caused by data scarcity, and combines generative enhancement technology to expand high - quality training samples. A multi - objective elastic scheduling mechanism is synchronously constructed to achieve the dynamic optimization and feasibility verification of production strategies in the digital twin environment, and finally form an intelligent decision - making system with closed - loop optimization. Brief Description of the Drawings

[0064] Figure 1 is a flowchart of an intelligent decision - making analysis method for a digital - twin - based manufacturing system proposed by the present invention;

[0065] Figure 2 is of the present invention Figure 1 is a flowchart of the sub - steps of S100;

[0066] Figure 3 is of the present invention Figure 1 is a flowchart of the sub - steps of S200;

[0067] Figure 4 is of the present invention Figure 1 is a flowchart of the sub - steps of S300;

[0068] Figure 5 is of the present invention Figure 1 is a flowchart of the sub - steps of S400;

[0069] Figure 6 is a structural block diagram of an intelligent decision - making analysis system for a digital - twin - based manufacturing system proposed by the present invention.

[0070] In the figure: 101, meta - learning base model construction module; 102, adversarial domain adaptation module; 103, data enhancement generation module; 104, scheduling optimization decision module. Detailed Embodiments

[0071] Now, the subject matter described herein will be discussed with reference to exemplary embodiments. It should be understood that discussing these embodiments is only to enable those skilled in the art to better understand and thus implement the subject matter described herein. Without departing from the scope of protection of the content of this specification, changes can be made to the functions and arrangements of the elements discussed. Each example can omit, substitute, or add various processes or components as needed. Additionally, the features described relative to some examples can also be combined in other examples.

[0072] As Figures 1 - 5 shown, an intelligent decision - making analysis method for a digital - twin - based manufacturing system includes the following steps:

[0073] S100, Meta - learning Pre - training (offline): By extracting multiple groups of simulated tasks from the data in the source domain (mature market), train a base model with fast adaptation ability. The base model first fine - tunes the parameters on the support set of each small task, and then updates the initial parameters reversely according to the performance on the query set of all tasks, finally obtaining a meta - knowledge base that can quickly adapt to the new market;

[0074] It should be noted that the mature market refers to the cross - border e - commerce target area with the following characteristics:

[0075] Data Completeness: Having more than 100,000 historical order data, covering the complete product life cycle (R & D, production, delisting);

[0076] Demand Stability: The demand fluctuation coefficient (standard deviation / mean) is less than 0.3, and the seasonal index deviation is less than 15%;

[0077] Supply Chain Maturity: A multi - level warehousing network has been established, and the average order fulfillment time limit ≤ 5 days;

[0078] In an embodiment of the present invention, it specifically includes the following steps:

[0079] S110, Meta - task Construction: Sample meta - tasks from the source domain data , and each task contains:

[0080] Support Set:

[0081] ;

[0082] where the support set is used for internal parameter optimization;

[0083] Query Set:

[0084] ;

[0085] where the query set is used for external parameter optimization

[0086] where, is the th meta - task, is the number of support set samples ( ), is the number of query set samples ( ), is the number of meta - tasks, is the input feature vector, is the corresponding label value, is the source domain data set;

[0087] S120, Internal Parameter Adaptation: For each meta-task , based on the current meta-parameters , calculate the support set loss and update it;

[0088] ;

[0089] Task Loss Calculation:

[0090] ;

[0091] Parameter Update:

[0092] ;

[0093] Among them, is the internal learning rate ( ), is the meta-parameter vector, is the adaptation parameter of task , is the model function with parameter , is the loss function on the support set, is the gradient operator with respect to parameter ;

[0094] S130, External Meta-parameter Optimization: Evaluate the performance of the adapted model on the query sets of all tasks, and update the meta-parameters by backpropagation;

[0095] ;

[0096] Meta-loss Calculation:

[0097] ;

[0098] Meta-gradient Update:

[0099] ;

[0100] Among them, is the external learning rate ( ), is the loss function on the query set, is the total meta-learning loss, is the adapted model with parameter , represents the meta-parameter vector to be optimized, is the partial derivative of the meta-loss with respect to the meta-parameters, is the k-th input sample in the query set, is the k-th label value in the query set;

[0101] S140, Convergence determination: Repeat S110 - S130 until the stop condition is met;

[0102] The stop conditions are as follows:

[0103] ;

[0104] where, is the convergence threshold ( ), is the maximum number of iterations ( ), is the current iteration number, is the meta - loss value at the th iteration, is the meta - loss value of the previous iteration;

[0105] S200, Adversarial domain adaptation (online): On the small - sample target domain (emerging market), adopt an adversarial training strategy to align the source - domain and target - domain features. The feature extractor is forced to learn domain - invariant representations, while the domain classifier attempts to distinguish the data sources. At the same time, dynamically adjust the weights of the prediction loss and the adversarial loss to prevent negative transfer caused by overfitting of small samples;

[0106] It should be noted that the emerging market refers to the cross - border e - commerce expansion area with the following characteristics:

[0107] Data scarcity: The amount of available effective historical data is less than 1000, and key fields are missing (such as user portraits, reasons for returns);

[0108] High demand volatility: The demand volatility coefficient exceeds 0.5, and the daily order volume change range can reach 300% due to the influence of social media;

[0109] Imperfect supply chain: Rely on a single logistics channel, the average customs clearance time ≥ 7 days, and door - to - door delivery cannot be achieved in more than 30% of the regions;

[0110] In an embodiment of the present invention, it specifically includes the following steps:

[0111] S210, Model initialization: Load the meta - pre - trained parameters as the initial parameters;

[0112] ;

[0113] where, is the initialized model parameter, is the optimal meta - parameter obtained by meta - learning;

[0114] S220, Domain - adversarial feature alignment: Achieve domain invariance in the feature space through the Gradient Reversal Layer (GRL);

[0115] Feature extraction:

[0116] ;

[0117] Domain prediction:

[0118] ;

[0119] Domain classification loss:

[0120] ;

[0121] Among them, is a feature extractor network with parameter ; is a 256-dimensional feature vector extracted, is the domain label probability predicted by the domain classifier, is the true domain label, 0 represents the source domain, and 1 represents the target domain, is the number of samples in each batch( ), is the domain classifier parameter, following a normal distribution , is a domain classifier network with parameter ; is the loss function of domain adversarial, is the true domain label of the i-th sample, is the predicted domain label probability of the i-th sample;

[0122] S230, Joint optimization: Synchronously optimize the prediction task loss and the domain adversarial loss:

[0123] ;

[0124] Prediction task loss (only the source domain has labels):

[0125] ;

[0126] Weight adjustment:

[0127] ;

[0128] Among them, is the total loss function, is the loss function of the prediction task, is the loss function of the domain adversarial, is the batch size of the source domain samples( ), is the weight coefficient of the domain adversarial loss (initial value 0.1), is the weight adjustment coefficient( ) is the average absolute error of the current target domain, and is the average absolute error of the initial source domain;

[0129] S240, Parameter update: Alternately update the parameters:

[0130] Fix while updating :

[0131] ;

[0132] Fix while updating :

[0133] ;

[0134] Among them, are the parameters of the feature extractor, are the parameters of the domain classifier, is the learning rate of adversarial learning ( ), is the gradient with respect to the parameter , is the gradient with respect to the parameter ;

[0135] S250, Few-shot early stopping strategy: Monitor the loss of the target domain validation set , and terminate the training when there is no decrease for 5 consecutive iterations;

[0136] ;

[0137] The monitoring metrics are as follows:

[0138] Validation set: Randomly partition 20% of the samples from ( = 20);

[0139] Loss calculation:

[0140] ;

[0141] Among them, is the loss value on the target domain validation set, is the current training epoch, is the number of validation set samples ( , which is 15% of the target domain data ), is the input feature of the validation set, is the true label of the validation set, represents the training termination signal, which is sent when the early stopping condition is triggered. denotes taking the minimum value among all historical rounds less than the current round t. denotes the historical training round index, used to compare the current loss with the historical optimal loss. denotes the input features of the target domain samples. denotes the true label of the target domain.

[0142] S300, Synthetic Data Augmentation: Based on the scarce data in the target domain and the domain knowledge graph, use the conditional diffusion model to generate a large amount of synthetic data, and construct an augmented dataset through distribution consistency testing and pre-training model screening.

[0143] In one embodiment of the present invention, it specifically includes the following steps:

[0144] S310, Domain Knowledge Conditional Encoding: Encode the prior knowledge of the target domain (such as regional consumption characteristics, product category tree) into a conditional vector.

[0145] ;

[0146] Wherein, is the conditional vector, is the knowledge embedding layer (parameters ). is the regional encoding, is the category encoding, is the embedding layer parameter matrix.

[0147] S320, Conditional Diffusion Model Training: Train the denoising diffusion probabilistic model (DDPM) with the target domain samples and the condition as the input.

[0148] Forward process:

[0149] ;

[0150] Backward process:

[0151] ;

[0152] Wherein is implemented by UNet, and the optimization objective is:

[0153] ;

[0154] Wherein, is the noise coefficient ([[]] ), is the noise prediction network, is the number of diffusion steps ([[]] ), is the original input data. is the data after adding noise at the t-th step, is the mean prediction network, is the variance prediction network, is the identity matrix, represents the Gaussian distribution, represents the data state after t-step diffusion, represents the data state after (t - 1)-step diffusion, is the conditional probability distribution of the forward diffusion process, represents the transition probability of single-step forward diffusion, represents the probability distribution of the conditional reverse diffusion process, represents the loss function of the DDPM model, represents for time step , initial data and random noise to take the expectation, represents the square of the Euclidean norm, represents the product from t = 1 to T;

[0155] S330, Diversity-Controllable Sample Generation: Sample enhanced data from the diffusion model, constrained to be consistent with the target domain distribution;

[0156] ;

[0157] Implementation Method: Use classifier guidance sampling and inject conditional gradients during the reverse process:

[0158] ;

[0159] Among them, is the k-th generated sample, is the conditional generation distribution, is the maximum mean discrepancy metric, is the real data sample, is the guidance coefficient, is the MMD threshold, represents with respect to the gradient operator, is the unconditional generation distribution, representing the generation model distribution at time t, represents the conditional probability distribution, representing the posterior probability of condition given the generation state , represents the conditional generation distribution, representing the generation model distribution at time t under condition c, represents the set of generated samples, represents the set of real samples;

[0160] S340, Generate sample quality verification: Verify through a pre-trained verifier Screen valid samples;

[0161] ;

[0162] Among them, is the verification model, is the similarity threshold, , is the enhanced dataset, is the KL divergence, is the verification model parameter, represents the delimiter symbol in the KL divergence calculation, represents the feature representation of the generated samples by the verification model, represents the feature representation of the real samples by the verification model;

[0163] S350, Enhanced dataset construction: Merge real and synthetic data to construct the final training set:

[0164] ;

[0165] The required balancing strategy is constructed as follows:

[0166] Expand to , Select ;

[0167] Among them, is the finally constructed enhanced dataset, is the target domain dataset, is the final dataset size;

[0168] S400, Elastic scheduling optimization: Load the enhanced data and the adaptation model in the digital twin, run the multi-objective deep reinforcement learning algorithm, dynamically adjust the weights of order fulfillment rate, inventory cost, etc. based on real-time feedback, and finally select the optimal scheduling strategy from the Pareto optimal solution set;

[0169] In one embodiment of the present invention, it specifically includes the following steps:

[0170] S410, Environment initialization: Load the virtual factory model and inject the enhanced data ;

[0171] ;

[0172] Among them, is the digital twin environment, is the Unity3D factory model, is the finally constructed enhanced dataset;

[0173] S420, Multi-objective Deep Reinforcement Learning (MO-DDPG) Policy Training: Construct an Actor-Critic architecture and optimize three objectives:

[0174] Actor network : Input state and output action (equipment start / stop, logistics path);

[0175] Critic network : Evaluate the action value, including three output heads:

[0176] ;

[0177] Objective function:

[0178] ;

[0179] Among them, is the Actr network, and the parameter is , is the Critic network, and the parameter is , is the state vector, is the action vector, is the state distribution generated by the behavior policy, is the variance penalty coefficient, , is the variance of the action, is the cost evaluation value, is the quality of service evaluation value, is the feasibility evaluation value, represents the objective function, which is used to optimize the Actor network parameters, is the expectation operator, which is used to calculate the average return, represents the variance of the action, which is used to encourage action diversity;

[0180] S430, Dynamic Weight Adjustment Mechanism: Adjust the objective weights according to the real-time feedback of the digital twin;

[0181] ;

[0182] Among them, is the weight of the th objective at the th moment, is the normalized score of the th objective at the th moment, is the temperature parameter ( ), is the current value of the th target, is the target value of the th target;

[0183] Specific threshold values: service rate ≥ 0.95, cost ≤ 0.3, feasibility ≥ 0.9;

[0184] S440, virtual-physical closed-loop verification: Execute the policy in the digital twin , and verify the feasibility;

[0185] ;

[0186] Among them, is the feasibility score of the policy , is the number of verification rounds, is the indicator function, is the result of the

[0187] round of verification;

[0188] S450, Pareto front policy selection: Select the optimal policy from the non-dominated solution set :

[0189] ;

[0190] Among them, is the optimal policy, is the Pareto optimal policy set, is the weight of the th target, is the Q value of the th target, is the feasibility score of the policy , is the policy action output in the state

[0191] Optimal policy output:

[0192] ;

[0193] Among them is the action of the production line start / stop decision, is the resource allocation ratio, is the production capacity allocation weight;

[0194] For example Figure 6As shown in the figure, based on the above method, an intelligent decision-making analysis system for a manufacturing system based on digital twins is proposed, including the following modules:

[0195] Meta-learning base model construction module 101: Automatically construct multiple sets of simulation tasks from the historical data of the mature market, and obtain a meta-learning model with fast adaptation ability through collaborative training of internal and external loops, enabling it to quickly capture cross-market common laws under a small amount of target data and laying a foundation for small-sample migration;

[0196] Adversarial domain adaptation module 102: Introduce an adversarial training mechanism on a very small amount of data in the target domain, dynamically balance prediction accuracy and domain-invariant feature learning, and eliminate the distribution differences between the source domain and the target domain through gradient reversal and weight adaptation strategies to avoid performance collapse of the model in emerging markets;

[0197] Data augmentation generation module 103: Generate synthetic data that conforms to the characteristics of the target domain based on the conditional diffusion model and the domain knowledge graph, and combine distribution consistency verification and pre-trained model screening to expand high-quality training samples and solve the problem of model overfitting caused by data scarcity;

[0198] Scheduling optimization decision module 104: Simulate multi-objective scheduling strategies in the digital twin environment, dynamically optimize the order fulfillment rate, inventory cost, and feasibility score through reinforcement learning, and finally output a flexibly produced plan verified by virtualization to ensure the reliable execution of the strategy in the physical system.

[0199] Based on the above method and system, the following is an example:

[0200] A multinational enterprise "Global Tech" plans to enter the East African smartphone market, but faces:

[0201] Data scarcity: Only 3 months of local sales data (about 100 items) have been collected;

[0202] Demand difference: African users prefer mobile phones with long battery life (battery capacity greater than 5000mAh), which is significantly different from the European and American markets;

[0203] Supply chain restrictions: A new assembly plant needs to be set up in Kenya, but there is no experience in production capacity allocation;

[0204] System application process:

[0205] Meta-learning base model construction: Load 100,000 historical data from the European and American markets, automatically generate 1200 sets of simulation tasks (such as "festival promotions", "chip shortage responses"), and the trained model can identify cross-market common laws (such as "the correlation between price sensitivity and GDP");

[0206] Adversarial Domain Adaptation: Input 100 pieces of data from Africa (including features such as battery capacity and local payment methods). The system detects a distribution shift in the "battery life demand" feature (KL divergence = 7.8) and automatically strengthens the feature alignment in this dimension. After 8 hours of training, the prediction error decreases from the initial 42% to 15%.

[0207] Data Augmentation Generation: Generate 10,000 virtual orders based on the cooperative data of African operators. The generated data includes special scenarios such as purchase delays caused by electricity price fluctuations during the dry season and a mobile payment ratio of 85%.

[0208] Scheduling Optimization Decision: Digital twin simulation shows that directly copying the production scheduling plan of the Vietnamese factory will result in 30% of the equipment being idle.

[0209] System Recommendation Strategy:

[0210] Prioritize deploying the production line for large battery models in the Kenyan factory;

[0211] Establish a direct link with the logistics center in Rwanda to shorten the delivery time;

[0212] Dynamically reserve 15% of the production capacity to handle sudden community group purchase orders;

[0213] By implementing the recommended strategy, the following improvement effects are obtained, as shown in Table 1:

[0214] Table 1: Improvement Effects

[0215]

[0216] Through this system, Global Tech has achieved a breakthrough in the East African market share from 0 to 17% within 6 months, verifying the commercial value of the digital twin decision-making system in small-sample scenarios.

[0217] The above describes the embodiments of the present invention. However, the present invention is not limited to the above specific embodiments. The above specific embodiments are merely illustrative and not restrictive. Under the inspiration of the present invention, those of ordinary skill in the art can also make many forms, all of which fall within the protection scope of the present invention.

Claims

1. An intelligent decision-making analysis method for a manufacturing system based on digital twin, characterized in that, It includes the following steps: S100: By extracting multiple groups of simulation tasks from the source domain data, train a base model with fast adaptation ability. The base model first fine-tunes the parameters on the support set of each small task, and then updates the initial parameters in reverse according to the performance on the query set of all tasks, and finally obtains a meta-knowledge base that can quickly adapt to the new market; S200: On the target domain small samples, adopt an adversarial training strategy to align the source domain and target domain features. The feature extractor is forced to learn domain-invariant representations, while the domain classifier tries to distinguish the data sources. At the same time, dynamically adjust the weights of the prediction loss and the adversarial loss to prevent negative transfer caused by overfitting of small samples; S300: Based on the scarce data in the target domain and the domain knowledge graph, use the conditional diffusion model to generate a large number of synthetic data, and construct an enhanced data set through distribution consistency testing and pre-trained model screening; S400: Load the enhanced data and the adapted model in the digital twin, run the multi-objective deep reinforcement learning algorithm, and dynamically adjust the weights of the order fulfillment rate and inventory cost in combination with real-time feedback, and finally select the optimal scheduling strategy from the Pareto optimal solution set; In S400, it includes the following steps: S410: Environment initialization: Load the virtual factory model and inject the enhanced data; S420: Multi-objective deep reinforcement learning strategy training: Construct an Actor-Critic architecture and optimize three objectives: Actor network: input state , output action ; Critic network: Evaluate the action value, including three output heads: ; Objective function: ; Among them, is the Actr network, and the parameter is , is the Critic network, and the parameter is , is the state vector, is the action vector, is the state distribution generated by the behavioral policy, is the variance penalty coefficient, is the variance of the action, is the cost evaluation value, is the quality of service evaluation value, is the feasibility evaluation value, represents the objective function, is the expectation operator, represents the variance of the action; S430: Dynamic weight adjustment mechanism: Adjust the target weights according to the real-time feedback of the digital twin; S440: Virtual-physical closed-loop verification: Execute the strategy in the digital twin to verify the feasibility; S450: Pareto front strategy selection: Select the optimal strategy from the non-dominated solution set.

2. The intelligent decision-making analysis method of a manufacturing system based on digital twin according to claim 1, characterized in that, In S100, it includes the following steps: S110, Meta-task construction: Sample meta-tasks from the source domain data, where each task contains a support set and a query set; S120: Internal parameter adaptation: For each meta-task, calculate the support set loss and update it based on the current meta-parameters; S130: External meta-parameter optimization: Evaluate the performance of the adapted model on the query set of all tasks, and update the meta-parameters by backpropagation; S140: Convergence determination: Repeat S110 - S130 until the stop condition is met.

3. The intelligent decision-making analysis method of a manufacturing system based on digital twin according to claim 2, wherein In S140, the stop conditions are as follows: ; wherein, is the convergence threshold, is the maximum number of iterations, , is the current iteration number, is the -th iteration meta-loss value, is the meta-loss value of the previous iteration.

4. The intelligent decision-making analysis method of a manufacturing system based on digital twin according to claim 1, characterized in that, In S200, it includes the following steps: S210: Model initialization: Load the meta-pre-trained parameters as the initial parameters; S220: Domain adversarial feature alignment: Achieve domain invariance in the feature space through the gradient reversal layer; S230: Joint optimization: Synchronously optimize the prediction task loss and the domain adversarial loss; S240: Parameter update: Alternately update the parameters of the feature extractor and the domain classifier: S250: Small sample early stopping strategy: Monitor the loss of the target domain validation set and terminate the training when there is no decrease for 5 consecutive iterations.

5. The intelligent decision-making analysis method of a manufacturing system based on digital twin according to claim 4, characterized in that, In S240, the calculation formula for parameter update is as follows: Fixed Update simultaneously : ; Fixed Update simultaneously : ; Among them, are the parameters of the feature extractor, are the parameters of the domain classifier, is the learning rate of adversarial learning, , is the gradient with respect to the parameter , is the gradient with respect to the parameter , is the loss function of domain adversarial, is the total loss function.

6. The intelligent decision-making analysis method of a manufacturing system based on digital twin according to claim 5, characterized in that, The calculation formula for the small sample early stopping strategy is as follows: Among them, is the loss value on the target domain validation set, is the current training round, represents the training termination signal, which is sent when the early stopping condition is triggered, means taking the minimum value among all historical rounds less than the current round t, represents the historical training round index, which is used to compare the current loss with the historical optimal loss.

7. The intelligent decision-making analysis method of a manufacturing system based on digital twin according to claim 1, characterized in that, In S300, it includes the following steps: S310: Domain knowledge conditional encoding: Encode the prior knowledge of the target domain into a conditional vector; S320: Conditional diffusion model training: Train the denoising diffusion probability model with the target domain samples and conditions as the input; S330, Diversity-Controllable Sample Generation: Sample enhanced data from the diffusion model, with constraints consistent with the target domain distribution; S340, Generated Sample Quality Verification: Screen valid samples through a pre-trained verifier; S350, Enhanced Dataset Construction: Merge real and synthetic data to construct the final training set.

8. The intelligent decision-making analysis method of a manufacturing system based on digital twin according to claim 1, wherein, The calculation formula for the optimal strategy is as follows: ; Among them, is the optimal strategy, is the Pareto optimal strategy set, is the weight of the th objective, is the Q value of the th objective, is the feasibility score of the strategy , is the system state, is the action output of the strategy in the state ; The output of the optimal strategy is as follows: ; Among them is the action for the start / stop decision of the production line, is the resource allocation ratio, is the production capacity allocation weight.

9. An intelligent decision-making analysis system for a manufacturing system based on digital twin, characterized in that Execute the steps in an intelligent decision-making analysis method for a digital twin-based manufacturing system as described in any one of claims 1-8, including: Meta-Learning Base Model Construction Module: Automatically construct multiple groups of simulation tasks from the historical data of mature markets, and obtain a meta-learning model with fast adaptation ability through collaborative training of internal and external loops, enabling it to quickly capture cross-market common laws with a small amount of target data; Adversarial Domain Adaptation Module: Introduce an adversarial training mechanism with a very small amount of data in the target domain, dynamically balance prediction accuracy and domain-invariant feature learning, and eliminate the distribution differences between the source domain and the target domain through gradient reversal and weight adaptation strategies; Data Augmentation Generation Module: Generate synthetic data that conforms to the characteristics of the target domain based on the conditional diffusion model and the domain knowledge graph, and expand high-quality training samples by combining distribution consistency verification and pre-trained model screening; Scheduling Optimization Decision Module: Simulate multi-objective scheduling strategies in a digital twin environment, dynamically optimize the order fulfillment rate, inventory cost, and feasibility score through reinforcement learning, and finally output a flexible production plan verified by virtualization.

Citation Information

Patent Citations

  • Decision control method and system for digital twin information of intelligent factory based on 5G driving

    CN114637262A

  • Intelligent clothing industry production regulation and control method and system based on data analysis

    CN118798494A