Method for solving scheduling problem of distributed heterogeneous replacement flow shop

By combining the initial population strategy, local search operator and Q-learning learning mechanism methods, the problem that traditional scheduling methods are difficult to deal with special workpieces and complex factory constraints is solved, and efficient and adaptable distributed heterogeneous replacement flow workshop scheduling is achieved, improving production efficiency and stability.

CN120069483AActive Publication Date: 2025-05-30LIAOCHENG UNIV

Patent Information

Application Number
CN202510542127.4
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2025-04-28
Publication Date
2025-05-30
Estimated Expiration
2045-04-28

AI Technical Summary

Technical Problem

The traditional distributed heterogeneous replacement flow workshop scheduling method is difficult to effectively deal with special workpieces and complex factory constraints, resulting in insufficient feasibility and adaptability of the scheduling solution, and it is difficult to meet the actual needs of printed circuit board production and manufacturing.

Method used

A method combining initial population strategy, local search operator and Q-learning learning mechanism is adopted to generate high-quality and diverse solutions through efficient initialization population, diversity selection and dynamic local search optimization, especially considering the production scenarios of special artifacts, and optimizing the scheduling scheme.

Benefits of technology

It effectively improves the feasibility and adaptability of the scheduling plan, optimizes the overall production performance, and improves the production efficiency and the stability of the production line.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120069483A_ABST
    Figure CN120069483A_ABST
Patent Text Reader

Abstract

The invention relates to the technical field of distributed flow shop scheduling, and belongs to a solving method for a distributed heterogeneous replacement flow shop scheduling problem. Comprising the following steps: determining to take minimization of total weighting completion time as a solving target, initializing a population, and optimizing the population to obtain an optimized population; selecting a solution from the initial population by using a ternary tournament method, and executing six local search operators in sequence, and executing each local search operator for five times; replacing the solution with the maximum target value in the optimized population with a new solution to obtain a new population, if no improvement is continuously carried out for five times, triggering a disturbance mechanism to update the new population, combining the new population after disturbance with the new population before disturbance, sorting each solution in the combined population from small to large according to the target value of the solution, and selecting part to continue iteration; and continuing to execute the next round of search until the limited maximum time is reached, and outputting the solution with the minimum target value of the solution. The method has the positive effects of improving the production efficiency and the stability of the production line.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present invention relates to the technical field of distributed flow shop scheduling, and specifically belongs to a method for solving the distributed heterogeneous permutation flow shop scheduling problem. Background Art

[0002] Under the background of the rapid development of the global market economy, enterprises are facing increasingly severe market competition. The traditional centralized production mode gradually exposes the problem of insufficient flexibility and is difficult to meet the market demand for multi-variety and small-batch production. To adapt to this change, the manufacturing mode evolves towards distribution and heterogeneity to improve production efficiency and resource utilization rate. In this transformation process, efficient scheduling methods become the key link to ensure the smooth production process, optimize resource allocation, and enhance the competitiveness of enterprises.

[0003] In recent years, the distributed heterogeneous permutation flow shop scheduling problem has received extensive attention. This problem involves the collaborative production of multiple factories, and it is necessary to comprehensively consider the release time of workpieces, the scheduling constraints of heterogeneous factories, and the optimal strategy for cross-factory task allocation. At the same time, due to the complexity of actual factories, two special types of workpieces need to be considered specifically. One is the workpiece from VIP customers that needs to be preferentially guaranteed, and the other is the infrequently used workpiece that is about to be phased out. Due to the complexity of workpiece sequencing, factory selection, and machine allocation, traditional scheduling methods are difficult to meet the needs of actual production. Therefore, more efficient optimization strategies are urgently needed to improve the feasibility and adaptability of the scheduling plan, thereby optimizing the overall production performance. Summary of the Invention

[0004] The present invention aims to provide a method for solving the distributed heterogeneous permutation flow shop scheduling problem, which solves the problem that the current research on the distributed heterogeneous permutation flow shop scheduling problem does not conform to the actual production scenario and does not consider special workpieces, so as to achieve the purpose of taking the actual production scenario as the research premise, making production scheduling more reasonable, and thus effectively improving production efficiency and the stability of the production line.

[0005] A method for solving the distributed heterogeneous permutation flow shop scheduling problem provided by the present invention is characterized by including the following steps: Step 1, analyze the characteristics of the distributed heterogeneous permutation flow shop scheduling problem in printed circuit board production and manufacturing, determine the objective of minimizing the total weighted completion time, and initialize the parameters, including the population size PSize , time parameters t ; Step 2, initialize the population, use the rule of the sum of the shortest processing time and the release time, the maximum weight rule, and the minimum release time rule, and combine the greedy insertion method to generate the first three solutions of the initial population, use the composite method to generate the fourth solution of the initial population, and generate the remainingPSize - Four solutions are obtained, and the backtracking optimization method is applied to the initial population to obtain the optimized population; Step 3: Select a solution from the initial population using the ternary tournament method and sequentially execute six local search operators, each local search operator is executed 5 times. Among them, after each execution of the local search operator, the Q-learning learning mechanism is executed. If the new solution obtained is better than the original solution, the Q value is increased; if the new solution obtained is worse than the original solution, the Q value is decreased; if the new solution obtained is equal to the original solution, the Q value remains unchanged. After 5 local search operators are executed, the Q values corresponding to each local search operator are sorted, and the local search operator corresponding to the highest Q value is selected and executed once. If the objective value of the new solution obtained after the selected local search operator is executed is smaller, the new solution is retained; Step 4: Replace the solution with the largest objective value in the optimized population with the new solution to obtain a new population. If there is no improvement for 5 consecutive times, the perturbation mechanism is triggered to update the new population, and the perturbed new population and the new population before perturbation are combined to obtain a combined population with a size of 2 × PSize , Sort each solution in the combined population in ascending order of the objective value of the solution, and select the first PSize solutions and continue the iteration; Step 5: Iteratively execute Step 3 and Step 4, continue to perform the next round of search until the limited maximum time is reached, output the solution with the smallest objective value, and terminate the execution.

[0006] Furthermore, in Step 2, the rule for the sum of the shortest processing time and the release time is to determine the priority according to the sum of the total processing time and the release time of the workpiece, and sort the workpieces in ascending order of the sum of the total processing time and the release time of the workpiece; The maximum weight rule is to determine the priority according to the weight of the workpiece, and sort the workpieces in descending order of the weight of the workpiece; The minimum release time rule is to determine the priority according to the release time of the workpiece, and sort the workpieces in ascending order of the release time of the workpiece; The greedy insertion method means that for the workpiece sequence sorted by each sorting rule, a workpiece is sequentially selected and inserted into the position that minimizes the objective value of the solution in all factory sequences. The process of selecting workpieces and inserting is repeated until all the workpieces in the workpiece sequences sorted by the sum of the shortest processing time and the release time rule, the maximum weight rule, and the minimum release time rule are inserted, and the first three solutions are obtained; The implementation process of the composite method includes the following steps (1) Sort all workpieces according to the maximum weight rule to obtain the initial workpiece sequence; (2) Extract workpieces one by one from the initial workpiece sequence. When extracting each workpiece, check whether there are other workpieces with the same weight as the extracted workpiece; (3) If there are no other workpieces with the same weight, directly insert the extracted workpiece into the position in all factory sequences that minimizes the objective value of the solution, and continue to extract the next workpiece; (4) If there are other workpieces with the same weight, put the workpieces with the same weight into a new workpiece sequence, reorder them according to the rule of the sum of the shortest processing time and the release time, and insert the workpieces after the new ordering into the positions in all factory sequences that can minimize the objective value of the obtained solution one by one until all the workpieces in the new workpiece sequence are completed in the insertion process; (5) According to the order of the workpieces in the initial workpiece sequence, select a workpiece that has not been extracted by the initial workpiece sequence and the new workpiece sequence, and repeat steps (3) and (4) until all workpieces are completed in the insertion process, thereby generating the fourth solution; The random method randomly sorts all workpieces, and then inserts the sorted workpieces into the positions in all factory sequences that minimize the objective value of the solution one by one; In the backtracking optimization method, select a solution with the smallest objective value in the initial population as the initial candidate set. If there are multiple solutions with the same and smallest objective values, all of them are added to the initial candidate set as the new candidate set. Randomly select a solution from the initial population, remove the workpieces in it one by one, and reinsert the removed workpieces into the positions in all factory sequences that can minimize the objective value of the obtained solution. During the insertion process, if a better solution is generated, replace the original solution with the better solution. If the better solution is not greater than the solution with the smallest objective value in the initial population, add it to the new candidate set to obtain the preferred candidate set, and select the first untested solution from the preferred candidate set, and repeat the process of removing and inserting workpieces. When the solutions obtained continuously three times are all greater than the solution with the smallest objective value in the initial population, select a solution that has not been selected yet from the preferred candidate set to start exploration. If there is no solution that meets the requirements in the preferred candidate set at this time, end the backtracking process and update the solution with the smallest objective value.

[0007] Further, in step 3, the ternary tournament method means randomly selecting 3 solutions from the initial population, comparing the objective values of the 3 solutions, and determining the one with the smallest objective value among the 3 solutions. If there are multiple solutions with the same and smallest objective values, select the first selected solution with the smallest objective value to execute the local search operator; The local search operators include the critical factory optimization operator, the large impact workpiece optimization operator, the disruption and reconstruction operator, the two-stage optimization operator, the special workpiece optimization operator, and the inter-factory exchange operator. The critical factory refers to the factory with the largest objective value of the solution among all factories. If there are multiple factories with the same objective value of the solution, the factory with the smallest factory number is selected. The objective value of the solution of a factory means adding the product of the weight of each workpiece assigned to a factory and the corresponding completion time of each workpiece. Among them, The operation process of the critical factory optimization operator is to remove all workpieces in the critical factory and randomly insert the removed workpieces into the position in the sequence of all factories that can minimize the objective value of the obtained solution. If the quality of the solution is improved through the removal and insertion operations, the obtained new solution is retained; otherwise, the original solution is restored, and the removal and insertion operations are performed at most five times. Finally, the best solution among the five removal and insertion operations is output; The operation process of the large impact workpiece optimization operator is to remove the top 50% of the workpieces with the greatest impact on the objective value of the solution from the current solution and randomly insert them into the position in the sequence of all factories that can minimize the objective value of the obtained solution. If the quality of the solution is improved through the removal and insertion operations, the obtained new solution is retained; otherwise, the original solution is restored, and the removal and insertion operations are performed at most five times. Finally, the best solution among the five removal and insertion operations is output; The operation process of the disruption and reconstruction operator is to randomly remove a part of the workpieces and re-insert the removed workpieces into the position in the sequence of all factories that can minimize the objective value of the obtained solution. If the quality of the obtained solution is improved through the removal and insertion operations, the obtained new solution is retained; otherwise, the original solution is restored, and the removal and insertion operations are performed at most five times. Finally, the best solution among the five removal and insertion operations is output; The operation process of the two-stage optimization operator is divided into two stages. The first stage is the same as the disruption and reconstruction operator. In the second stage, the workpieces in all factories are exchanged in turn. If the exchange operation improves the quality of the solution, the obtained new solution is retained; otherwise, the solution before the exchange is restored, and the exchange operation is performed at most five times. Finally, the best solution among the five exchange operations is output; The operation process of the special workpiece optimization operator is to randomly select a factory and exchange the positions of two special workpieces and other workpieces in turn. If the obtained solution is improved, it is updated; otherwise, the original solution is restored.

[0008] The operation process of the inter-factory exchange operator is to select the top two factories with the largest objective value of the solution of the factory, exchange the latter half of the workpieces in the processing sequences of the two factories, respectively transfer the exchanged workpieces to the new factories, and re-insert them into the position in the new factory sequence that can minimize the objective value of the obtained solution. If the obtained solution is improved, it is updated; otherwise, the original solution is restored; The execution process of the Q-learning learning mechanism includes recording the results of each local search operator each time when 6 local search operators are executed 5 times in step 3. If the new solution obtained is better than the original solution, the Q value is rewarded. If the new solution obtained is worse than the original solution, the Q value is punished. If the new solution obtained is equal to the original solution, the Q value remains unchanged. After executing the five local searches, the Q values corresponding to each local search operator are sorted, and the local search operator corresponding to the highest Q value is selected to execute once. If the objective value of the new solution is better, the best solution is updated.

[0009] Furthermore, the perturbation mechanism includes two perturbation operators, the shift perturbation operator and the swap perturbation operator. Each time, a perturbation operator is randomly selected for execution. The shift perturbation operator means randomly selecting a workpiece in the current solution and moving the selected workpiece to a position different from the selected workpiece in the current solution. The swap perturbation operator means randomly selecting two workpieces and swapping the positions of the two selected workpieces in the solution.

[0010] Furthermore, the limited maximum time is t × n × m, n where is the total number of workpieces, m and is the total number of machines in each factory.

[0011] A method for solving the distributed heterogeneous permutation flow shop scheduling problem provided by the present invention designs an efficient initial population strategy. Through a heuristic method based on problem characteristics, an initial population with high quality and diversity is generated, and a backtracking optimization method is executed on the initial population; the diversity of the selected solutions is ensured through the ternary tournament method, and the quality of the selected solutions is further improved; by combining 6 different local search operator strategies and the Q-learning learning mechanism, the local search operator corresponding to the highest Q value is dynamically selected for execution, thereby enhancing the search ability and convergence performance of the method. In summary, the application of the present invention particularly considers the production scenarios with special workpieces, can effectively solve the scheduling problem in the production and manufacturing of printed circuit boards, continuously optimize the objective value of the solution, and has the positive effect of improving production efficiency and enhancing the overall performance of the production line. Brief Description of the Drawings

[0012] Figure 1 is the implementation flowchart of the present invention; Figure 2 is the implementation flowchart of the composite method of the present invention; Figure 3 is the Gantt chart of the embodiment of the present invention; Figure 4 is the mean value chart of the present invention and 5 existing comparison algorithms. Detailed Embodiment

[0013] As Figure 1 and Figure 2 shown, a method for solving the distributed heterogeneous permutation flow shop scheduling problem provided by the present invention is mainly implemented through the following steps.

[0014] Step 1: Analyze the characteristics of the distributed heterogeneous permutation flow shop scheduling problem in printed circuit board production and manufacturing, determine the objective of minimizing the total weighted completion time, and initialize the parameters, including the population size PSize , time parameters t .

[0015] Step 2: Initialize the population. Use the rule of the sum of the shortest processing time and the release time, the maximum weight rule, and the minimum release time rule, and combine the greedy insertion method to generate the first three solutions of the initial population. Use the composite method to generate the fourth solution of the initial population, and generate the remaining PSize -4 solutions of the initial population by the random method, and apply the backtracking optimization method to the initial population to optimize it to obtain the optimized population. Among them: The rule of the sum of the shortest processing time and the release time is to determine the priority according to the sum of the total processing time and the release time of the workpiece, and sort the workpieces in ascending order according to the sum of the total processing time and the release time of the workpiece; The maximum weight rule is to determine the priority according to the weight of the workpiece, and sort the workpieces in descending order according to the weight of the workpiece; The minimum release time rule is to determine the priority according to the release time of the workpiece, and sort the workpieces in ascending order according to the release time of the workpiece; The greedy insertion method means that for the workpiece sequence sorted by each sorting rule, select a workpiece in turn and insert it into the position in all factory sequences that makes the objective value of the solution the smallest. Repeat the process of selecting workpieces and inserting until all the workpieces in the workpiece sequences sorted by the rule of the sum of the shortest processing time and the release time, the maximum weight rule, and the minimum release time rule are inserted to obtain the first three solutions; The implementation process of the composite method includes the following steps, (1) Sort all workpieces according to the maximum weight rule to obtain the initial workpiece sequence; (2) Extract workpieces one by one from the initial workpiece sequence. When extracting each workpiece, check whether there are other workpieces with the same weight as the extracted workpiece; (3) If there are no other workpieces with the same weight, directly insert the extracted workpiece into the position in all factory sequences that makes the objective value of the solution the smallest, and continue to extract the next workpiece; (4) If there are other workpieces with the same weight, place the workpieces with the same weight into a new workpiece sequence, reorder them according to the rule of the sum of the shortest processing time and the release time, and insert the workpieces after the new reordering into the position in all factory sequences that can minimize the objective value of the obtained solution, until all the workpieces in the new workpiece sequence have completed the insertion process; (5) According to the order of the workpieces in the initial workpiece sequence, select a workpiece that has not been extracted by the initial workpiece sequence and the new workpiece sequence, and repeat steps (3) and (4) until all the workpieces have completed the insertion process, thereby generating the fourth solution; The random method randomly sorts all the workpieces, and then inserts the sorted workpieces into the position in all factory sequences that can minimize the objective value of the solution one by one; In the backtracking optimization method, select a solution with the minimum objective value in the initial population as the initial candidate set. If there are multiple solutions with the same minimum objective value, all of them are added to the initial candidate set as the new candidate set. Randomly select a solution from the initial population, remove the workpieces in it one by one, and reinsert the removed workpieces into the position in all factory sequences that can minimize the objective value of the obtained solution. During the insertion process, if a better solution is generated, replace the original solution with the better solution. If the better solution is not greater than the solution with the minimum objective value in the initial population, add it to the new candidate set to obtain the preferred candidate set, and select the first untested solution from the preferred candidate set, and repeat the process of removing and inserting workpieces. When the solutions obtained continuously three times are all greater than the solution with the minimum objective value in the initial population, select a solution that has not been selected yet from the preferred candidate set and start exploration. If there is no solution that meets the requirements in the preferred candidate set at this time, end the backtracking process and update the solution with the minimum objective value.

[0016] Step 3: Select a solution from the initial population using the ternary tournament method and sequentially execute 6 local search operators, with each local search operator being executed 5 times. Among them, after each execution of the local search operator, the Q-learning learning mechanism is executed. The ternary tournament method means randomly selecting 3 solutions from the initial population, comparing the objective values of the 3 solutions, and determining the one with the smallest objective value among the 3 solutions. If there are multiple solutions with the same and smallest objective values, select the first selected solution with the smallest objective value to execute the local search operator. The local search operators include the critical factory optimization operator, the large impact workpiece optimization operator, the disruption and reconstruction operator, the two-stage optimization operator, the special workpiece optimization operator, and the inter-factory exchange operator. The critical factory refers to the factory with the largest objective value of the solution among all factories. If there are multiple factories with the same objective value of the solution, select the factory with the smallest factory number. The objective value of the solution of a factory means adding the product of the weight of each workpiece assigned to a factory and the corresponding completion time of each workpiece. Among them, the operation process of the critical factory optimization operator is to remove all workpieces in the critical factory and randomly insert the removed workpieces into the position in all factory sequences that can make the objective value of the obtained solution the smallest. If the quality of the solution is improved through the removal and insertion operations, retain the obtained new solution; otherwise, restore the original solution. Perform the removal and insertion operations at most five times, and finally output the best solution among the 5 removal and insertion operations; The operation process of the large impact workpiece optimization operator is to remove the first 50% of the workpieces with the greatest impact on the objective value of the solution from the current solution and randomly insert them into the position in all factory sequences that can make the objective value of the obtained solution the smallest. If the quality of the solution is improved through the removal and insertion operations, retain the obtained new solution; otherwise, restore the original solution. Perform the removal and insertion operations at most five times, and finally output the best solution among the 5 removal and insertion operations; The operation process of the disruption and reconstruction operator is to randomly remove a part of the workpieces and re-insert the removed workpieces into the position in all factory sequences that can make the objective value of the obtained solution the smallest. If the quality of the obtained solution is improved through the removal and insertion operations, retain the obtained new solution; otherwise, restore the original solution. Perform the removal and insertion operations at most five times, and finally output the best solution among the 5 removal and insertion operations; The operation process of the two-stage optimization operator is divided into two stages. The first stage is the same as the disruption and reconstruction operator. In the second stage, the workpieces within all factories are sequentially exchanged. If the exchange operation improves the quality of the solution, retain the obtained new solution; otherwise, restore the solution before the exchange. Perform the exchange operation at most five times, and finally output the best solution among the 5 exchange operations; The operation process of the special workpiece optimization operator is to randomly select a factory and sequentially exchange the positions of two special workpieces and other workpieces. If the obtained solution is improved, update it; otherwise, restore the original solution; The operation process of the factory - to - factory exchange operator is as follows: select the top two factories with the largest objective values of the solutions of the factories, exchange the latter - half workpieces in the processing sequences of the two factories. Respectively, the exchanged workpieces are transferred to the new factories and re - inserted into the positions in the new factory sequences that can minimize the objective value of the obtained solution. If the obtained solution is improved, it is updated; otherwise, the original solution is restored.

[0017] The execution process of the Q - learning learning mechanism includes: when executing 6 local search operators 5 times in step 3, record the results of each local search operator each time. According to the comparison result between the obtained new solution and the original solution, update the Q - value corresponding to the local search operator. Specifically: If the obtained new solution is better than the original solution, the Q - value is rewarded ; If the obtained new solution is worse than the original solution, the Q - value is punished ; If the obtained new solution is equal to the original solution, the Q - value remains unchanged; After executing the five - time local search, sort the Q - values corresponding to each local search operator, and select the local search operator with the highest corresponding Q - value to execute once. If the objective value of the new solution is better, update the best solution.

[0018] Step 4: Replace the solution with the largest objective value in the optimized population with the new solution to obtain a new population. If there is no improvement for 5 consecutive times, trigger the perturbation mechanism to update the new population, and combine the perturbed new population with the new population before perturbation to obtain a combined population with a size of 2 × PSize The combined population , Sort each solution in the combined population in ascending order according to the objective value of the solution, and select the first PSize solutions to continue the iteration. Among them, the perturbation mechanism includes two perturbation operators, the shift perturbation operator and the exchange perturbation operator. Each time, randomly select one perturbation operator to execute. The shift perturbation operator means randomly select a workpiece in the current solution and move the selected workpiece to a position different from the selected workpiece in the current solution. The exchange perturbation operator means randomly select two workpieces and exchange the positions of the two selected workpieces in the solution.

[0019] Step 5: Iteratively execute step 3 and step 4, continue to execute the next - round search until the limited maximum time t × n × m is reached, and output the solution with the smallest objective value to terminate the execution. Among them, n is the total number of workpieces, m is the total number of machines in each factory.

[0020] To better prove the effectiveness of the present invention, the following will further describe and explain the present invention through experimental analysis of a series of examples of the present invention.

[0021] The test data includes 225 large-scale instances, which are created based on the total number of factories f , the total number of machines in each factory m , and the total number of workpieces n . Specifically, f ∈ {2, 3, 4}, n ∈ {20, 40, 60, 80, 100}, m ∈ {3, 5, 7}, and there are 5 different test cases for each parameter combination { f , n , m}. In the first factory, the processing time of each workpiece is randomly generated within the range of [10, 40]. In subsequent factories, the processing time of each workpiece on each machine either remains unchanged or randomly increases by 0.05, 0.1, or 0.2 on the basis of the previous factory. The release time of the workpiece is randomly generated within the range of [0, 5 n , and at least f workpieces have a release time of 0. When allocating factories, two special workpieces have been fixedly allocated to the corresponding specific factories and can only perform operations within the specific factories. CPU time is used as the termination criterion for the comparison algorithm, and the termination criterion is set to t × n × m milliseconds, where t is a multiple, set to 60. To better solve and optimize the distributed heterogeneous permutation flow shop scheduling problem based on printed circuit board production, in terms of parameter settings, PSize = 5, and in the local search operator, the number of workpieces randomly removed by the destruction and reconstruction operator is 6.

[0022] To verify the effectiveness of the proposed theoretical results, an instance in the simulation data was selected. The manufacturer consists of two factories, each with two machines. There are 10 workpieces in the current order that need to be processed, and the specific information of their processing time, weight, and release time is shown in Table 1: Table 1 Specific Information of the Manufacturer

[0023] In Table 1, represents the processing time of workpiece j on machine k in factory i , j is the workpiece number,j = 1, 2, ..., n , k is the factory number, k = 1, 2, ..., f , i is the machine number, i = 1, 2, ..., m , is the workpiece j 's release time, is the workpiece j 's weight. Among them, workpiece 4, workpiece 6, and workpiece 10 are special workpieces. Workpiece 4 and workpiece 10 can only be processed in factory 1, and workpiece 6 can only be processed in factory 2.

[0024] Figure 3 is the Gantt chart of the embodiment of the present invention, used to show the workpiece arrangement of different machines in two factories. The vertical axis of the chart represents different factories and machines, and the horizontal axis represents the processing time. Each workpiece task is represented in the form of a bar chart. The left endpoint and the right endpoint of the bar respectively represent the start time and the completion time of the workpiece. represents the workpiece j in the factory k on the machine i 's completion time, and the unit is unit time. The completion times of workpiece 1 to workpiece 10 are 172, 118, 49, 132, 111, 80, 125, 82, 39, 158 respectively. The workpiece processing sequence in factory 1 is 3, 8, 2, 4, 10, and the workpiece processing sequence in factory 2 is 9, 6, 5, 7, 1.

[0025] To calculate the objective value of the solution, we need to multiply the weight of each workpiece by the completion time corresponding to each workpiece and add up all the products. The following is the specific calculation process, which can be obtained from Figure 3 the completion times of the workpieces in and the weights of the workpieces in Table 1: The completion time of workpiece 1 is 172 unit times, and the weight is 4. The product is 4 × 172 = 688; The completion time of workpiece 2 is 118 unit times, and the weight is 5. The product is 5 × 118 = 590; The completion time of workpiece 3 is 49 unit times, and the weight is 7. The product is 7 × 49 = 343; The completion time of workpiece 4 is 132 unit times, and the weight is 4. The product is 4 × 132 = 528; The completion time of workpiece 5 is 111 unit times, and the weight is 5. The product is 5 × 111 = 555; The completion time of workpiece 7 is 125 unit time, with a weight of 4, and the product is 4×125 = 500; The completion time of workpiece 8 is 82 unit time, with a weight of 5, and the product is 5×82 = 410; The completion time of workpiece 9 is 39 unit time, with a weight of 9, and the product is 9×39 = 351; The completion time of workpiece 10 is 158 unit time, with a weight of 1, and the product is 1×158 = 158; Adding these products together, we get the objective value of the solution: 688 + 590 + 343 + 528 + 555 + 560 + 500 + 410 + 351 + 158 = 4683.

[0026] The experimental results and analysis of this example are as follows. After setting the parameters for the algorithm (QMA) for solving and optimizing the distributed heterogeneous permutation flow shop scheduling problem in the present invention, it is experimentally compared with an improved fruit fly algorithm (DFFO), three improved IG algorithms (IIG, CMSIG, IGP), and an improved evolutionary algorithm (NEA). In order to make the existing 5 comparison algorithms adapt to the problems to be solved, necessary modifications are required, including using unified instances, adopting the same objective value, handling special workpieces, etc., and following the details of their respective original algorithms, selecting the algorithm type, f 、 n and m as analysis factors, and comparing the instance results. In order to evaluate the performance of the algorithms, each instance is independently run 5 times, and the relative percentage increase (RPI) is calculated as the evaluation criterion. The calculation formula of the RPI value is , T is the objective value of the solution obtained by a certain algorithm, is the minimum objective value of the solutions obtained by the six comparison algorithms. Obviously, the smaller the RPI value, the better the performance of the algorithm. At the same time, the average RPI (ARPI) is also used as the evaluation criterion. The comparison results of the ARPI values of the six algorithms are shown in Table 2: Table 2 ARPI values of six algorithms Type CMSIG IIG IGP DFFO NEA QMA =2 4.649 10.025 4.352 2.589 9.226 1.942 =3 4.030 9.236 3.899 3.313 9.228 2.312 =4 3.912 8.213 2.869 3.568 8.895 2.397 =20 1.119 1.547 1.538 1.057 6.370 1.177 =40 2.843 7.010 2.261 2.617 9.472 2.131 =60 3.965 9.896 3.529 3.432 9.961 2.499 =80 5.843 12.981 4.631 4.064 10.021 2.485 =100 7.216 14.357 6.575 4.613 9.759 2.795 =3 2.336 10.946 4.721 4.252 10.381 2.624 =5 5.466 8.368 3.443 2.797 8.949 2.092 =7 4.790 8.161 2.957 2.421 8.020 1.936 Average value 4.197 9.158 3.707 3.156 9.116 2.217 It can be seen from the data in Table 2 that for all scales of data, the QMA algorithm performs better than the existing 5 comparison algorithms in most problems, showing good performance. Moreover, for different total numbers of factories, the QMA algorithm of the present invention shows better performance, indicating that the QMA algorithm can effectively handle the scheduling problems of distributed factories.

[0027] Figure 4 is the mean value graph of the present invention and the existing algorithms. From Figure 4It can be seen that the QMA algorithm of the present invention is significantly superior to the existing five comparison algorithms statistically, demonstrating good adaptability and optimization ability. Generally speaking, the QMA algorithm of the present invention shows good performance in data analysis.

[0028] In summary, the QMA algorithm of the present invention demonstrates excellent performance, successfully solves a distributed heterogeneous permutation flow shop scheduling problem considering release time and special jobs, provides an innovative solution, offers strong support for optimal scheduling in actual production scenarios, and provides a feasible method for improving production efficiency and reducing costs.

Claims

1. A method for solving the distributed heterogeneous permutation flow shop scheduling problem, characterized in that: The following steps are included: Step 1: Analyze the characteristics of the distributed heterogeneous permutation flow shop scheduling problem in printed circuit board manufacturing, determine the solution goal to minimize the total weighted completion time, and initialize the parameters, including the population size. PSize , time parameter t ; Step 2: Initialize the population, use the shortest processing time and release time rule, the maximum weight rule, and the minimum release time rule, combined with the greedy insertion method to generate the first three solutions of the initial population, use the composite method to generate the fourth solution of the initial population, and use the random method to generate the remaining solutions of the initial population. PSize -4 solutions, and apply the backtracking optimization method to the initial population to obtain the optimized population; Step 3: Use the ternary tournament method to select a solution from the initial population, and execute 6 local search operators in sequence, each of which is executed 5 times. After each execution of the local search operator, the Q-learning mechanism is executed. If the new solution is better than the original solution, the Q value is increased; if the new solution is worse than the original solution, the Q value is reduced; if the new solution is equal to the original solution, the Q value remains unchanged. After the 5 local search operators are executed, the Q values ​​corresponding to each local search operator are sorted, and the local search operator with the highest corresponding Q value is selected to execute once. If the target value of the new solution obtained after the selected local search operator is executed is smaller, the new solution is retained. Step 4: Replace the solution with the largest target value in the optimized population with the new solution to obtain a new population. If there is no improvement for 5 consecutive times, the perturbation mechanism is triggered to update the new population, and the new population after the perturbation is combined with the new population before the perturbation to obtain a population of size 2 × PSize The combined population , Sort each solution in the combined population from small to large according to the target value of the solution, and select the first PSize solution, and continue to iterate; Step 5, iteratively execute steps 3 and 4, and continue the next round of search until the maximum time is reached, output the solution with the smallest target value, and terminate the execution.

2. The method for solving the distributed heterogeneous permutation flow shop scheduling problem according to claim 1 is characterized in that: The shortest sum of processing time and release time rule is to determine the priority according to the sum of the total processing time and release time of the workpiece, and sort the workpieces from small to large according to the sum of the total processing time and release time of the workpiece; The maximum weight rule is to determine the priority according to the weight of the workpiece, and sort the workpieces from the largest to the smallest weight; The minimum release time rule is to determine the priority according to the release time of the workpiece, and sort the workpieces from the smallest to the largest release time; The greedy insertion method means that for each workpiece sequence sorted by the sorting rule, one workpiece is selected in turn and inserted into the position in all factory sequences where the target value of the solution is minimized. The process of selecting and inserting workpieces is repeated until all the workpieces in the workpiece sequence sorted by the shortest processing time and release time rule, the maximum weight rule, and the minimum release time rule are inserted, and the first three solutions are obtained.

3. The method for solving the distributed heterogeneous permutation flow shop scheduling problem according to claim 2 is characterized in that: The implementation process of the composite method includes the following steps: (1) sorting all artifacts according to the maximum weight rule to obtain an initial artifact sequence; (2) extracting artifacts one by one from the initial artifact sequence, and when extracting each artifact, checking whether there are other artifacts with the same weight as the extracted artifact; (3) if there are no other artifacts with the same weight, directly insert the extracted artifact into the position in all factory sequences that minimizes the target value of the solution, and continue to extract the next artifact; (4) if there are other artifacts with the same weight, put the artifacts with the same weight into a new artifact sequence, and re-sort them according to the rule of the sum of the shortest processing time and the release time, and insert the newly sorted artifacts into the positions in all factory sequences that minimize the target value of the solution, until all artifacts in the new artifact sequence complete the insertion process; (5) according to the order of artifacts in the initial artifact sequence, select a artifact that has not been extracted by the initial artifact sequence and the new artifact sequence, and repeat steps (3) and (4) until all artifacts complete the insertion process, generating a fourth solution.

4. The method for solving the distributed heterogeneous permutation flow shop scheduling problem according to claim 3 is characterized in that: In the backtracking optimization method, a solution with the smallest target value in the initial population is selected as the initial candidate set. If there are multiple solutions with the smallest and identical target values, they are all added to the initial candidate set as the new candidate set. A solution is randomly selected from the initial population, and the workpieces therein are removed one by one. The removed workpieces are reinserted into the position in all factory sequences where the target value of the obtained solution can be minimized. During the insertion process, if a better solution is generated, the original solution is replaced by the better solution. If the better solution is not greater than the solution with the smallest target value in the initial population, it is added to the new candidate set to obtain the preferred candidate set, and the first untested solution is selected from the preferred candidate set. The process of removing and inserting workpieces is repeated. When the solutions obtained for three consecutive times are all greater than the solution with the smallest target value in the initial population, a solution that has not been selected is reselected from the preferred candidate set to start exploration. If there is no solution that meets the requirements in the preferred candidate set at this time, the backtracking process is terminated and the solution with the smallest target value is updated.

5. A method for solving the distributed heterogeneous permutation flow shop scheduling problem according to claim 4, characterized in that: The ternary tournament method randomly selects three solutions from the initial population, compares the target values ​​of the three solutions, and determines the one with the smallest target value among the three solutions. If there are multiple solutions with the same target value and they are all the smallest, the first selected solution with the smallest target value is selected to execute the local search operator, where: The local search operators include the key factory optimization operator, the large-impact workpiece optimization operator, the destruction and reconstruction operator, the two-stage optimization operator, the special workpiece optimization operator, and the inter-factory exchange operator. The key factory refers to the factory with the largest target value of the factory solution among all factories. If there are multiple factories with the same target value of the solution, the factory with the smallest factory number is selected. The target value of the factory solution refers to the sum of the weight of each workpiece assigned to a factory and the product of the completion time corresponding to each workpiece, where The operation process of the key factory optimization operator is to remove all the workpieces in the key factory and randomly insert the removed workpieces into the position in all factory sequences that can minimize the target value of the obtained solution. If the quality of the solution is improved by the removal and insertion operation, the new solution is retained; otherwise, the original solution is restored, and the removal and insertion operations are performed at most five times, and finally the best solution among the five removal and insertion operations is output; The operation process of the large-impact workpiece optimization operator is to remove the top 50% of the workpieces that have the greatest impact on the target value of the solution from the current solution, and randomly insert them into the position in all factory sequences that can minimize the target value of the solution. If the quality of the solution is improved by the removal and insertion operation, the new solution is retained; otherwise, the original solution is restored, and a maximum of five removal and insertion operations are performed, and finally the best solution among the five removal and insertion operations is output; The operation process of the destructive reconstruction operator is to randomly remove a part of the workpieces and reinsert the removed workpieces into the position in all factory sequences that can minimize the target value of the obtained solution. If the quality of the obtained solution is improved by the removal and insertion operation, the new solution is retained; otherwise, the original solution is restored, and a maximum of five removal and insertion operations are performed, and finally the best solution among the five removal and insertion operations is output; The operation process of the two-stage optimization operator is divided into two stages. The first stage is the same as the destruction and reconstruction operator. In the second stage, all the workpieces in the factory are exchanged in turn. If the exchange operation improves the quality of the solution, the new solution is retained; otherwise, the solution before the exchange is restored. A maximum of five exchange operations are performed, and the best solution among the five exchange operations is finally output. The operation process of the special workpiece optimization operator is to randomly select a factory, exchange the positions of the two special workpieces and other workpieces in turn, and update if the solution obtained is improved, otherwise restore the original solution; The operation process of the inter-factory exchange operator is to select the first two factories whose solution target values ​​are the largest, exchange the second half of the workpieces in the processing sequences of the two factories, transfer the exchanged workpieces to the new factories, and reinsert them into the new factory sequence at a position where the target value of the solution can be minimized. If the solution is improved, it is updated, otherwise the original solution is restored.

6. A method for solving the distributed heterogeneous permutation flow shop scheduling problem according to claim 5, characterized in that: The Q-learning learning mechanism execution process includes, when step 3 executes 6 local search operators 5 times, recording the results of each local search operator each time, and if the new solution is better than the original solution, the Q value is rewarded , if the new solution is worse than the original solution, the Q value is penalized If the new solution is equal to the original solution, the Q value remains unchanged. After executing five local searches, sort the Q values ​​corresponding to each local search operator, select the local search operator with the highest Q value and execute it once. If the target value of the new solution is better, update the best solution.

7. A method for solving the distributed heterogeneous permutation flow shop scheduling problem according to claim 6, characterized in that: The perturbation mechanism includes two types of perturbation operators, the shift perturbation operator and the exchange perturbation operator. One perturbation operator is randomly selected to execute each time. The shift perturbation operator refers to randomly selecting a workpiece in the current solution and moving the selected workpiece to a position different from the selected workpiece in the current solution. The exchange perturbation operator refers to randomly selecting two workpieces and exchanging the positions of the two selected workpieces in the solution.

8. The method for solving the distributed heterogeneous permutation flow shop scheduling problem according to claim 1 is characterized in that: The maximum time limit is t × n × m,n is the total number of workpieces, m is the total number of machines in each factory.

Citation Information

Patent Citations

  • Multi-target distributed hybrid flow shop scheduling method

    CN115933568A

  • Distributed blocking flow shop scheduling optimization system based on reinforcement learning

    CN116700176A

  • Distributed heterogeneous flow shop scheduling method based on improved hybrid memetic algorithm

    CN117035364A

  • Distributed heterogeneous flow shop scheduling method based on hybrid initialization memetic algorithm

    CN117077975A

  • Solving method for batch scheduling of distributed reentrant heterogeneous hybrid flow shop

    CN117829550A

Cited By

  • Building energy control method and system

    CN120972706A

  • A building energy control method and system

    CN120972706B

  • Full-active-scheduling multi-workshop combined scheduling method and system and storage medium

    CN122219381A