The present application relates to the field of
drone control technology and discloses a method,
system and medium for cooperative task allocation among multiple drones. The method comprises: S1, defining the problem of cooperative task allocation among multiple drones; S2, solving the
optimal control strategy for the problem of cooperative task allocation among multiple drones; S3, solving the dual
optimization problem based on the primal-dual theory; S4, solving the convex dual
optimization problem using the Q-function and Schur
Complementation theory transforms the convex dual
optimization problem into a semidefinite
programming problem. S5, based on the properties of matrix congruence, transforms the semidefinite
programming problem into a model-free semidefinite
programming problem. S6, using a
solver, obtains the
optimal control strategy. S7, by varying the weight coefficients and repeating S2-S6, obtains the optimal
energy loss function bound. This application, based on the Q-learning method, requires only a small amount of data collection, not a precise dynamics model or
extensive data, to obtain the optimal strategy for multi-UAV cooperative task allocation.