This invention discloses a method and
system for coordinating neighbor discovery conflicts in
wireless ad hoc networks based on Q-learning. The method includes: constructing a
wireless ad hoc
network model and initializing its parameters; based on the initialized
model parameters, determining the activation state of time slots according to the working sequence, running and updating relevant data, and determining whether to enter the next learning round based on the cumulative number of time slot activations; if entering the next learning round, calculating the average
reward value and determining the node state for the next learning round, updating the Q-table, and selecting actions according to the ε-greedy policy to achieve neighbor
information transmission and reception for nodes in the
wireless ad hoc
network model. This invention can effectively coordinate
beacon conflicts and significantly improve the neighbor discovery efficiency of nodes. As a method and
system for coordinating neighbor discovery conflicts in wireless ad hoc networks based on Q-learning, this invention can be widely applied in the field of communication technology.