Distributed dynamic spectrum access method based on multi-agent reinforcement learning

A dynamic spectrum access, multi-agent technology, applied in neural learning methods, character and pattern recognition, instruments, etc., can solve the problems of primary user interference and low throughput of communication systems, reduce access conflicts, and improve utilization Efficiency, the effect of scalable network scale

CN113923794APending Publication Date: 2022-01-11NAT UNIV OF DEFENSE TECH
0 Cites 0 Cited by

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Publication Date
2022-01-11

Smart Images

  • Figure 1
    Figure 1
  • Figure 2
    Figure 2
  • Figure 3
    Figure 3
Patent Text Reader

Abstract

The invention discloses a distributed dynamic spectrum access method based on multi-agent reinforcement learning, and the method enables a multi-user distributed dynamic spectrum access problem to be modeled into a multi-agent Markov cooperative game model, and constructs a centralized training and distributed execution multi-agent reinforcement learning framework. The multi-agent reinforcement learning framework comprises an offline training module and an online execution module, the online execution module performs spectrum access of the cognitive user by using a learned access strategy, and the offline training module dynamically updates the online execution module according to a spectrum access result of the cognitive user. The invention provides a multi-user cooperative spectrum access method with self-adaptive communication environment and extensible network scale, which reduces access conflicts among cognitive users while avoiding interference to authorized users, thereby maximizing the access success rate of the cognitive users and improving the utilization efficiency of the spectrum.
Need to check novelty before this filing date? Find Prior Art

Description

technical field

[0001] The invention relates to the technical field of wireless communication networks, in particular to a distributed dynamic spectrum access method and system based on multi-agent reinforcement learning. Background technique

[0002] In cognitive wireless networks, cognitive users use the overlay method to opportunistically access spectrum holes of authorized users for data transmission. Distributed multi-user dynamic spectrum access faces two major challenges: one is to avoid the interference of cognitive users to the primary user, that is, when the primary user occupies the licensed spectrum for data transmission, the cognitive user cannot access the corresponding spectrum; the other is Avoid access conflicts between cognitive users, that is, prevent more than two cognitive users from accessing the same spectrum hole, resulting in unsuccessful data transmission. Due to the limited perception ability of a single cognitive node, only part of the channel st...

Examples

Embodiment Construction

[0023] The following will clearly and completely describe the technical solutions in the embodiments of the present invention with reference to the accompanying drawings in the embodiments of the present invention. Obviously, the described embodiments are only part of the embodiments of the present invention, not all of them. Based on the embodiments of the present invention, all other embodiments obtained by persons of ordinary skill in the art without creative efforts fall within the protection scope of the present invention.

[0024] In addition, the technical solutions of the various embodiments of the present invention can be combined with each other, but it must be based on the realization of those skilled in the art. When the combination of technical solutions is contradictory or cannot be realized, it should be considered as a combination of technical solutions. Does not exist, nor is it within the scope of protection required by the present invention.

[0025] Unless ...