Skip to main navigation Skip to search Skip to main content

Greedy Reinforcement Learning for Anti-Jamming Wireless Communications

  • Pei Gen Ye
  • , Yuan Gen Wang
  • , Jin Li
  • , Liang Xiao
  • , Guopu Zhu
  • Guangzhou University
  • Xiamen University
  • Shenzhen Institute of Advanced Technology

Research output: Contribution to journalConference articlepeer-review

Abstract

In this article, we propose a(t, ?)-greedy reinforcement learning algorithm for anti-jamming wireless communications, which chooses previous action with probability t and applies ?-greedy with probability 1-t. The key idea of our algorithm is that the more valuable the previous action is, the higher probability of directly performing it at the current time slot without learning. For this purpose, the average utility of several previous actions is first calculated as a threshold for the valuable action judgment. Then, probability t is formulated as a Gaussian-like function with respect to the difference between the threshold and the utility of the previous action, which makes the wireless devices find the optimal action at a faster speed in the early stage, and eventually ensures the convergence. As a concrete example, the proposed algorithm is implemented in a wireless communication system against multiple jammers. Simulation results show that compared with ?-greedy, the (t, ?)-greedy obtains faster convergence rate and slightly higher signalto-interference-plus-noise ratio when being applied to Qlearning, deep Q-networks (DQN), double DQN (DDQN), and prioritized experience reply based DDQN (PDDQN). The source code is available at https://github.com/GZHUDVL/tau-epsilon-greedy-RL.

Original languageEnglish
Article number9322486
JournalProceedings - IEEE Global Communications Conference, GLOBECOM
DOIs
StatePublished - 2020
Externally publishedYes
Event2020 IEEE Global Communications Conference, GLOBECOM 2020 - Virtual, Taipei, Taiwan, Province of China
Duration: 7 Dec 202011 Dec 2020

Keywords

  • ?-greedy
  • Reinforcement learning
  • action repetition
  • jamming attacks
  • wireless communications

Fingerprint

Dive into the research topics of 'Greedy Reinforcement Learning for Anti-Jamming Wireless Communications'. Together they form a unique fingerprint.

Cite this