Adaptive Scheduling: A Reinforcement Learning Whittle Index Approach for Wireless Sensor Networks
We propose a Reinforcement Learning (RL)-based scheduling framework for Restless Multi-Armed Bandit (RMAB) problems, centred on a Whittle Index Q-Learning policy with Upper Confidence Bound (Whittle index Q-Learning (WIQL)-upper confidence bound (UCB)) exploration. Unlike existing approaches that re...
Gardado en:
| Principais autores: | , , |
|---|---|
| Formato: | Artigo |
| Idioma: | Inglês |
| Publicado: |
IEEE
2026-01-01
|
| Series: | IEEE Access |
| Assuntos: | |
| Acceso en liña: | https://ieeexplore.ieee.org/document/11433425/ |
| Tags: |
Sen Etiquetas, Sexa o primeiro en etiquetar este rexistro!
|
