Abstract
Algorithms based on deep learning and the Bellman iteration serve as a basis for most state-of-the-art approaches in the field of reinforcement learning. Deep Q-learning stands out as the predominant example. In scenarios involving high-dimensional input data, like pixel observations, the architecture of a deep Q-Network typically features a sequence of convolutional layers succeeded by a set of linear layers. In this setting, the activations in the network's final hidden layer can be seen as the latent representation, encompassing all the compressed information. We show that this learned representation is prone to saturation or contraction, leading to vanishing gradients, a reduction of information content and sub-optimal convergence. In addition, the temporal evolution of the latent representation in RL is analyzed by characterizing its entropy. Finally, a set of methods is proposed to alter the latent representation during learning by influencing its entropy. Three entropy-enhancing techniques are compared, which show a strong empirical relation between representation entropy and downstream performance.
| Original language | English |
|---|---|
| Title of host publication | 2025 IEEE Symposium on Computational Intelligence in Image, Signal Processing and Synthetic Media (CISM) |
| Subtitle of host publication | [Proceedings] |
| Publisher | Institute of Electrical and Electronics Engineers Inc. |
| Pages | 1-7 |
| Number of pages | 7 |
| ISBN (Electronic) | 9798331508357 |
| ISBN (Print) | 9798331508364 |
| DOIs | |
| Publication status | Published - 2025 |
| Event | 2025 IEEE Symposium on Computational Intelligence in Image, Signal Processing and Synthetic Media, CISM 2025 - Trondheim, Norway Duration: 17 Mar 2025 → 20 Mar 2025 |
Conference
| Conference | 2025 IEEE Symposium on Computational Intelligence in Image, Signal Processing and Synthetic Media, CISM 2025 |
|---|---|
| Country/Territory | Norway |
| City | Trondheim |
| Period | 17/03/25 → 20/03/25 |
Bibliographical note
Publisher Copyright:© 2025 IEEE.
Keywords
- reinforcement learning
- representation learning
Fingerprint
Dive into the research topics of 'Analyzing Latent Entropy in Deep Q-Learning'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver