Memory-Augmented RL & Recurrent Policies Reading List
Curated by Mouhssine Rifaki | Stanford Electrical Engineering | Last updated August 2026
Partial observability demands memory. This list covers recurrent policies, memory-augmented architectures, and transformers as memory.
Memory-Augmented RL & Recurrent Policies: 10 key papers
- Neural Turing Machines
Graves et al. arXiv 2014.
- Hybrid computing using a neural network with dynamic external memory
Graves et al. Nature 2016.
- Model-Free Episodic Control
Blundell et al. arXiv 2016.
- Neural Episodic Control
Pritzel et al. arXiv 2017.
- Been There, Done That: Meta-Learning with Episodic Recall
Ritter et al. arXiv 2018.
- Unsupervised Predictive Memory in a Goal-Directed Agent
Wayne et al. arXiv 2018.
- Relational recurrent neural networks
Santoro et al. arXiv 2018.
- Deep Attention Recurrent Q-Network
Sorokin et al. arXiv 2015.
- Memory-based control with recurrent neural networks
Heess et al. arXiv 2015.
- Working Memory Graphs
Loynd et al. arXiv 2019.
← Back to main page