Awesome Agent Memory Papers

April 22, 2026 · View on GitHub

Stars Last updated Papers

A curated list of papers on memory for LLM / multimodal agents — methods, benchmarks, and surveys — covering episodic, semantic, procedural, and multimodal memory, with both parametric (internal) and retrieval-based (external) storage, learned via prompting, supervised finetuning, or reinforcement learning.

90 papers · 7 surveys · 31 benchmarks · 52 methods · last updated 2026-04-21

Interactive dashboard with multi-tag filtering: https://yyyujintang.github.io/Awesome-Agent-Memory-Papers/

Contributions welcome — open an issue or PR with new papers.

Contents

Surveys

Benchmarks

Evaluation suites for agent memory, split by interaction mode.

QA-based Memory Evaluation

Web Navigation

Desktop / Mobile GUI

Embodied & Game Environments

General Long-Horizon / Office

Methods

Each paper is placed in exactly one primary section (Multimodal > Procedural > Episodic > Semantic > External > Internal). Tag badges on each entry show the full tag vector — use the website for true multi-axis filtering.

Multimodal Memory

Procedural Memory

Episodic Memory

Semantic Memory

Internal / Parametric Memory

Other Methods

Tag Legend

AxisValues
CategorySurvey · Benchmark · Method
Benchmark TypeQA · Web · GUI · Embodied · Long-Horizon
StorageInternal (parametric — weights / latent tokens) · External (non-parametric — retrieval)
LearningPrompt-based · RL-based · SFT · Training-free
Memory TypeEpisodic · Semantic · Procedural · Multimodal

Citation

If this list is useful in your work, please consider starring the repo.