A survey of Graph-based Agent Memory | A curated list of resources (surveys, papers, benchmarks, and opensource projects) on graph-based agent memory.
See the codeThis repository provides a comprehensive collection of research papers, benchmarks, and open-source projects on Graph-based Agent Memory. It includes contents from our survey paper 📖"Graph-based Agent Memory: Taxonomy, Techniques, and Applications" and will be continuously updated.
🤗 You are very welcome to contribute to this repository by launching an issue or a pull request. If you find any missing resources or come across interesting new research works, please don’t hesitate to open an issue or submit a PR!
📫 Contact us via emails: chang.yang@connect.polyu.hk, qinggang.zhang@polyu.edu.hk

Comparison between Traditional Agent Memory and Graph-based Agent Memory.




| Dataset | Scenario | Modality | Feature | Paper | Repo |
|---|---|---|---|---|---|
| LoCoMo | Interaction | Text+Image | Long conversational memory | [Paper] | [Website] |
| LongMemEval | Interaction | Text | Long-term interactive memory | [Paper] | [Github] |
| MemoryAgentBench | Interaction | Text | Multi-turn interactions | [Paper] | [Github] |
| MEMTRACK | Interaction | Text+Code+Logs | Long-term interactive memory | [Paper] | [Website] |
| MADial-Bench | Interaction | Text | Memory-augmented dialogue generation | [Paper] | [Github] |
| MemSim | Interaction | Text | Bayesian memory simulation | [Paper] | [Github] |
| ChMapData | Interaction | Text | Memory-aware proactive dialogue | [Paper] | [Github] |
| MSC | Interaction | Text | Multi-session chat | [Paper] | [Website] |
| MMRC | Interaction | Text+Image | Multi-modal real-world conversation | [Paper] | [Github] |
| MemBench | Interaction | Text | Interactive scenarios | [Paper] | [Github] |
| StoryBench | Interaction | Text | Interactive fiction memory | [Paper] | [Website] |
| DialSim | Interaction | Text | Multi-dialogue understanding | [Paper] | [Website] |
| RealMem | Interaction | Text | Project-oriented long-term memory interaction | [Paper] | [Github] |
| PersonaMem | Personalization | Text | Dynamic user profiling | [Paper] | [Github] |
| PerLTQA | Personalization | Text | Social personalized interactions | [Paper] | [Website] |
| MemoryBank | Personalization | Text | User memory updating | [Paper] | [Github] |
| MPR | Personalization | Text | User personalization | [Paper] | [Github] |
| PrefEval | Personalization | Text | Personal preferences | [Paper] | [Website] |
| LOCCO | Personalization | Text | Chronological conversations | [Paper] | [Github] |
| WebChoreArena | Web | Text+Image | Tedious web browsing | [Paper] | [Github] |
| MT-Mind2Web | Web | Text | Conversational web navigation | [Paper] | [Github] |
| WebShop | Web | Text+Image | E-commerce web interaction | [Paper] | [Github] |
| WebArena | Web | Text+Image | Web interaction | [Paper] | [Github] |
| MMInA | Web | Text+Image | Multihop web agent | [Paper] | [Website] |
| NQ | LongContext | Text | Natural question answering | [Paper] | [Website] |
| TriviaQA | LongContext | Text | Large-scale question answering | [Paper] | [Website] |
| PopQA | LongContext | Text | Adaptive retrieval augmentation | [Paper] | [Github] |
| HotpotQA | LongContext | Text | Explainable multi-hop QA | [Paper] | [Website] |
| 2wikimultihopQA | LongContext | Text | Multi-hop QA | [Paper] | [Github] |
| Musique | LongContext | Text | Multi-hop QA | [Paper] | [Github] |
| LongBench | LongContext | Text | Long-context understanding | [Paper] | [Github] |
| LongBench v2 | LongContext | Text | Long-context multitasks | [Paper] | [Github] |
| RULER | LongContext | Text | Long-context retrieval | [Paper] | [Github] |
| BABILong | LongContext | Text | Long-context reasoning | [Paper] | [Github] |
| MM-Needle | LongContext | Text+Image | Multimodal needle retrieval | [Paper] | [Website] |
| HaluMem | LongContext | Text | Memory hallucination eval | [Paper] | [Github] |
| MemoryBench | Continual | Text | Continual learning | [Paper] | [Github] |
| LifelongAgentBench | Continual | Text | Lifelong learning | [Paper] | [Website] |
| StreamBench | Continual | Text | Continuous online learning | [Paper] | [Website] |
| Evo-Memory | Continual | Text | Test-time learning | [Paper] | [Website] |
| Ego4D | Environments | Video+Audio | Egocentric episodic memory | [Paper] | [Website] |
| EgoLife | Environments | Video+Audio | Long-context life QA | [Paper] | [Website] |
| ALFWorld | Environments | Text | Household tasks | [Paper] | [Website] |
| BabyAI | Environments | Text | Language navigation | [Paper] | [Website] |
| ScienceWorld | Environments | Text | Multi-step science experiments | [Paper] | [Github] |
| AgentGym | Environments | Text | Multiple environments | [Paper] | [Website] |
| AgentBoard | Environments | Text | Multi-round interaction | [Paper] | [Github] |
| SWE-Bench | Tool/Gen | Text+Code | Code repair | [Paper] | [Website] |
| GAIA | Tool/Gen | Text | Deep research tasks | [Paper] | [Website] |
| xBench-DS | Tool/Gen | Text+Image | Deep-search evaluation | [Paper] | [Website] |
| ToolBench | Tool/Gen | Text→API | API tool use | [Paper] | [Github] |
| GenAI-Bench | Tool/Gen | Text+Image | Visual generation eval | [Paper] | [Website] |
@article{yang2026graph,
title={Graph-based Agent Memory: Taxonomy, Techniques, and Applications},
author={Yang, Chang and Zhou, Chuang and Xiao, Yilin and Dong, Su and Zhuang, Luyao and Zhang, Yujing and Wang, Zhu and Hong, Zijin and Yuan, Zheng and Xiang, Zhishang and others},
journal={arXiv preprint arXiv:2602.05665},
year={2026}
}
A survey of Graph-based Agent Memory | A curated list of resources (surveys, papers, benchmarks, and opensource projects) on graph-based agent memory.
See the codeThis repository provides a comprehensive collection of research papers, benchmarks, and open-source projects on Graph-based Agent Memory. It includes contents from our survey paper 📖"Graph-based Agent Memory: Taxonomy, Techniques, and Applications" and will be continuously updated.
🤗 You are very welcome to contribute to this repository by launching an issue or a pull request. If you find any missing resources or come across interesting new research works, please don’t hesitate to open an issue or submit a PR!
📫 Contact us via emails: chang.yang@connect.polyu.hk, qinggang.zhang@polyu.edu.hk

Comparison between Traditional Agent Memory and Graph-based Agent Memory.




| Dataset | Scenario | Modality | Feature | Paper | Repo |
|---|---|---|---|---|---|
| LoCoMo | Interaction | Text+Image | Long conversational memory | [Paper] | [Website] |
| LongMemEval | Interaction | Text | Long-term interactive memory | [Paper] | [Github] |
| MemoryAgentBench | Interaction | Text | Multi-turn interactions | [Paper] | [Github] |
| MEMTRACK | Interaction | Text+Code+Logs | Long-term interactive memory | [Paper] | [Website] |
| MADial-Bench | Interaction | Text | Memory-augmented dialogue generation | [Paper] | [Github] |
| MemSim | Interaction | Text | Bayesian memory simulation | [Paper] | [Github] |
| ChMapData | Interaction | Text | Memory-aware proactive dialogue | [Paper] | [Github] |
| MSC | Interaction | Text | Multi-session chat | [Paper] | [Website] |
| MMRC | Interaction | Text+Image | Multi-modal real-world conversation | [Paper] | [Github] |
| MemBench | Interaction | Text | Interactive scenarios | [Paper] | [Github] |
| StoryBench | Interaction | Text | Interactive fiction memory | [Paper] | [Website] |
| DialSim | Interaction | Text | Multi-dialogue understanding | [Paper] | [Website] |
| RealMem | Interaction | Text | Project-oriented long-term memory interaction | [Paper] | [Github] |
| PersonaMem | Personalization | Text | Dynamic user profiling | [Paper] | [Github] |
| PerLTQA | Personalization | Text | Social personalized interactions | [Paper] | [Website] |
| MemoryBank | Personalization | Text | User memory updating | [Paper] | [Github] |
| MPR | Personalization | Text | User personalization | [Paper] | [Github] |
| PrefEval | Personalization | Text | Personal preferences | [Paper] | [Website] |
| LOCCO | Personalization | Text | Chronological conversations | [Paper] | [Github] |
| WebChoreArena | Web | Text+Image | Tedious web browsing | [Paper] | [Github] |
| MT-Mind2Web | Web | Text | Conversational web navigation | [Paper] | [Github] |
| WebShop | Web | Text+Image | E-commerce web interaction | [Paper] | [Github] |
| WebArena | Web | Text+Image | Web interaction | [Paper] | [Github] |
| MMInA | Web | Text+Image | Multihop web agent | [Paper] | [Website] |
| NQ | LongContext | Text | Natural question answering | [Paper] | [Website] |
| TriviaQA | LongContext | Text | Large-scale question answering | [Paper] | [Website] |
| PopQA | LongContext | Text | Adaptive retrieval augmentation | [Paper] | [Github] |
| HotpotQA | LongContext | Text | Explainable multi-hop QA | [Paper] | [Website] |
| 2wikimultihopQA | LongContext | Text | Multi-hop QA | [Paper] | [Github] |
| Musique | LongContext | Text | Multi-hop QA | [Paper] | [Github] |
| LongBench | LongContext | Text | Long-context understanding | [Paper] | [Github] |
| LongBench v2 | LongContext | Text | Long-context multitasks | [Paper] | [Github] |
| RULER | LongContext | Text | Long-context retrieval | [Paper] | [Github] |
| BABILong | LongContext | Text | Long-context reasoning | [Paper] | [Github] |
| MM-Needle | LongContext | Text+Image | Multimodal needle retrieval | [Paper] | [Website] |
| HaluMem | LongContext | Text | Memory hallucination eval | [Paper] | [Github] |
| MemoryBench | Continual | Text | Continual learning | [Paper] | [Github] |
| LifelongAgentBench | Continual | Text | Lifelong learning | [Paper] | [Website] |
| StreamBench | Continual | Text | Continuous online learning | [Paper] | [Website] |
| Evo-Memory | Continual | Text | Test-time learning | [Paper] | [Website] |
| Ego4D | Environments | Video+Audio | Egocentric episodic memory | [Paper] | [Website] |
| EgoLife | Environments | Video+Audio | Long-context life QA | [Paper] | [Website] |
| ALFWorld | Environments | Text | Household tasks | [Paper] | [Website] |
| BabyAI | Environments | Text | Language navigation | [Paper] | [Website] |
| ScienceWorld | Environments | Text | Multi-step science experiments | [Paper] | [Github] |
| AgentGym | Environments | Text | Multiple environments | [Paper] | [Website] |
| AgentBoard | Environments | Text | Multi-round interaction | [Paper] | [Github] |
| SWE-Bench | Tool/Gen | Text+Code | Code repair | [Paper] | [Website] |
| GAIA | Tool/Gen | Text | Deep research tasks | [Paper] | [Website] |
| xBench-DS | Tool/Gen | Text+Image | Deep-search evaluation | [Paper] | [Website] |
| ToolBench | Tool/Gen | Text→API | API tool use | [Paper] | [Github] |
| GenAI-Bench | Tool/Gen | Text+Image | Visual generation eval | [Paper] | [Website] |
@article{yang2026graph,
title={Graph-based Agent Memory: Taxonomy, Techniques, and Applications},
author={Yang, Chang and Zhou, Chuang and Xiao, Yilin and Dong, Su and Zhuang, Luyao and Zhang, Yujing and Wang, Zhu and Hong, Zijin and Yuan, Zheng and Xiang, Zhishang and others},
journal={arXiv preprint arXiv:2602.05665},
year={2026}
}