Agent / long-term memory
HippoRAG
HippoRAG: Neurobiologically Inspired Long-Term Memory for Large Language Models
Superseded baseline#9 of 63 most-superseded · first seen May 23, 2024
Superseded — cited as a baseline and beaten by newer methods
4 papers critique it · 2 beat it on benchmarks
What papers say
Verbatim critique sentences, each from a paper that cites HippoRAG as a baseline.
Some recent approaches (e.g. RAPTOR sarthi2024raptor, GraphRAG GraphRAG, HippoRAG gutierrez2024hipporag, and MemTree memtree) recognize the importance of memory structurality, yet none simultaneously embodies the flexibility and dynamicity during memory structure development.
“Every index-based method (one that pre-builds a structured store such as a graph, summary notes, or multi-store cache) lags long context on at least one benchmark”
“HippoRAG's performance drops most on large-scale discourse understanding due to its lack of query-based contextualization”
“it relies on a single unified index for the entire events with fixed Top-k retrieval”
What to use instead
Recent methods in the same sub-problem, not yet superseded in the knowledge base — arXiv benchmark leaders, not vetted production recommendations.
- HingeMemHingeMem: Boundary Guided Long-Term Memory with Query Adaptive Retrieval for Scalable DialoguesApr 8, 2026
- Jan 13, 2026
- Nov 25, 2025
- Generative Semantic Workspace (GSW)Beyond Fact Retrieval: Episodic Memory for RAG with Generative Semantic WorkspacesNov 10, 2025
- Oct 7, 2025
- PREMemPre-Storage Reasoning for Episodic Memory: Shifting Inference Burden to Memory for Personalized DialogueSep 13, 2025