DBScholar

Back to papers

DepCache: A KV Cache Management Framework for GraphRAG with Dependency Attention

Summary: Dependency attention: graph-aware attention that prunes token-pair interactions to structural dependencies and reuses computations along relational paths to reduce inference cost. DepCache: KV-cache reuse aligned across graph-augmented prompts with a locality-aware replacement policy, yielding 1.5–5× throughput and up to 3.2× time-to-first-token reduction without accuracy loss. (summarized by gpt-5-mini on Feb 11 2026)

Paper ID
7564
Venue
SIGMOD
Year
2026
Pagerank
5.093636e-05
Overall Rank
10,357 | 28.95%
DOI
10.1145/3769778

Incoming Non-self Citations Over Time

No non-self incoming citations found for this paper in this database.

Authors

BibTeX Citation

@inproceedings{yuan_sigmod26,
        title = {{DepCache: A KV Cache Management Framework for GraphRAG with Dependency Attention}},
        author = {Yuan, Hao and Ai, Xin and Wang, Qiange and Li, Peizheng and Yu, Jiayang and Chen, Chaoyi and Yang, Xinbo and Zhang, Yanfeng and Fu, Zhenbo and Wen, Yingyou and Yu, Ge},
        series = {{SIGMOD} '26},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3769778},
        url = {https://dl.acm.org/doi/10.1145/3769778},
        year = {2026}
}

Incoming Citations (Sorted by Pagerank)

Showing 0 of 0 citing papers.

Rank Citing Paper Year Venue Pagerank
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 9 of 9 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers