The following publications are possibly variants of this publication:
- Edge-Assisted KV-Cache Collaborative Loading for Split LLM InferenceShiQi Zhang, Yunlong Xie, Lin Sun, Haibo Wu, Lili Zhang. icic 2027: 190-201 [doi]
- EdgeServe: efficient deep learning model caching at the edgeTian Guo 0001, Robert J. Walls, Samuel S. Ogden. edge 2019: 313-315 [doi]
- Energy-efficient edge-cloud collaborative intelligent computing: a dual-agent approachShoulu Hou, Mingyu Huo, Tao Wang, Xiulei Liu. computing, 108(6):94, June 2026. [doi]
- Socially trusted collaborative edge computing in ultra dense networksLixing Chen, Jie Xu. edge 2017: [doi]
- KV Cache Reuse for Elastic LLM Inference on Edge DevicesPeishuo Wang, Zhenzhe Zheng 0001, Xiaoyao Huang, Jie Wu 0001, Fan Wu 0006, Guihai Chen. icdcs 2026: 294-304 [doi]
- An efficient mobile-edge collaborative system for video photorealistic style transferAng Li, Chunpeng Wu, Yiran Chen, Bin Ni. edge 2019: 344-345 [doi]
- Kelle: Co-design KV Caching and eDRAM for Efficient LLM Serving in Edge ComputingTianhua Xia, Sai Qian Zhang. micro 2025: 18-33 [doi]