Depthweave-Kv

Momentum

2 papers in the last four weeks, against 1 the four weeks before. 0.0% of all new papers.

Jul 6Week of Sep 21

Latest papers 18

All topics
CardsList
  1. TierKV: Long-Context On-Device LLMs via Predictive Multi-Tier KV Caching

    Sep 18, 2026Zhihao Shu, Md Musfiqur Rahman Sanim, Jie Hu +4Kv-Cache ManagementDepthweave-Kv

  2. AgentKV: Phase-Aware KV Eviction for Agentic LLMs

    Sep 14, 2026Taowen Tony Liu, Jeffrey T. H. Wong, Can Xiao +3Depthweave-KvKey-Value Cache Eviction

  3. CommitKV: Lifecycle-Aware KV Cache Compression via Commit Transitions for Multi-Turn Agents

    Aug 8, 2026Weizhong Huang, Jinchao Zhang, Xiawu ZhengKey-Value Cache CompressionDepthweave-Kv

  4. Chess_db: A framework for working with large chess game datasets

    Jul 23, 2026Nicos Angelopoulos, Jan WielemakerChessLarge Databases

  5. SMetric: Rethink LLM Scheduling for Serving Agents with Balanced Session-centric Scheduling

    Jul 9, 2026Jiahao Wang, Kaizhan Lin, Kaixi Zhang +7SchedulersLarge Language Model Serving

  6. DepthWeave-KV: Token-Adaptive Cross-Layer Residual Factorization for Long-Context KV Cache Compression

    Jul 7, 2026Anna Cordoba, Adam Puente Tercero, Nerea Angulo Hijo +4Key-Value Cache CompressionDepthweave-Kv

  7. FreqDepthKV: Frequency-Guided Depth Sharing for Robust KV Cache Compression in Long-Context LLM Inference

    Jul 7, 2026Anna Córdoba, Adam Puente Tercero, Nerea Angulo Hijo +4Key-Value Cache CompressionDepthweave-Kv

  8. KV-Control: Parameter-Efficient K/V Injection for Trajectory-Controlled Text-to-Motion

    Jun 4, 2026Tengjiao Sun, Pengcheng Fang, Xiaoyu Zhan +4Depthweave-Kv

  9. Linear Scaling Video VLMs for Long Video Understanding

    May 29, 2026Cristobal Eyzaguirre, Jiajun Wu, Juan Carlos NieblesLong-Video BenchmarksSpatio-Temporal Attention

  10. NestedKV: Nested Memory Routing for Long-Context KV Cache Compression

    May 26, 2026Hong Chen, Xiang Liu, Yubo Gao +5Key-Value Cache CompressionDepthweave-Kv

  11. KVBuffer: IO-aware Serving for Linear Attention

    May 18, 2026Longwei Zou, Lin ZhongKimi Delta AttentionDepthweave-Kv

  12. HeatKV: Head-tuned KV-cache Compression for Visual Autoregressive Modeling

    May 14, 2026Jonathan Cederlund, Axel Berg, William Isaksson +3Key-Value Cache CompressionVisual Autoregressive Models

  13. Make Your LVLM KV Cache More Lightweight

    May 1, 2026Xihao Chen, Yangyang Guo, Roger ZimmermannDepthweave-KvRecent Vision-Language Models

  14. DepthKV: Layer-Dependent KV Cache Pruning for Long-Context LLM Inference

    Apr 27, 2026Zahra Dehghanighobadi, Asja FischerDepthweave-KvKey-Value Cache Eviction

  15. SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference

    Apr 23, 2026Hongyao Liu, Liuqun Zhai, Junyi Wang +1LLM Inference OptimizationEdge Devices

  16. DASH-KV: Accelerating Long-Context LLM Inference via Asymmetric KV Cache Hashing

    Apr 21, 2026Jinyu Guo, Zhihan Zhang, Jiehui Xie +7Key-Value Cache CompressionDepthweave-Kv