Skip to yearly menu bar Skip to main content


Spotlight

Long-Context Modeling with Dynamic Hierarchical Sparse Attention for Memory-Constrained LLM Inference

Siheng Xiong ⋅ Joe Zou ⋅ Faramarz Fekri ⋅ Yae Jee Cho

Abstract

Lay Summary

Log in and register to view live content