Skip to content

Author

Ashkan Shahbazi

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Oct 2026

OVAL: Output-Aware Local Page Bases for KV Cache Retrieval

Long context inference with large language models becomes increasingly expensive as attention must operate over an ever growing KV cache. Page sparse attention reduces this cost by representing each KV page compactly and retrieving only a subset for each query. Existing retrieval methods are designed to estimate attent...

Ashkan Shahbazi, Chayne Thrash, Soheil Kolouri · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.