Skip to content

3 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

Human-Anchored Inference for Ranking New Models with Large Language Model Judges

Human pairwise comparisons provide a reference for evaluating large language models (LLMs), but collecting sufficient judgments for each new release is costly and time-consuming. LLM judges offer a scalable alternative, although their comparisons may differ systematically from human preferences and across judges. We st...

Xin Zhou, Si-Nian Zhang, Zhan-Yan Yang et al. · 0 citations
#machine learning Book Open access Jul 2026

GroupKV: Hierarchical KV Cache Management for Long-Context Diffusion LLM Inference

GroupKV is presented, a lightweight hierarchical KV cache management system for long-context dLLM inference that observes that under block-wise decoding, tokens within the same generation block tend to access highly overlapping and spatially concentrated context regions, making group-level sparse selection effective.

Jin-Hao Wang, Zhe-Xin Hu, Kang-Jie Zhou et al. · 2 citations · ⚡1
#large language models Book Open access Jul 2026

GroupKV: Hierarchical KV Cache Management for Long-Context Diffusion LLM Inference

Diffusion large language models (dLLMs) are emerging as a promising generative paradigm that complements autoregressive decoding. In long-context settings, KV cache bloat and offloading transfer overhead have become primary bottlenecks in inference systems. Meanwhile, the periodic full-sequence recomputation and locali...

Jin-Hao Wang, Zhe-Xin Hu, Kang-Jie Zhou et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.