Skip to content

Author

Yiren Zhao

We have 8 of 89 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Oct 2026

Securing Computer-Use Agents Against Branch Steering Attacks

Modern Computer Use Agents (CUAs) directly interact with graphical user interfaces and execute third-party web tools, exposing them to indirect prompt injection across every rendered page and tool response. While the Dual-LLM pattern is the primary system-level architecture offering formal security guarantees - using a...

Giulio Zingrillo, Hanna Foerster, Ilia Shumailov et al. · 0 citations
#artificial intelligence Preprint Oct 2026

iS-KV: Online Low-Rank KV Cache Compression via Block-Incremental SVD

Long chain-of-thought reasoning substantially increases KV-cache memory during autoregressive decoding, as every generated token introduces new key and value states and causes the cache to grow linearly with decoding length. Existing KV-cache compression methods typically control this growth through token eviction, but...

Yi-Ren Zhao, Guang-Hui Song, Tianrui Qin et al. · 0 citations
#machine learning Preprint Sep 2026

AgentKV: Phase-Aware KV Eviction for Agentic LLMs

This work proposes AGENTKV, which maintains a small query buffer per phase and scores cached keys against their union and implements AGENTKV in a persistent multi-turn serving path that carries compressed KV state across turns and compacts retained KV pages online.

T. Liu, Jeffrey T. H. Wong, Can Xiao et al. · 1 citation
#robotics Preprint Jul 2026

Source-Lifted Flow Matching for Intervenable Multimodal Imitation

Flow-matching policies are promising for imitation learning because they model complex multimodal action distributions. However, their stochasticity is largely passive: repeated sampling may yield diverse behaviors, but users cannot directly choose among valid continuations from the same state. We propose Source-Lifted...

He Zhang, Ying Sun, Pengteng Li et al. · 0 citations
Preprint Aug 2026

How Should Vision-Language-Action Models Use Proprioceptive State?

Five representative interfaces are implemented -- discrete state prompt, VLM prefix, action prefix, state expert, and feature modulation -- under matched implementation details, and evaluated on 45 atomic tasks spanning three task families plus 20 composite tasks.

Yiren Zhao, Ziyang Chen, Zi-Yang Rao et al. · 0 citations
Preprint Aug 2026

OasisKV: Scaling In-Decode KV Cache Beyond HBM with Lookahead Sparse Prefetching

OasisKV is presented, a memory-centric LLM inference system design that alleviates HBM capacity pressure by decoupling full KV-cache storage from HBM during LLM decoding and observes that future important tokens can be predicted accurately in advance using lookahead tokens drafted by speculative decoding (SD).

Can Xiao, Sukmin Cho, J. We et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.