Skip to content

Author

Tong Xiao

We have 6 of 23 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#natural language process... Preprint Sep 2026

Complementary Roles of Activation and Parametric Memory in Few-Shot Learning

At test time, large language models (LLMs) can encode historical information in activation memory (i.e., KV caches) and parametric memory (i.e., updated parameters). While activation memory is generally considered effective for factual recall and parametric memory for learning new tasks, their interplay remains unclear...

Miao-He Niu, Run-Song Zhao, Xin-Yu Liu et al. · 0 citations
#small language model Preprint Sep 2026

Reading Right, Answering Wrong: How Visual Configuration Changes Affect Evidence Use in VLMs

Findings show that configuration changes can affect how models use information they can still read, and attention interventions in LLaVA-NeXT suggest that configuration changes can weaken the use of readable information during answering.

Ding-Yang Lin, Ying-Feng Luo, Cheng-Long Wang et al. · 0 citations

BabelArena: A Large-Scale Multilingual Benchmark for LLM Agents

BabelFlow is introduced, a benchmark-general agentic workflow that adapts existing agent benchmarks to new languages by analyzing runtime dependencies, coordinating structure-preserving translation, and combining multi-layer verification with human review to preserve task and evaluation semantics.

Peng Kuang, Yu-Chun Fan, Jiang-Nan Li et al. · 0 citations
Preprint Aug 2026

Learning from Environmental Feedback: Credit Assignment across Multiple Timescales for Agentic Reinforcement Learning

Environmental Feedback-based Credit Assignment (EFCA), a multi-timescale credit assignment approach for long-horizon agentic RL that complements the long-term outcome signal with two environment-grounded process signals: a short-term feedback signal that captures the immediate effect of the current action and a medium-...

Yifu Huo, Shunjie Xing, Chenglong Wang et al. · 1 citation
Jul 2026

FlowCTS: On-policy Continuous Trajectory Supervision of Flow Models

Flow Continuous Trajectory Supervision (FlowCTS), which matches subsequent student and reference trajectories initialized from the same student-visited state to derive a temporally weighted velocity-matching upper bound and discretize it into practical objectives parameterized by the number of supervision steps.

Kaiyang Ye, Yuan Ge, Junxia Zhang et al. · 0 citations
Preprint Jul 2026

ToFu: A White-Box, Token-Efficient Agent Harness for Researchers

ToFu is presented, an agentic harness for researchers that reads your codebase, edits files, runs commands, and integrates with your development tools and provides a white-box agentic harness that allows researchers to inspect, modify, and evaluate its orchestration logic, tool-use behavior, and harness design.

Junhao Ruan, Yuan Ge, Bei Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.