Skip to content

Recursive Reasoning or Statistical Extrapolation? In-Context Learning in Multi-Agent Interdependent Decision-Making

Sep 2026 · 0 citations · 27 references
Computer Science Economics

TL;DR

This work extends the mechanistic study of ICL to strategic multi-agent settings, introduces REE as a diagnostic tool for distinguishing reasoning from extrapolation, and provides a reusable framework for probing the boundaries of LLM reasoning in recursive belief tasks.

Abstract

In-context learning (ICL) enables large language model (LLM) agents to improve decisions using interaction history, yet it remains unclear whether such improvement reflects refined internal reasoning or mere extrapolation of statistical patterns. To disentangle these mechanisms, we study LLM agents in multi-agent incomplete-information games that require recursive belief reasoning. By constructing a public goods game and manipulating the statistical structure of historical feedback, we evaluate decision quality against a history-independent rational expectations equilibrium (REE) benchmark. Our experiments reveal that when historical statistical patterns are disrupted, the benefits of longer context largely vanish, degrading decision quality to the no-context baseline in a way sharply amplified by stronger strategic interdependence. These results suggest that, in such strategic environments, ICL behavior is more consistent with statistical extrapolation than with strategic reasoning. Our work extends the mechanistic study of ICL to strategic multi-agent settings, introduces REE as a diagnostic tool for distinguishing reasoning from extrapolation, and provides a reusable framework for probing the boundaries of LLM reasoning in recursive belief tasks.

View source

Similar papers

Preprint Aug 2026

Contextual Information Policy Optimization for Search Agents

Search agents extend large language models beyond static parametric memory by enabling them to acquire and use external evidence during multi-step reasoning. For knowledge-intensive tasks involving complex or evolving information, their reliability depends not only on retrieving relevant evidence but also on using it t...

Xingyu Guo, Wei Chen, Lin-Lin Yang et al. · 0 citations
#artificial intelligence Preprint Sep 2026

SAGE: Structured Strategic Reasoning for Efficient LLM Game Playing

A strong LLM strategic agent should reason prospectively over uncertain futures, adapt its strategy to opponents'behavioral tendencies, and continuously recalibrate its decision process from interaction experience. However, incorporating these sources in free-form reasoning could lead to unsupported strategic assumptio...

Zhi-Wei Chen, Tian-Chun Wang, Zhong-Tao Rao et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Agentic Multi-Turn Reasoning: A Fairness Approach

Recent advances in Large Language Models (LLMs) have enabled agentic systems capable of solving complex tasks through multi-turn planning, tool use, verification, and memory updates. However, learning agentic systems remains difficult due to two fundamental challenges, i.e., (1) long-horizon credit assignment, where su...

Thanh-Dat Truong, Sankalp Pandey, Hugh Churchill et al. · 0 citations
#artificial intelligence Review Aug 2026

What is Missing from AI Post-Training AI: An Empirical Analysis

Analyzing a large corpus of publicly released post-training trajectories, it is found that across different tasks, the agent's training strategy is locked in at the very beginning, and the entire remaining budget is spent on local adjustments within the selected strategy.

J. Lim, Xin Huang, Hao Peng et al. · 1 citation · ⚡1
#artificial intelligence Preprint Sep 2026

Thinking Before Thinking: Scaling Agentic Inference Through Meta-Reasoning

As agents take on longer and more complex problems, controlling the execution becomes a task in its own right. Each step in the run brings new control choices, like which partial work to build on, whether to start fresh, or when to stop. We introduce agentic meta-reasoning, an inference-time harness that makes these ch...

Paras Dahal, A. Bakhtin, Taco Cohen et al. · 0 citations
Conference Aug 2026

Subjective Decision-Making in Multi-Agent Reinforcement Learning with Cumulative Prospect Theory and Social Value Orientation

Cyber physical systems such as autonomous vehicles operate in highly dynamic environments where interactions with other autonomous and human agents is inevitable. Reinforcement learning (RL) is a well-established paradigm to allow agents to learn behaviors through interactions with the environment when a model of the e...

John Lewis, Josiah Mesler, Bhaskar Ramasubramanian · 0 citations

Related blog posts

MIT News · Artificial Intelligence Sep 29, 2026

Who we become when we talk to machines

Professor Sherry Turkle’s new book, “Artificial Intimacy,” offers a withering critique of chatbots and the antisocial dynamics she believes they encourage.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.