Skip to content

Category

artificial intelligence

14,158 papers

#artificial intelligence Preprint Open access Oct 2026

Safe Actions Alone Do Not Ensure Safe Agents: Identifying Unfulfilled Obligations with Guard Models

Guard models are increasingly used to safeguard LLM-based agents, primarily by identifying actions that agents are forbidden to perform. However, identifying forbidden actions alone is insufficient to ensure agent safety. In this paper, we argue that agent safety also depends on identifying required yet unperformed saf...

Youwei Feng, Yitong Zhang, Yuetong Liu et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

What Output-Only Review Cannot Verify: Study Contracts for Research Agents

Some defects in an AI-generated study can be identified from its artifacts; others require knowledge of what was approved before execution. We propose study contracts that bind declared experimental choices, run obligations and claim scope to recorded execution evidence, and distinguish this contract-relative verificat...

Eitan Waks, Ben Glocker · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Where Draft Trees Lose Target Mass: Exit-Guided Speculative Decoding

Tree-based speculative decoding verifies multiple draft continuations in one target-model pass, but finite trees built from draft scores face a fundamental draft-target mismatch. We ask whether better exact verification can increase acceptance on a fixed tree and how target feedback can improve the tree itself. Through...

Shijing Hu, Xuancheng Ren, Zhihui Lu et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Scalable AI Uncertainty Quantification via Generalized Laplace Active Subspaces

Reliable uncertainty quantification (UQ) is essential for deploying neural networks in scientific and high-stakes applications, but full Bayesian inference over the network parameters is computationally infeasible. We propose a low-rank generalized Laplace approximation for neural-network UQ based on a small number of...

Wouter N. Edeling, Peter V. Coveney · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Onboard Marine Anomaly Detection on $\Phi$sat-2: From Simulation-Based Development to In-Orbit Demonstration

Onboard Artificial Intelligence can improve responsiveness and bandwidth efficiency of Earth Observation systems by processing data directly on the satellite. This paper presents the experience gained from the development, onboard integration, and post-launch adaptation of a lightweight marine anomaly detection pipelin...

Clotilde Szywala, Thomas Goudemant, Marjorie Bellizzi et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

MemTrial: Learning When to Trust Memory in LLM Portfolio Agents

Large language model (LLM) agents for portfolio management learn from experience: they credit each experience in their memory with the outcome of the decisions that used it. In financial markets, however, this outcome mostly reflects the market move shared by all decisions on that date, so the credit tracks the market...

Guanghao Wu, Zhuo Cai, Shoujin Wang · 0 citations
#artificial intelligence Preprint Open access Oct 2026

MultiWorldBench: Do Independently Controlled Views Describe One Shared World?

Multiplayer world models must ensure that independently controlled views remain consistent with one shared and persistent world. We introduce MultiWorldBench, a diagnostic Minecraft benchmark containing 495 case configurations across seven task suites and ten capabilities, including independent control, cross-view moti...

Zhangbo Xu, Ruoxi Zhang, Rui Hu et al. · 0 citations
#artificial intelligence Preprint Oct 2026

Internalizer: Portable Context-to-Parameter Mapping for Very Large Language Models

Hypernetworks that map a context directly to a LoRA adapter let a large language model carry that context in its weights, but prior work has demonstrated them only on base models of up to 14 billion parameters. We present the Internalizer, a state-of-the-art, portable Context-to-Parameter Mapping hypernetwork that gene...

Peter Devine, Nick Ryan, Benjamin Sirb et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

DeltaReplay: Task-Relative Memory Reuse for Mobile GUI Agents

Memory-augmented mobile GUI agents store successful execution trajectories and reuse them in later tasks, but a stored trajectory rarely matches a new task exactly. The new task may use different parameters, share only some of its steps with a stored trajectory, or have no relevant record in memory. Forcing the agent t...

Yudong Bai, Yihong Chen, Quanming Yao et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Intervention anchors and scientific verification in synthetic vascular predictive representations

Complete orthogonal predictive coordinates do not by themselves bind a latent direction to a named intervention. We present a mathematical and synthetic audit motivated by vascular device-vessel suitcordance. Capacity-matched least-squares predictors were exactly equivalent under complete fixed output transforms, where...

Lingsen You, Yujun Guo, Xinyu Zhong et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

A 3D Characterization Framework for Intelligent Sequential Decision Making

Puzzles are widely used to evaluate the reasoning capabilities of artificial intelligence (AI) systems for sequential decision making, yet approaches originating from different paradigms are rarely compared under unified conditions. To address this gap, we introduce a three-dimensional characterization framework that e...

Sadig Gojayev, Carolina Fortuna · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Constrained Command-Conditioned Reinforcement Learning with Bandit Strategy Selection in Real-Time Strategy Games

Deep reinforcement learning agents reach strong performance in real-time strategy games but can be brittle against opponents outside their training distribution. Separating strategic command selection from learned unit control allows different strategies to be selected for different opponents while reusing the same exe...

Nick Leenders, Roy Lindelauf, Joost van Oijen et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Sep 29, 2026

Who we become when we talk to machines

Professor Sherry Turkle’s new book, “Artificial Intimacy,” offers a withering critique of chatbots and the antisocial dynamics she believes they encourage.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.