Skip to content

Category

artificial intelligence

14,192 papers

#artificial intelligence Preprint Open access Oct 2026

AI Safety Considerations for Agents With Limited Time to Act

In the wake of the increasingly public discussion about AI alignment, recent work has tried to propose specific AI architectures that behave safely. However, the proposed arguments that seemingly demonstrate proved alignment mostly neglect the environment the agent needs to act in. We discuss theoretical bounds for age...

Leo Zeitler, Jack Richings, Victoria Nockles · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Stale, Misattributed, or Late: Where Personal Memory Fails Before Generation

Personal memory for language agents is usually judged by whether the final an- swer is correct. That score hides errors that arise before generation: the memory block may contain an obsolete value, a fact about the wrong person, or no use- ful fact before the serving deadline. We measure these failures directly. Using...

Haonan Deng, Park Sinchaisri · 0 citations
#artificial intelligence Preprint Open access Oct 2026

OOM-RL II: Reality Is an Oracle, Not a Debugger Provenance-Constrained Diagnosis in Continually Evolving Agent-Engineered Systems

Reality may establish that an outcome occurred without identifying which evolving procedure produced it or why. This distinction matters in production ML systems whose code, configuration, and artifacts change while external feedback accumulates. We examine it in a human-directed, agent-engineered quantitative trading...

Kun Liu, Liqun Chen · 0 citations
#artificial intelligence Preprint Open access Oct 2026

GAGR-Lab: Evaluating Joint Spatial-Geometric and Analytic Function Reasoning

Joint spatial-geometric and analytic function reasoning requires translating a perceived spatial configuration into a symbolic function whose executed curve satisfies geometric constraints. We present GAGR-Lab, a framework for measuring this capability through Cartesian game scenes, explicit function semantics, and aut...

Jingyao Zhang, Yun Li, Lu Han · 0 citations
#artificial intelligence Preprint Oct 2026

Agentic AI-Assisted Modeling for Production Scheduling: Assessment in Constraint Programming

Developing optimization models for production scheduling requires substantial expert effort. Research on large language models (LLMs) has followed two directions: specialized approaches for automated modeling, mostly for mixed-integer linear programming, which often rely on dedicated training or problem-specific archit...

Ángel Sánchez-Fernández, J. Pernas-Álvarez, D. Crespo-Pereira · 0 citations
#artificial intelligence Preprint Oct 2026

UniSkill: Learning Actor-Aligned Skill Proposals for an Evolving Policy

Large language model agents can improve across tasks by retaining reusable skills distilled from prior interactions. Recent work jointly optimizes task execution and skill extraction, enabling the policy and skillbank to co-evolve. However, as the actor continues learning, rewarding skill proposals through their reuse...

Yi-Fei Lu, Cheng Liu, Dian-Zhi Yu et al. · 0 citations
#artificial intelligence Preprint Oct 2026

RewardWeaver: Long-Horizon Interactive Learning for Language Agents via Self-Evolving Reward Adaptation

Reinforcement learning with verifiable rewards (RLVR) has driven substantial progress in domains where task outcomes can be reliably evaluated, but long-horizon interaction remains challenging due to sparse terminal feedback and difficult credit assignment. Process rewards provide denser supervision, yet the capabiliti...

Heng-Bo Xiao, Bo-Yao Zhang, Pu-Rui Liu et al. · 0 citations
#artificial intelligence Preprint Oct 2026

SkillSandbox: Skill Verification via Dynamic Scenario Synthesis

Self-evolving agents distill task-solving experience into skills for future reuse, but these skills can encode incorrect procedures or non-transferable knowledge. It is therefore critical to verify each skill's reusability: whether its guidance remains useful beyond the experience from which it was distilled. Such veri...

Serin Kim, Kwangwook Seo, Dok-Yung Song et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

HGP:An on-device personalized agent memory via hybrid graph storage

LLM-based agents face challenges in personalized interactive tasks due to heterogeneous, multi-typed, and implicitly constrained long-term traces. Existing memory mechanisms struggle with accurate routing and retrieval, especially on-device where personalization is critical. Most methods use single-vector representatio...

Ran Zhou, Xueming Han, Jiaheng Liu et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Loud Failures, Quiet Failures: Fault Detection and Recovery in Tool-Using Language Model Agents

Tool-using agents are usually scored on whether they finish a task while the tools work. Deployments are less forgiving: services time out, endpoints disappear, parameter names change, and results come back well formed but wrong. Prior work has shown that language models over-trust tool outputs that fail silently; we a...

Obada Kraishan · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Learning to Accumulate Knowledge with Mutual Information

Large language model (LLM) agents can improve their performance by reusing knowledge distilled from past interactions. However, curating new experiences into a knowledge bank that becomes more useful as it grows remains challenging. Effective knowledge accumulation should limit redundant overlap among entries and ensur...

Yuyang Zhao, Lizi Liao, Leyang Shen et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

What the Sleeve Feels: Explainable Machine Learning for Textile Pressure-Based Postural Screening

Pressure-sensing smart textiles convert body-surface contact into a dense, image-like signal closely tied to posture and movement, making them a promising low-cost route to wearable posture screening. Realizing that promise, however, requires more than classification accuracy: a deployable system must generalize to wea...

Limon Bin Hossain, Md Sadib Rahman Ananta · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Sep 29, 2026

Who we become when we talk to machines

Professor Sherry Turkle’s new book, “Artificial Intimacy,” offers a withering critique of chatbots and the antisocial dynamics she believes they encourage.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.