Skip to content

Category

artificial intelligence

14,190 papers

#artificial intelligence Preprint Open access Oct 2026

Training Advisors for LLM Agents from Task Outcomes

Large language model agents tackle multi-step tasks by interleaving reasoning and tool calls with observations from the environment. Prior work has shown that natural-language feedback can help these agents revise their decisions during task execution. We introduce Caddie, a method for training critics to provide natur...

Sergei Polezhaev, Barys Liskavets, Ori Press et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Stream-Based Active Learning with Cooperative Neural Networks for Data-Efficient Partial Inverse Design: An Automotive Glass Run Channel Case Study

Inverse design in engineering often runs into a simple problem. Each labeled training sample must be produced through expensive simulation, so building a large dataset is slow and costly. This study addresses that problem for partial inverse design, where only some design variables are specified and the rest must be in...

Agung Nugraha, Hyerin Kwon, Heungjun Im et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Fully Interpretable Minimal Transformers: From Geometry to Algorithm

We present a framework for building and interpreting minimal transformer models. By constraining a transformer's embedding dimension and head size to 2, we enable full two-dimensional visualization of its internal representations. Embeddings, query/key/value transforms, attention outputs, residual streams, and decision...

Raneem Mahajne, Toviah Moldwin · 0 citations
#artificial intelligence Preprint Open access Oct 2026

A Deafening Silence: Catastrophic Forgetting Lives in the Output Embeddings of Tokens the Data Never Speaks

Continual pre-training and fine-tuning in Large Language Models (LLMs) inevitably induce catastrophic forgetting, typically mitigated by replay using often-inaccessible original data. In this data-free regime, we analyze where forgetting occurs and why. Systematic parameter freezing across five settings up to 1.4B reve...

Jonghyun Han, Younghoon Song, Jongyoul Park · 0 citations
#artificial intelligence Preprint Open access Oct 2026

UltraText Bench: A Comprehensive Bilingual Benchmark for Evaluating Visual Text Rendering in Image Generation

Dense visual text requires image generators to reproduce long strings across multiple regions with correct placement and legibility. As short-string rendering improves, evaluation must test sustained performance across more demanding scenes. We introduce UltraText Bench, a bilingual benchmark for prompt-only generation...

Deyuan Liu, Yihao Hu, Jingxuan Zhang et al. · 0 citations
#artificial intelligence Preprint Oct 2026

DisParQ: Self-Supervised Part Concepts for Interpretable Vision Foundation Models

Concept-based vision models represent images through an intermediate layer of human-inspectable concepts, so what a model relies on can be traced to those concepts. However, those models are often limited to fixed categories or depend on language to define their concepts. We introduce DisParQ (Discrete Parts with Quant...

Adam Pardyl, Siddhartha Gairola, Sukrut Rao et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

MIRROR: From Imitation to Internalization in LLM Personalization

The demand for personalized LLMs is shifting from style imitation toward content quality. We investigate whether self-distillation can bridge this gap in existing fine-tuning paradigm. To address this limitation, we introduce MIRROR(Meta- personalization by Internalizing Reference-Revealed On-policy Reflections), a nov...

Huayi Lai, Jicheng Yang, Min Yi et al. · 0 citations
#artificial intelligence Preprint Oct 2026

Decoupling Logic from Persona: Structural Immunity of Edge LLM Agents to Context Pollution

Small language-model agents on edge devices must hold a persona and reason correctly at once, inside one context window that fills with conversational history and persona instructions. We study what happens to the logical part of such an agent when that history is long, misleading and persona-heavy (persona-logic inter...

Masaaki Nakatsu, Ren-Xiong Wang · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Artificial intelligence pathways from weather to climate

Deep learning has made rapid advances in weather forecasting: autoregressive models trained on atmospheric reanalyses now rival dynamical models across nowcasting, medium-range, and subseasonal-to-seasonal lead times, producing well-calibrated ensemble forecasts at reduced cost. We review these advances and consider th...

Tom Beucler, J. David Neelin, Hui Su et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Beyond Policy Support: Interaction Constrained Offline Reinforcement Learning for Autonomous Driving

Offline reinforcement learning enables reward-driven policy improvement from fixed datasets without requiring online exploration, making it particularly attractive in safety-critical domains. A central challenge, however, is distribution shift: policy optimization may favor actions that are weakly supported by the offl...

Mahmoud Selim, Cristina Cipriani, Karl Henrik Johansson · 0 citations
#artificial intelligence Preprint Open access Oct 2026

PARC-Loc: Text-to-Point-Cloud Localization with Partial Assignment and Relational Consistency

Text-to-point-cloud localization estimates a position in a city-scale 3D map from descriptions of surrounding objects. Existing coarse-to-fine methods retrieve submaps using aggregate learned compatibility and then localize within a selected submap. However, repetitive or similar urban objects can inflate the embedding...

Shengkai Ma, Zhenyu Hou, Weihua Cao · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Unrolled Flow Models for Reasoning

Flow matching enables language generation in few steps, but whether additional integration steps improve reasoning remains unclear. We prove that a flow parameterized by a two-layer Transformer can solve graph reachability, with the required number of integration steps increasing with the target's distance from the roo...

Faissal Izermine, Hanru Bai, Oscar Davis et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Sep 29, 2026

Who we become when we talk to machines

Professor Sherry Turkle’s new book, “Artificial Intimacy,” offers a withering critique of chatbots and the antisocial dynamics she believes they encourage.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.