Skip to content

Category

artificial intelligence

14,156 papers

#artificial intelligence Preprint Oct 2026

MetaEncoder: Exploring the Limit of Bi-Encoders for Multimodal System One Decision Making with Natural Language Interface

System One models output constrained decisions and probability distributions rather than free-form text generation. While prevailing paradigms rely on structured schema objects to encode state, intent, and candidate choices, we revisit a fully natural language-based System One interface. In this framework, both the use...

Jian-Peng Cheng, Guang-Yu Sun, Aashu Singh et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

From Geometry to Generalization: Why Row Normalization Can Beat Adam and Muon

Different optimizers can fit the same training data while selecting classifiers with substantially different geometries, but whether this difference provably affects population performance remains unclear. We show that row-wise normalization can achieve strictly higher population accuracy than full-batch Adam, a proxy...

Jihwan Kim, Dogyoon Song, Chulhee Yun · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Characterizing Overconfident Failure in LLM-Based Code Generation

Large language models (LLMs) are increasingly used for automated code generation, but generated programs can appear syntactically plausible while still failing execution-based correctness checks. Existing validation methods, such as testing and program analysis, remain essential but are often incomplete, costly, or app...

Ravishka Rathnasuriya, Wei Yang · 0 citations
#artificial intelligence Preprint Open access Oct 2026

REMORY: Learning Residual Memory for Context Compaction

Long-horizon agents compact their history to continue within a finite context window, but a textual summary alone may not support every subsequent decision. We introduce REMORY, a neural memory network that supplements the summary with a bounded sequence of soft memory tokens. Given the history and summary, the network...

Hanchen Xia, Baoyou Chen, Yutang Ge et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

DynaTE: Accelerating Diffusion LLMs via Dynamic Token Execution

Diffusion-based LLMs (dLLMs) have recently emerged as a promising alternative to autoregressive (AR) LLMs by enabling bidirectional parallel refinement, alleviating the sequential decoding bottleneck of AR generation. However, their parallel iterative refinement mismatches AR accelerators optimized for sequential decod...

Minghan Jiang, Jiayi Wang, Shuaiting Li et al. · 0 citations
#artificial intelligence Review Oct 2026

How to post-train on a surrogate: Envelope sampling mitigates reward hacking

Large language models (LLMs) are commonly post-trained against LLM judges and other cheap surrogates because the true reward, such as human preference, is too expensive to query at scale. This practice often leads to reward hacking, where reinforcement learning against a miscalibrated surrogate leads to undesirable sid...

Sanjit Dandapanthula, Shuvom Sadhuka, Samir Khan et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Gated Memory: Admission-Controlled Memory Formation for Conversational AI

Personalized conversational AI relies on long-term memory systems that extract facts from user utterances and store them in persistent vector stores. Despite progress in retrieval, deduplication, and lifecycle management, the formation stage, the moment a fact is first written to storage has received almost no principl...

Preeti Saraswat, Divya Neelagiri, Ajay Manoj · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Neuro-Memory Fuzzy Inference System for Mimicking Human-like Car Following Behavior

This study presents the Neuro-Memory Fuzzy Inference System (NeMeFIS), a hierarchical machine learning architecture that asymmetrically models acceleration and deceleration in car following behavior by integrating five human memory types procedural, working, episodic, semantic, and declarative. By linking external vari...

Nazmul Haque, Md Asif Raihan. Md. Hadiuzzaman · 0 citations
#artificial intelligence Preprint Open access Oct 2026

SafeInferCom: Safe Inference-Time Compute via Verifier-Guided Mid-Generation Intervention for Robotic Task Planning

Large Reasoning Language Models (LRLMs) enable multi-step reasoning for robotic task planning, but continued reasoning can overwrite valid intermediate plans or leave constraint violations unresolved, reducing planning reliability and wasting inference-time computation. We develop an inference-time monitor that exposes...

Weizhe Xu, Jialiang Fan, Mengyu Liu et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

RAG-Stress: Probing the Limits of Evidence Reliance in Retrieval-Augmented Generation

Following retrieved evidence does not guarantee factual correctness: misleading evidence can induce a model to replace an answer it previously gave correctly. Standard accuracy measures obscure this behavior by combining answer replacement with preexisting errors. We introduce RAG-Stress, a controlled diagnostic protoc...

Shunyuan Zhou, Hao Chen, Tianyu Wang et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Higher-Order Action Supervision Makes A Strong Policy Class

Modern data-driven decision-making methods, such as imitation learning (IL) and reinforcement learning (RL), have achieved great success in solving many complex tasks. However, these methods often suffer from serious control instability and robustness issues when applied in real-world applications such as robotics and...

Peng Cheng, Yunxian Hou, Zhi Zhou et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

VAMR: Multi-Question Agentic Reasoning for Efficient Long-Form Video Understanding

Long-form video understanding often involves multiple questions about different aspects of the same recording. Yet existing video agents typically process each question through an isolated tool-use trajectory. This repeatedly restarts video exploration and memory construction, missing opportunities to acquire evidence...

Runquan Gui, Hanzhu Chen, Zehao Wang et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Sep 29, 2026

Who we become when we talk to machines

Professor Sherry Turkle’s new book, “Artificial Intimacy,” offers a withering critique of chatbots and the antisocial dynamics she believes they encourage.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.