Skip to content

Category

natural language processing

6,613 papers

#artificial intelligence Preprint Oct 2026

Cluster Validation Indices as Self-Supervised Objectives for Text Representation Learning

Self-supervised fine-tuning refines the embedding space of a pretrained language encoder without labels. However, the commonly used approaches are computationally expensive. Specifically, contrastive learning-based methods need multiview data and in-batch negative examples, while negative-free approaches require auxili...

Kishor Kumar Bhaumik, Nícolas Roque dos Santos, Neil Shah et al. · 0 citations
#artificial intelligence Review Oct 2026

Viva La Vida: Verification and Accumulation Failures in Multi-Agent Proof Search

When an agentic prover works on an open problem, there is no proof assistant to fall back on: its verifier and lemma library are ultimately language models judging model outputs. We instrumented such a system end to end and analyzed $51{,}754$ traced observations across three full runs ($186$ hours, \$$5{,}694$). We fi...

Ben-Ji Xu, Ken Zheng, Noah Han · 0 citations
#artificial intelligence Preprint Oct 2026

More Value per Key: Asymmetric Sparse Attention for Faster LLM Decoding

Autoregressive generation in Large Language Models (LLMs) is constrained by the memory and computational demands of attention mechanisms. Sparse attention methods mitigate this cost by selecting only high-probability entries of the attention matrix. We observe that in many such methods, this renders the probability-val...

Noam Elata, Itay Lamprecht, Mikey Shechter et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Extracting Persona Subspaces Through Iterative Nullspace Projection For Modulation

Large Language Models (LLMs) can adopt distinct personas to tune their semantics, expertise, and perspective to different users and tasks. Precise control over these traits is critical to ensure safety and reliability in model behavior. Existing methods like activation steering and prompt-based persona induction reduce...

Ananya Malik, Mai ElSherief · 0 citations
#artificial intelligence Preprint Open access Oct 2026

From Probe Scores to Alarm Policies: Operational Validity of Activation Monitors for Language-Model Agents

Activation probes can predict safety-relevant properties of language models with high area under the receiver-operating-characteristic curve (AUROC), but deployed agent monitors make thresholded alarm decisions under tight false-alarm budgets. These are different estimands. We introduce an Operational Validity Contract...

Xueping Gao · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Understanding and Mitigating Hallucination Escape in Tool-Using LLM Agents

Large language models (LLMs) increasingly serve as autonomous agents that invoke external tools. However, this capability introduces tool hallucination, selecting incorrect tools or generating invalid calls. Existing mitigation methods report substantial improvements, yet we identify a previously overlooked failure mod...

Peigui Qi, Kunsheng Tang, Yide Song et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

XTurnix: Large-Scale Self-Supervised Turn Control through Two-State Binary Decisions

General turn-taking behavior in real-time dialogue systems requires deciding whether to keep listening or start responding while listening, and whether to continue or stop while speaking. Existing turn detectors use heterogeneous, task-specific label spaces and are often trained on limited annotations or evaluated on i...

Zhanxun Liu, Yifan Duan, Hengtao Wu et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

GlitchPatch: Repairing Glitch Tokens in Frozen Language Models via Local Retokenization

Glitch tokens are anomalous vocabulary entries that can cause large language models (LLMs) to produce outputs inconsistent with their inputs. Existing repair methods require access to model internals, making them impractical for frozen checkpoints. We investigate whether glitch tokens can be repaired outside the model...

Kunsheng Tang, Peigui Qi, Yide Song et al. · 0 citations
#artificial intelligence Preprint Oct 2026

Questioning the Questions: Sustaining Self-Evolution in Reasoning Models

Self-evolving reasoning models learn from their own generated questions, yet repeated self-training can lead to performance collapse. In this paper, we investigate why performance deteriorates over successive rounds and how to sustain self-evolution. Our analysis identifies two recurring quality problems in self-genera...

Jin-Yuan Li, Chengsong Huang, Lang-Lin Huang et al. · 0 citations
#artificial intelligence Preprint Oct 2026

First-Order Steering: Translating Weight Adaptation into Activation Steering

Activation steering exploits interpretable directions in the residual stream to enable inference-time manipulation of model behavior. Composing steering vectors to apply multiple target behaviors simultaneously is important in various fields-including AI alignment and safety-but remains a challenge for existing activat...

Sri Pranav Kunda, Alexander Kurz, T. Dominik et al. · 0 citations
#artificial intelligence Preprint Oct 2026

Rethinking Self-Distillation for Multi-Teacher Capability Merging

Combining capabilities of multiple expert models trained starting from the same base checkpoint has become increasingly common in frontier language-model post-training. Recent trends suggest that multi-teacher on-policy distillation (MOPD) outperforms conventional off-policy methods. However, despite the higher inferen...

Roy Xie, Dan Friedman, Feng Nan et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

AI-Enabled Quality Assurance for Multiple-Choice Assessment Items

Generating multiple-choice questions is increasingly scalable, but establishing their assessment quality remains difficult. We present a focused narrative review of automated item-writing flaw detection, revision, psychometric screening, and NLP benchmark auditing. Database searches, citation retrieval, and nominated s...

Steven Moore, Nicholas Diana · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Sep 24, 2026

Estimating suicide risk from text

A new language-processing tool could help identify the highest-risk individuals from natural language, enabling swifter interventions.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.