Skip to content

Category

natural language processing

6,613 papers

#natural language process... Preprint Open access Oct 2026

Long-Horizon Textual World Modeling through Structured Reasoning

World models must predict how an environment evolves under sequences of actions, enabling agents to compare possible futures and reason about counterfactual actions before acting. Long-horizon prediction is commonly obtained by recursively applying a one-step transition model, but intermediate errors can compound over...

Fangxin Wang, Xiang Gao, Yuguang Yao et al. · 0 citations
#natural language process... Preprint Oct 2026

JEV versus LLMs: Accuracy, Cost and Calibration on Seven Political Science Replications

Large language models (LLMs) annotate and scale political text or constructs by generating text tokens. A new class of models, which TypeSafe markets as"System One"models, instead returns decisions and probability distributions across a user-supplied fixed answer set. A commercial model, JEV, is advertised as having a...

S. Denney, Matthew DiGiuseppe · 0 citations
#natural language process... Preprint Open access Oct 2026

Molecules of a Story: Community Detection in PMI-weighted Narrative Networks

Automatically extracted narrative networks -- graphs with entities as nodes and their relations as edges -- have proven useful for revealing central narrative structures through salient entities and their connections (Tangherlini et al. 2020; Labatut and Bost 2019). But a narrative is more than those central structures...

Kasper Fyhn, Rebekah Baglini · 0 citations
#natural language process... Preprint Open access Oct 2026

Before Agent Tells The Lie: Has Deception Already Been Represented?

Large language model (LLM)-based agents can exhibit deceptive behavior during task execution, including hiding failures, fabricating results, or falsely signaling task completion. Existing monitoring approaches mainly detect deception after it appears in observable actions or outputs. In this paper, we investigate whet...

Xinling Li, Dadi Guo, Qingyu Liu et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

Test-Time Adaptation of Reasoning Strategies with Bayesian Nonparametric Memory

While modern large language models (LLMs) have been trained to reason through verbalized chains-of-thought, the generation cost grows substantially due to suboptimal paths to reach the final answer. Furthermore, as new insights are discovered while observing various input queries (e.g. through self-reflection), limited...

Keshav Ramji, Tahira Naseem, Ram\'on Fernandez Astudillo · 0 citations
#natural language process... Preprint Oct 2026

AECP: Artifact-Exclusive Communication Protocol for Multi-Agent Code Generation

As AI agents increasingly tackle complex repository-level coding tasks, distributing work across multiple agents is a natural way to scale beyond the capabilities of a single agent. To coordinate their interdependent work, these agents share findings and agree on interfaces between modules. However, exchanged informati...

Jia-Qi Xue, Yan-Jun Wang, Xiangci Li et al. · 0 citations
#natural language process... Preprint Oct 2026

Behavior-Preserving KV Cache Compression

KV caches are a major bottleneck in long-context inference and long-form generation with large language models. Existing training-free eviction policies largely rely on proxy importance signals, such as attention mass, to decide which past tokens to retain. We argue that cache compression should instead preserve the pr...

Doo Hwan Hwang, Junyoung Jang, Jun-Ho Na et al. · 0 citations
#natural language process... Preprint Oct 2026

Do Speech Representations Preserve Regional Accent Across Read and Spontaneous Speech?

Regional accent cues can be captured under matched conditions, but it remains unclear whether they persist between read and spontaneous speech. We study RVG1, with 500 German speakers from nine regions, comparing ten speech representations on regional classification and continuous geolocation under matched conditions a...

P. A. Pérez-Toro, Tomás Arias-Vergara, Annette Schwarz et al. · 0 citations

Breaking Bureaucracy: Evaluating open-source LLMs for legal document review

In this paper, we evaluate open-source generative LLMs on legal Natural Language Inference (NLI). Legal inspectorial processes take place in specific domains and often deal with confidential data. This creates a need for working with local models that do not require labeled training data. We evaluate our models on the...

Farrukh Baratov, Niki van Stein, Suzan Verberne · 0 citations
#natural language process... Preprint Oct 2026

DeferKV: Rethinking Eviction Timing for One-Shot KV Cache Compression

Long-context large language models (LLMs) have demonstrated strong capabilities across a wide range of tasks, but the growing KV cache introduces substantial memory and inference overhead. Existing one-shot KV cache compression methods typically commit to irreversible eviction immediately after prefill, before any sign...

Zhe Wang, Jia-Kai Li, Yu-Jia Sun et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

Probabilistic Race and Ethnicity Prediction Using Group-Specific Name Lists

Statistically valid estimation of racial and ethnic disparities often requires inferring the probability that an individual belongs to a particular racial or ethnic group given only their name and geographic location. The standard approach, Bayesian Improved Surname Geocoding (BISG), relies on group population frequenc...

Kyla Chasalow, Noah Dasanaike, Kosuke Imai · 0 citations
#natural language process... Preprint Open access Oct 2026

Shared Stopping Decisions Change Answers in HQQ Cache Quantization

Language-model systems batch questions for throughput, but unrelated questions should not change a target's answer when its input and numerical execution are fixed. We study compression of the key and value cache, which stores attention representations reused during generation. With request-local groups, Transformers'...

Seunghui Jwa, Minsu Oh, Chanjun Park et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Sep 24, 2026

Estimating suicide risk from text

A new language-processing tool could help identify the highest-risk individuals from natural language, enabling swifter interventions.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.