Skip to content

Category

natural language processing

6,429 papers

#natural language process... Preprint Open access Oct 2026

Learning to Act with Task Progress: Distilling Small Agents from Compact Teacher Supervision

Learning from large-model demonstrations offers a way to train small agents that can complete recurring tasks without calling a large model at every step. A central design choice is what to retain from teacher trajectories that contain reasoning, actions, and information about task progress. We introduce Task-Progress...

Wenxi Gan · 0 citations
#natural language process... Preprint Open access Oct 2026

Nobody Truly Agrees on Sentiment: Humans, Bespoke Tools, and LLMs Struggle with Social Media Texts

Social media is a rich source of real-time public sentiment, but widely used sentiment analysis tools are often applied without understanding their limitations. In this study, we evaluate the inter-rater reliability of three bespoke sentiment analysis tools (TextBlob, VADER, and Twitter-roBERTa-base) and three large la...

Himarsha R. Jayanetti, Sivakanesan Dhanushkanda, Shuai Hao et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

HySPE: Positional Encoding via Symplectic Dual Shears

We introduce Hyperbolic Symplectic Positional Encoding (HySPE), grounding positional attention in non-compact symplectic transformations. While canonical Rotary Position Embedding (RoPE) parameterizes the compact, elliptic branch of $\Sp(2,\R)$ via rotations, HySPE operationalizes its hyperbolic branch via a damped sym...

Zhongping Ji · 0 citations
#natural language process... Preprint Open access Oct 2026

InterView-C: A Synchronized Multimodal Corpus of VR Avatar-Mediated Survey Interviews

We present InterView-C, a German multimodal corpus of 27 survey interviews conducted entirely in virtual reality, with both interlocutors represented by avatars. The corpus aligns spoken interaction with synchronized behavioral data, including gaze, head and body movement, facial behavior, hand and finger tracking. Its...

Patrick Schrottenbacher, Leon Hammerla, Lydia Kleine et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

LLM4Impact: Integrating Heterogeneous Information for Scientific Impact Prediction

Predicting the future impact of a newly published paper is challenging because it must be inferred from heterogeneous evidence available at publication time. Existing approaches often rely on a single source of information or combine multiple sources without accounting for their different predictive roles. In this pape...

Yong Cao, Markus Flicke, Haoyu He et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

Mechanics of Long-Context Hybrid Models Part 1.1: From Hybrid Attention to Hybrid Position

The architectural design of Large Language Models (LLMs) is shifting from traditional full-attention-only models to hybrid models, which combine different attention modules to improve long-context efficiency and performance in length extrapolation and context extension. To explain why hybrid models work and how to desi...

Xiaoran Liu, Ziwei He, Xipeng Qiu · 0 citations
#natural language process... Preprint Open access Oct 2026

I would rather quit NLP than read another paper like this: The rise of antithesis in NLP papers

For better or worse, LLMs are by now used routinely for scientific writing.\footnote{This paper is no exception; we did use AI to assist with writing some of the sections (see Acknowledgments).} Many have noticed that recent models fill papers with unnecessary antithesis, stating over and over what the work does not do...

Olga Zamaraeva, Adri\'an Gude, Roi Santos-R\'ios et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

Cache the Encoder Within:Compact, Reusable Memory across LLM Queries

Repeated queries over shared documents incur redundant encoding, while caching model states introduces persistent storage costs. Building on CoMem's intermediate-state interface, EncBank treats a pretrained LLM's lower layers as a reusable document encoder and compactly stores their outputs for an adapted upper-layer r...

Hanzuo Liu, Chunyu Liu, Chaofan Lin et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

The Long Road to the Same Answer: Cognitive Bias Under Escalating Reasoning Budgets in Large Language Models

Reasoning models allocate extra computation at inference time and present their answers as the product of deliberate thought. If this deliberation works the way dual-process accounts of human cognition suggest, longer thinking should weaken the classic decision biases that fast, intuitive judgment produces. Using 30 vi...

Obada Kraishan · 0 citations
#natural language process... Preprint Open access Oct 2026

EASE: Entropy-Adaptive Distribution Shaping for Evading AI-generated Text Detectors

AI-generated text (AIGT) detection can be sensitive to the decoding choices of the source large language model (LLM). We observe that perturbing next-token logits or adjusting sampling temperature can reduce detection performance, providing a clear signal of detector vulnerability to decoding-time distribution changes....

Jicheng Zhou, Kahim Wong, Jialong Wang et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

Judging in Latent Space: Efficient Generative Reward Modeling via Semantics-Preserving Compression

Reward modeling often requires jointly representing and reasoning over multiple evaluation criteria, yet verbalizing this process token by token can incur substantial inference cost. Recent work on latent reasoning suggests that continuous states may support this computation more compactly. We introduce LatentGRM, a la...

Mingqing Yuan (Soochow University), Xiaobo Liang (Soochow University), Junwei Yang (University of Cambridge) et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

Shaer: Controlled Arabic Poetry Generation with Meter Subform and Semantic Conditioning

Classical Arabic poetry generation requires simultaneously satisfying semantic, linguistic, and fine-grained prosodic constraints. Existing systems typically control broad poetic attributes but do not jointly model semantic intent, meter subform, and poem length. We present Shaer, a controllable Classical Arabic poetry...

Ahmad Abbas, Tamara Fakih, Nour Fakih et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Sep 24, 2026

Estimating suicide risk from text

A new language-processing tool could help identify the highest-risk individuals from natural language, enabling swifter interventions.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.