Skip to content

Category

natural language processing

6,613 papers

#natural language process... Preprint Open access Oct 2026

Large Language Model Orchestration under Heterogeneous Preferences via Explicit Persona Inference

LLM orchestration investigates how an orchestrator coordinates a group of autonomous agents to achieve common goals or maximize collective welfare. The agents are typically heterogeneous, each holding a private preference that it pursues but does not reveal. Inferring such hidden preferences from behavior has been a su...

Shuqing Shi, Ziyan Wang, Milind Tambe et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

HouseholdBench: Evaluating Large Language Models as Predictors of Household Economic Behavior

Large language models (LLMs) have the potential to meet a key goal in economics: a quantitative model of household decision making, across a variety of settings. Yet existing evaluations cover few surveys and outcomes, and do not study how households adjust to changing economic conditions. We introduce a new evaluation...

Jin Huang, Diego Ferreras Garrucho, Yutong Xie et al. · 0 citations
#natural language process... Preprint Oct 2026

Quality-Aware Self-Correcting Speech Translation on an Edge Device

We present a fully offline speech-to-speech translation pipeline that runs on a Jetson Nano (4 GB) and corrects its own weak translations without retraining. A Whisper-tiny ASR feeds an Opus-MT translator; multilingual BERT cosine similarity acts as a Quality Estimation (QE) gate, triggering a secondary-pass correction...

Z. Farooq, Diptesh Kanojia · 0 citations
#natural language process... Preprint Open access Oct 2026

Not What a Child Expressed: Auditing the Sign-to-Text Safety Interface in Child-Facing AI

Automatic sign language translation (SLT) has entered consumer products, turning American Sign Language into English text for dictation, messaging, and queries put to a conversational assistant. Child-facing AI and platform trust-and-safety tooling decide on text, using filters on minor accounts and grooming classifier...

Muhammad Rafiullah Memon, Viet Vo, Wanlun Ma et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

Closing Ambient Clinical Documentation Gaps with Automated Provider Queries

Provider queries are clarifying requests sent by clinical documentation specialists to physicians to close gaps in the clinical note and ensure accurate billing. Prior work automates note drafting, ICD-10 coding, and order extraction assuming a complete transcript, leaving these gaps unaddressed. We study whether an LL...

Joseph Paul Cohen, Raj Shah, Han-Chin Shing et al. · 0 citations

Who Wrote It Is Not Enough: Detecting Who Contributed the Insight

As LLMs increasingly assist scientific writing and peer review, detecting who wrote the text is no longer sufficient: we need to determine who contributed the underlying insight. We introduce Insight Provenance, the task of identifying whether a review insight originates from a human, an LLM, or their hybrid contributi...

Zhuo-Yang Zou, Abolfazl Ansari, Jia-Xi Yang et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

Kurate: Scalable Scientific Quality Analysis

Scientific search systems can find papers that are relevant to a question, but they generally do not assess the quality of the evidence that those papers provide. We present Kurate, a system that uses large language models (LLMs) to assess the quality of published studies. Kurate uses both the paper and its related doc...

Matthew J. Vowels, Jamie Cummins · 0 citations
#artificial intelligence Preprint Open access Oct 2026

TIDE 2.0: an open, model-agnostic engine for keyed de-identification of clinical notes

Clinical notes capture most of what is documented about a patient's care, but they cannot be used for research until protected health information (PHI) is removed. De-identification is often treated as a detection problem. Detection alone is not sufficient: redaction strips clinical content along with identifiers, date...

Jose D. Posada, Somalee Datta, Priya Desai · 0 citations
#natural language process... Preprint Open access Oct 2026

Identifying Introspection From the Inside

Large language models make claims about themselves that are both consequential and increasingly difficult to verify from behavior alone. How can we distinguish plausible confabulations from genuine introspection? In this paper, we identify mechanistic signatures of faithful self-report in a controlled setting. Using lo...

David I. Atkinson, Dillon Plunkett, David Bau · 0 citations
#natural language process... Preprint Open access Oct 2026

JudgeMoE: Distributional Aggregation for LLM-as-a-Judge

When an LLM judge scores an output, its score distribution retains uncertainty and disagreement information that is lost after scalar compression. We introduce JudgeMoE, a lightweight aggregator that assigns example-specific weights to cached judge score distributions and fuses them before computing a final score. A pr...

Yiqi Liu, Joseph James, Yang Wang et al. · 0 citations
#natural language process... Preprint Oct 2026

Turnslide: Scalable Multi-Turn Data Synthesis by Walking a Finite-State Machine

Small language models are inexpensive to serve and can run on private infrastructure, but base models are often not good enough at multi-turn tool calling, and fine-tuning them needs per-API data that rarely exists. Existing synthesis methods are too expensive for high-scale fine-tuning, as they often require mock oper...

Aaron Fainman, Gabriela Kadlecová, Maciej Gryka et al. · 0 citations
#natural language process... Preprint Oct 2026

WavePrune: One period is often enough for RoPE

Rotary Position Embedding (RoPE) encodes token positions by rotating each two-dimensional channel of the query and key vectors at a channel-specific frequency, making the attention logits invariant to a common shift of positions. However, this rotation is periodic, and it leads to position aliasing where relative posit...

Guan-Cheng Du, Luo-Tian Huang, Shao-Wen Wang et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Sep 24, 2026

Estimating suicide risk from text

A new language-processing tool could help identify the highest-risk individuals from natural language, enabling swifter interventions.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.