Skip to content

Category

natural language processing

6,613 papers

#natural language process... Preprint Open access Oct 2026

INMS: Memory Sharing for Large Language Model based Agents

While Large Language Model (LLM) based agents excel at complex tasks, their performance in open-ended scenarios is often constrained by isolated operation and reliance on static databases, missing the dynamic knowledge exchange of human dialogue. To bridge this gap, we propose the INteractive Memory Sharing (INMS) fram...

Hang Gao, Yongfeng Zhang · 0 citations
#natural language process... Preprint Open access Oct 2026

COMPASS 2.0: psychometric representational similarity analysis distinguishes symptom structure from personal signal

Language models can score psychiatric questionnaires from speech, but agreement with self-report may reflect the questionnaire rather than the person. We introduce psychometric representational similarity analysis, a framework for comparing the structure of speech-derived scores, self-report, item wording and theory, a...

Baihan Lin · 0 citations
#natural language process... Preprint Open access Oct 2026

Synthetic Cultural Agents from Aggregate Anchors

Population prompts are widely used to generate synthetic survey responses, but they combine information supplied at inference with associations already encoded during pretraining. We introduce an alternative construction that maps declared aggregate preference anchors into group-indexed choice policies. For each popula...

Augusto Gonzalez-Bonorino (Department of Economics, Arizona State University, EconLLM Lab) et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

DP-ES: Differentially Private Evolution Strategies for Prompt Optimization

Token-level differentially private (DP) prompt optimization methods such as DP-OPT can become unstable under tight privacy budgets: on GSM8K, DP-OPT obtains $49.5\pm28.5\%$ across 30 runs, and a logged search trajectory reveals prompt-template drift and noise-sensitive irreversible choices. We diagnose these as structu...

Ziniu Liu, Aiping Li, Yue Han et al. · 0 citations
#computer vision Preprint Oct 2026

Efficient Test-time Adaptation through Candidate Verification and Divergence Shifts

Vision-language models (VLMs) achieve strong zero-shot transferability but remain vulnerable to target-domain shifts at inference time. Test-time adaptation (TTA) offers a practical remedy, yet most existing VLM-TTA methods follow a prediction-side adaptation paradigm. They use test samples to adjust logits, prototypes...

Seungmin Oh, Seung-Hun Kang, Jongbin Ryu · 0 citations
#natural language process... Preprint Open access Oct 2026

VHDL-REPOBENCH: A Repository-Level Benchmark for Evaluating Large Language Models on VHDL Design Generation

Large Language Models (LLMs) are increasingly applied in hardware design automation, demonstrating strong potential in generating and understanding hardware description languages. However, most existing benchmarks focus on Verilog, with limited evaluation of VHDL, which remains widely used in industry and academia for...

Prashanth Vijayaraghavan, Akul Malhotra, Ashutosh Jadhav et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

templar: agentic induction and evolution of standardized radiology reporting templates from large-scale clinical corpora

Structured radiology reporting mitigates the heterogeneity of free-text reports, yet its benefits depend on high-quality reporting templates. In practice, such templates are conventionally built through labor-intensive expert consensus and therefore vary across institutions and lag behind evolving clinical practice. La...

Xiaotian Hu, Mingxuan Liu, Zhonghan Wang et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

SearchJev: A Fast and Calibrated System-1 Model for Search Agents

Search agents repeatedly make short decisions about relevance, evidence sufficiency, and search actions. Using generative language models for these decisions introduces latency and unreliable confidence. We present SearchJev, a fast and calibrated System-1 model that separates search decisions from System-2 reasoning a...

Congfeng Cao, Lipeng Zuo, Konstantinos Papakostas et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

StegoMemory: Agentic Memory Acts as Covert Steganographic Channel

Is agentic memory robust against stealthy steganographic attacks? We carry out a large-scale red-teaming exercise to test whether agents can encode attacker-controlled strings in one session and recover them in another without triggering safety oversight. Following SHADE-Arena-style tasks, we embed malicious side tasks...

Snehasis Mukhopadhyay, Arun Nair · 0 citations
#natural language process... Preprint Open access Oct 2026

The Same Zero: Why Identical ASR Can Imply Different Guarantees in LLM-Agent Security

LLM-agent security has produced a dense landscape of defenses - prompt hardening, content filters, permission gates, sandboxes - yet no framework tells a deployer what a defense actually guarantees, or where that guarantee comes from. We apply Verification Autonomy Levels (VAL) - L0: LLM self-declaration; L1: determini...

YaJie Yin · 0 citations
#natural language process... Preprint Open access Oct 2026

Large Language Models and Augmented Democracy

Artificial intelligence enables computational agents to represent political preferences and take part in collective decision-making. In this thesis, I investigate the opportunities and challenges of digital twins (DTs) based on Large Language Models (LLMs) as intermediaries in augmented democracy, focusing on individua...

Jairo Gudi\~no-Rosero · 0 citations
#natural language process... Preprint Open access Oct 2026

Factorized Delayed Streams Modeling for LLM-based Streaming ASR

Delayed Streams Modeling (DSM) enables LLM-based streaming automatic speech recognition (ASR) by aligning acoustic and text streams on a common timeline. DSM adds the padding token <p> and the word-start token <w> to the LLM vocabulary and predicts them together with normal text tokens using the same softmax. We first...

Tatsunari Takagi, Kai Washizaki, Atsushi Kojima et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Sep 24, 2026

Estimating suicide risk from text

A new language-processing tool could help identify the highest-risk individuals from natural language, enabling swifter interventions.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.