Skip to content

Category

natural language processing

6,613 papers

#artificial intelligence Preprint Open access Oct 2026

Cross-Cultural Value Attribution in Large Vision-Language Models

The rapid adoption of large vision-language models (LVLMs) in recent years has been accompanied by growing fairness concerns due to their propensity to reinforce harmful societal stereotypes. While significant attention has been paid to such fairness concerns in the context of social biases, relatively little prior wor...

Phillip Howard, Xin Su, Kathleen C. Fraser · 0 citations
#natural language process... Preprint Open access Oct 2026

Quantifying Cross-Lingual Transfer in Paralinguistic Speech Tasks

Paralinguistic speech tasks are often considered relatively language-agnostic, as they rely on extralinguistic acoustic cues rather than lexical content. However, prior studies report performance degradation under cross-lingual conditions, indicating non-negligible language dependence. Still, these studies typically fo...

Pol Buitrago, Oriol Pareras, Federico Costa et al. · 0 citations
#computer vision Preprint Open access Oct 2026

AVMeme Exam: A Multimodal Multilingual Multicultural Benchmark for LLMs' Contextual and Cultural Knowledge and Thinking

Internet audio-visual clips convey meaning through time-varying sound and motion, which extend beyond what text alone can represent. To examine whether AI models can understand such signals in human cultural contexts, we introduce AVMeme Exam, a human-curated benchmark of over one thousand iconic Internet sounds and vi...

Xilin Jiang, Qiaolin Wang, Junkai Wu et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

Zero-Shot Lombard Speech Synthesis with Controllable Style Embeddings

The Lombard effect plays a key role in natural communication, particularly in noisy environments or when addressing hearing-impaired listeners. We present a controllable text-to-speech (TTS) system capable of synthesizing Lombard-like speech in a zero-shot manner without requiring Lombard-specific training data. Our ap...

Seymanur Akti, Alexander Waibel · 0 citations
#natural language process... Preprint Open access Oct 2026

Enhancing High-order Interaction Awareness in LLM-based Recommender Model

Large language models (LLMs) have demonstrated prominent reasoning capabilities in recommendation tasks by transforming them into text-generation tasks. However, existing approaches either disregard or ineffectively model the user-item high-order interactions. To this end, this paper presents an enhanced LLM-based reco...

Xinfeng Wang, Jin Cui, Fumiyo Fukumoto et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

Evaluating Large Language Model Raters for German Open-Response Clinical Questions: A Physician-Annotated Benchmark Study of Agreement, Evaluator Bias, and Abstention

Background: Expert-annotated benchmarks for non-English open-response clinical questions are scarce. LLM-as-a-judge systems may scale evaluation but require validation. Objective: To introduce MedQADE, a standardized German open-response clinical benchmark with physician reference annotations, and evaluate LLM-as-a-j...

William Philipp, Finn Fassbender, Daniel Fister et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

Emotion Recognition in Sign Language Conversation

Emotion Recognition in Conversation is a core component of affective computing, while current sign language emotion datasets primarily focus on isolated sentences and lack conversational context. Models trained exclusively on these isolated utterances demonstrate degraded performance in real world scenarios because the...

Yusong Wang, Keyu Mao, Takao Obi et al. · 0 citations
#natural language process... Preprint Open access Oct 2026

Cooperative Profiles Predict Multi-Agent LLM Team Performance in AI for Science Workflows

Multi-agent systems built from teams of large language models (LLMs) are increasingly deployed for collaborative scientific reasoning and problem-solving. These systems require agents to coordinate under shared constraints, such as GPUs or credit balances, where cooperative behavior matters. Behavioral economics provid...

Shivani Kumar, Adarsh Bharathwaj, David Jurgens · 0 citations
#natural language process... Preprint Open access Oct 2026

Model in Distress: Sentiment Analysis on French Synthetic Social Media

Automated analysis of customer feedback on social media is hindered by three challenges: the high cost of annotated training data, the scarcity of evaluation sets, especially in multilingual settings, and privacy concerns that prevent data sharing and reproducibility. We address these issues by developing a generalizab...

Pierre-Carl Langlais, Pavel Chizhov, Yannick Detrois et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Unbiased Reward Modeling from Implicit Feedback for LLM Alignment

Despite the success of reinforcement learning from human feedback (RLHF), existing reward modeling methods largely rely on explicit feedback, which is costly to collect and difficult to scale. This work studies implicit reward modeling, learning reward models from implicit user feedback, such as clicks, copies and skip...

Hao Wang, Haocheng Yang, Licheng Pan et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Understanding Moral Reasoning Trajectories in Large Language Models: Toward Probing-Based Explainability

Large language models (LLMs) increasingly participate in morally sensitive decision-making, yet how they organize ethical frameworks across reasoning steps remains underexplored. We introduce moral reasoning trajectories, sequences of ethical framework invocations across intermediate reasoning steps, and analyze their...

Fan Huang, Haewoon Kwak, Jisun An · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Real-Time Generation of Game Video Commentary with Multimodal LLMs: Pause-Aware Decoding Approaches

Real-time video commentary generation provides textual descriptions of ongoing events in videos. It supports accessibility and engagement in domains such as sports, esports, and livestreaming. Commentary generation involves two essential decisions: what to say and when to say it. While recent prompting-based approaches...

Anum Afzal, Yuki Saito, Hiroya Takamura et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Sep 24, 2026

Estimating suicide risk from text

A new language-processing tool could help identify the highest-risk individuals from natural language, enabling swifter interventions.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.