Skip to content
Conference

Combining Emotion Perception and Keywords to Support Emotional Dialogue Research

Jul 2026 · 2026 6th International Conference on Electrical, Computer and Energy Technologies (ICECET) · pp. 1-6 · 0 citations · 18 references

Abstract

Recent advances in artificial intelligence have enabled the development of more natural and effective human-machine interactions. In this study, we propose an enhanced BART-based dialogue generation framework that integrates emotion labels and keyword annotations to support emotionally aware end-to-end conversations. The proposed system combines an independent RoBERTa-based emotion classification module with a terminology-aware annotation mechanism, in which emotional labels and key semantic terms are explicitly injected into the dialogue context. The framework is evaluated on a merged dataset constructed from two benchmark dialogue corpora, Empathetic Dialogues and DailyDialog, which have been relabeled into nine unified emotion categories. Automatic evaluation is conducted using BLEU, ROUGE, and perplexity metrics. Experimental results demonstrate that the proposed approach consistently outperforms baseline models, including GPT-2, DialoGPT, T5, UniLM, and the original BART model. In particular, the incorporation of keyword annotations significantly improves response relevance and informational completeness, while emotion labels further enhance emotional alignment. These findings suggest that explicitly modeling emotional cues and semantic focus can effectively improve the quality of emotional dialogue generation, offering a promising direction for the development of emotionally intelligent conversational agents.

View source

Similar papers

Preprint Jul 2026

Dialogue Summarization with Emotion Dynamics Using Topic- and Participant-Centric Decomposition

Existing text summarization research has focused much on monologic information (e.g., newspaper articles, reports) without accounting for the interaction between speakers or authors. In contrast, dialogues are a rich communication channel where multiple participants conduct back and forth exchanges to construct meaning. We propose a dialogue summarization framework that explicitly models both semantic and emotion dynamics using multimodal dialogue inputs, built on an adapted hierarchical Chain-of-Agents approach. We decompose dialogues from two perspectives: (1) topic segments based on the utterances of all participants, and (2) participant-specific utterance segments. These are used to generate corresponding summaries while incorporating automatically inferred emotions. Topic- and participant-level summaries are aggregated into a dialogue summary capturing semantic content and emotion trajectories. To evaluate beyond content accuracy, we introduce emotion trajectory metrics measuring how well summaries preserve emotional flow. Experiments with small language models on multimodal dialogue datasets show that our framework produces summaries with both semantic and emotion content. Further experiments on explicit emotion label availability highlight the efficacy of our proposed methodology and the opportunities in dialogue analysis using language models.

Linyun Xiang, Mark Antonius Neerincx, Stephanie Tan · 0 citations
Jul 2026

AtmosERC: Modeling Dialogue-Level Affective Atmosphere for Emotion Recognition in Conversation

Emotion Recognition in Conversation (ERC) aims to predict utterance-level emotions in dialogues and has largely advanced through context-centric modeling. However, global context is a heterogeneous signal, and not all contextual information is equally relevant to emotion prediction. This paper focuses on the affect-oriented component of this signal, termed dialogue-level affective atmosphere, which captures a latent tendency commonly reflected in conversational emotion patterns. To estimate and exploit this tendency, we propose AtmosERC, a graph-based ERC framework that models each dialogue as a conversational graph over utterances and speakers. A relation-aware graph extractor filters and fuses heterogeneous graph signals to produce dialogue-level and speaker-conditioned affective priors. The resulting compact prior guides lightweight sequential emotion prediction and can also be verbalized into prompt-level cues for LLM-based ERC without modifying backbone models. Experiments on four ERC benchmarks show that AtmosERC improves lightweight ERC, enhances LLM-based ERC as a plug-in cue, and yields more stable predictions under local emotional deviations.

Weijie Feng, Tong Zhang, Binbin Liu et al. · 0 citations

TalkFa: A Unified Benchmark for Farsi Dialogue Generation and Understanding

Farsi, spoken by more than 120 million people, lacks a comprehensive benchmark for dialogue generation and understanding. We introduce TALKFA, a unified benchmark comprising three complementary datasets: (1) WIKI-FADIAL, 4.2K Wikipedia-grounded dialogues for knowledge-grounded generation; (2) DAILYDIALOG-FA, 6.6K dialogues annotated for dialogue acts and emotions; and (3) PLAYDIAL-FA, 2.1K theatrical dialogues with sentiment labels. While LLMs assist data construction, every dialogue undergoes multi-stage review and revision by native Farsi speakers, and only the final human-approved dialogues are released. Experiments with six LLAMA and MISTRAL models show that LoRA substantially improves dialogue generation while requiring only 25-50% of the training data to recover over 90% of the final performance gains. Across classification tasks, FABERT achieves the best dialogue-act performance, LORA-MISTRAL-7B performs best on emotion recognition, and MISTRAL-24B achieves the highest sentiment score. Human evaluation and independent external validation demonstrate the reliability of the benchmark, while comparisons with GPT-4.1 as an LLM judge reveal that automatic metrics substantially overestimate dialogue quality. Zero-shot evaluation with frontier LLMs further shows that TalkFa remains a challenging benchmark. We will release all datasets, annotation guidelines, code, and checkpoints.

Neda Jamshidi, Kamyar Zeinalipour, F. Akbari et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.