Sentiment analysis is a fundamental problem in Natural Language Processing (NLP). Standard sentiment classification for the Arabic language remains challenging due to the high volume of dialectal Arabic. To advance research in this area, this paper proposes the Shared Task on Sentiment Analysis and Swapping in Arabic D...
Saad Ezzini, Shadi Abudalfa, Maram Alharbi et al.· 0 citations
The shared task is a shared task for evaluating hallucination detection and factual verification in Arabic question answering under challenging generalization settings, based on two Arabic datasets: HalluScore and HalluTruthQA.
Aisha Alansari, Abdessalam Bouchekif, A. Hasanaath et al.· 0 citations
A lightweight quality-assessment protocol is presented for LLM-generated synthetic training data and applied to 13,579 synthetic user reviews generated from GitHub issues across four open-source Android applications, high-lighting the need for hybrid human-AI verification when synthetic data is used in security-critica...
AraGenre is a shared task on hierarchical, definition-guided Arabic genre classification, motivated by the limited availability of annotated data in Arabic and other low-resource languages. Systems assign each Arabic text segment both a broad communicative genre and a fine-grained specific genre. The released training...
Mo El-Haj, Saad Ezzini, Shadi Abudalfa et al.· 0 citations
Interactive storytelling has shown great potential in supporting children’s creativity. However, existing robotic storytelling systems mostly treat the robot as the primary author or lack the ability to adapt to a child’s interaction state in real time. In this paper, we present LLM-Storyteller, an embodied robotic sys...
Fatimah Alali, Saad Ezzini· Frontiers in Robotics and AI· 0 citations
Dialectal Arabic machine translation (MT) remains challenging despite recent progress in Arabic language technologies, particularly because effective translation requires modeling not only semantic content but also dialectal variation, conversational context, speaker and addressee characteristics, and sociolinguistic a...
Abdellah El Mekki, AbdelRahim A. Elmadany, Samar M. Magdy et al.· 1 citation
This paper presents the Mawqif-XT, consisting of 996 manually annotated Arabic tweets collected from three public targets: Women Driving, E-Cars, and Trimester System, which provides a benchmark for evaluating cross-target generalization in Arabic stance detection.
BUL, a multi-dialect Arabic ASR dataset collected from 275 speakers in 11 Arab countries, includes structured dialect and sub-dialect coverage, as well as recordings of classical Arabic and modern standard Arabic spoken by participants in their native dialectal accents to support accent-aware modeling.
Ahmed Ashraf, Aisha Alansari, Fadel Al Abbas et al.· 1 citation
This work introduces L\"etzCross, a benchmark for cross-lingual page-level retrieval over Luxembourgish PDF documents, with document pages indexed as images and queries provided in English, French, German, and Luxembourgish, and examines single-language and multilingual fine-tuning.
Omar El Bachyr, Fred Philippy, Laura Bernardy et al.· 0 citations
SysName, a production-oriented pipeline that automates device configuration end-to-end for Modbus RTU, OPC-UA, Profibus DP, and CANopen, builds a hybrid dense-sparse retrieval index augmented by an ontology graph derived from ECLASS, AAS, and SOSA/SSN, using a BGE-M3 encoder with a cross-encoder reranker to surface rel...
A. Ganie, Saad Ezzini, Naveed Farooz Marazi· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.