Skip to content
Preprint

What Makes a Good Fiqh Retriever? Answer Retrieval for Arabic Islamic Jurisprudence

Aug 2026 · 0 citations · 33 references
Computer Science

TL;DR

This work builds a retrieval test collection for Arabic fiqh and uses it to evaluate dense, lexical, hybrid, fine-tuned, and madhhab-aware retrieval strategies, and presents an error analysis showing that the main challenge is distinguishing answer-bearing passages from topically similar passages that do not contain the requested ruling.

Abstract

Retrieval-Augmented Generation is used for Islamic question answering, but most systems are evaluated end-to-end, making retrieval failures difficult to isolate from generation failures. We study answer-bearing retrieval for Arabic fiqh, where a passage is relevant only if it states the ruling required by the question. We build a retrieval test collection for Arabic fiqh and use it to evaluate dense, lexical, hybrid, fine-tuned, and madhhab-aware retrieval strategies. The best retriever achieves 0.524 MRR@5, while fine-tuning improves performance to 0.553. Hybrid retrieval provides limited gains for strong models, whereas madhhab-aware filtering more than doubles MRR@5 on school-specific questions. We further present an error analysis showing that the main challenge is distinguishing answer-bearing passages from topically similar passages that do not contain the requested ruling.

View source

Similar papers

#natural language process... Preprint Aug 2026

GreekBarRetrieval: A Benchmark for Greek Statutory Retrieval

Statutory retrieval is necessary for citation-grounded legal question answering, but remains underexplored for Greek. We introduce GreekBarRetrieval, a public retrieval benchmark derived from, and complementing GreekBarBench, which did not include retrieval. The new benchmark comprises 283 bar-exam questions, each acco...

Ernest Beta, Odysseas S. Chlapanis, D. Galanis et al. · 0 citations
Preprint Aug 2026

Rhetorical-Role-Aware Retrieval-Augmented Generation for Legal Question Answering over Indian Supreme Court Judgments

This research paper proposes a Retrieval Augmented Generation framework that is specific to the legal field in order to assist interactive retrieval and reason about judgments from the Supreme Court of India and demonstrates strong performance on metrics including contextual recall and answer relevancy.

Sayed Ayaan Ahmed Sha, Sangeetha Sivanesan, A. Madasamy et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Comparing Chunking and Embedding Strategies for Turkish RAG Systems

This work compares Turkish document question answering across three chunking strategies, five embedding models, and two LLMs, over three documents with contrasting layouts, finding the faster LLM is not the more accurate one.

Mustafa Sertac Turkel, Fatma Nur Korkmaz, Ahmet Tugrul Bayrak · 0 citations
Preprint Aug 2026

HybridRAG-BN: A Retrieval-Augmented Framework with Fine-Tuned Verification for Bangla KBQA

This work proposes HybridRAG-BN, a retrieval-augmented framework for Bangla KBQA that integrates hybrid retrieval using BM25 and BGE-M3, answer generation using the GGUF version of Gemma-4-31B-Instruct, and a LoRA-fine-tuned Gemma-4-31B-Instruct model for answer verification and refinement.

Rathijit Aich, Nirjhar Das, Mahfuzulhoq Chowdhury · 0 citations
Open access Aug 2026

Enhanced Hybrid Retrieval-Augmented Model for Question Answering in High-Sensitivity Domains

Arabic question-answering systems in high-sensitivity domains require not only accurate retrieval but also reliable evidence grounding and effective hallucination mitigation, as incorrect or unsupported responses may have serious consequences. Existing retrieval and generation approaches do not fully integrate reliable...

A. Aloqla, Reda Salama, Wajdi Alghamdi et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.