Skip to content
Preprint

Subjective Multi-Bias Detection with Large Language Models

Aug 2026 · 0 citations · 30 references
Computer Science

TL;DR

This project delved into the pervasive challenge of bias detection within the text content by detecting three different types of multi-span biases in corpus WIKIBIAS with more than 4,000 sentence pairs from Wikipedia edits.

Abstract

In this project, we delved into the pervasive challenge of bias detection within the text content. More specifically, our focus lies on the identification of subjective bias, a type of bias that introduces improper attitudes or portrays a statement at odds with the actual truth. The subjective bias can jeopardize the authenticity and reliability of texts, leading to misconceptions and potential social tensions, especially when expressed through offensive language. Following prior work [1], we tackled with three different types of subjective biases in text: (1) framing bias with the use of one-sided words or phrases containing a particular point of view; (2) epistemological bias which includes subtle linguistic features that can affect the believability of the texts; (3) demographic bias with word/phrase usage under presuppositions of a particular demographic factor (i.e., gender or religion). In terms of the data we utilize, the input consists of texts that may harbor subjective biases. The output is a classification or annotation that reveals the presence or absence of such biases within the provided content. More specifically, we detected three different types of multi-span biases in corpus WIKIBIAS [2] with more than 4,000 sentence pairs from Wikipedia edits. The data is labelled by bias type for span pairs with the following categories: (1) framing bias, (2) epistemological bias, (3) demographic bias, and (4) no bias. The project codes are released at https://github.com/HoningJade/LLM-Bias-Type-Classification.

View source

Similar papers

Open access Sep 2026

BiasScope: Inference-Time Bias Detection and Mitigation for Fair Natural Language Processing (NLP) Using Retrieval-Augmented Generation and Prompt Engineering

BiasScope is a lightweight framework that uses prompt engineering and RAG to detect and mitigate model biases at inference time without changing weights, and provides an effective means of achieving fairness and accountability in production Natural Language Processing (NLP).

Rakeshkumarreddy Ambati · 0 citations
Review Open access Aug 2026

A Comprehensive Survey of News Bias Detection Datasets

The quality, relevance, and annotation guidelines of the data have a key influence on the accuracy and effectiveness of models that could perform automatic bias detection.

Pooja B. Bhise, S. Govilkar, Sheetal B. Gawande · 0 citations
#machine learning Preprint Sep 2026

When Noise Fabricates Bias: The Fragility of LLM-as-a-Judge Bias Measurement under Noisy Text

Large language models are increasingly used as judges to measure social bias in text, yet the passages they judge are often noisy, containing typos, informal spelling, and broken punctuation. The consequences of such surface noise for social bias measurement remain unclear. To investigate this question, we apply five r...

DongHyun Ryu, Jaehyeok Lee, Yeongjun Hwang et al. · 0 citations
Aug 2026

Quantifying Social Biases in Language Model Classifiers is Domain-Dependent

This work investigates whether large language models (LLMs) can automatically adapt template-based bias datasets to specific domains using zero-shot prompting and shows that domain-adapted templates capture real-world bias patterns more faithfully than standard templates.

Tamara Quiroga, Felipe Bravo-Marquez, Valentin Barrière · 0 citations
Open access Sep 2026

The Media Bias Detector: A framework for annotating and analyzing the news

News organizations introduce bias into their coverage via the choices they make about which topics to cover (or ignore) and how to frame the issues they do decide to cover. Here, we introduce the Media Bias Detector, a scalable computational framework that integrates large language models (LLMs) with near-real-time new...

Samar Haider, Amir Tohidi, Jenny S. Wang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.