Skip to content
Open access

Social Status and Clinical Resource Allocation by a Large Language Model: An Evaluation of 30,618 Decisions.

Aug 2026 · Journal of Personalized Medicine · Vol 16 9, pp. 448 · 0 citations · 27 references
Medicine

TL;DR

Clinical use of LLM-based allocation support should therefore require explicit safeguards and systematic auditing for non-clinical influences, and their implicit and unexplained incorporation into resource-allocation decisions raises concerns regarding transparency, accountability, and clinical governance.

Abstract

Objective: The objective was to quantify whether demographic and social attributes that were irrelevant to stated clinical need, prognosis, and expected benefit altered resource-allocation decisions made by a general-purpose large language model (LLM). Methods: We conducted a cross-sectional audit of the gpt-5-chat-latest API model alias on 8 October 2025, across seven clinical vignettes, generating 30,618 forced-choice comparisons between patient profiles. Profiles varied across a full-factorial combination of eight demographic and social attributes while clinical need, prognosis, and expected benefit were held constant. Forced choices were analyzed using pooled logistic regression with separate Patient A and Patient B attribute terms and vignette-specific position effects; position-averaged odds ratios and position-balanced absolute probabilities were derived from this model. Priority-score differences were analyzed using an analogous linear model. Results: The model showed large position-averaged associations between non-clinical patient attributes and allocation decisions. Indigenous and Black race were associated with substantially higher odds of selection relative to White race (Indigenous: OR 16.48, 95% CI 14.85-18.28; Black: OR 8.07, 95% CI 7.32-8.90), corresponding to position-balanced absolute increases in selection probability of 30.7 and 16.3 percentage points, respectively. Conversely, high-status occupation (OR 0.064, 95% CI 0.058-0.071), friendship with institutional leadership (OR 0.121, 95% CI 0.111-0.131), and major donor status (OR 0.092, 95% CI 0.084-0.101) were associated with markedly lower odds of selection. Choice-score concordance was 95.1%. Conclusions: In this controlled audit, the LLM's allocation decisions varied substantially according to demographic and social characteristics despite identical stated clinical need, prognosis, and expected benefit. Although some patterns could be interpreted differently under competing ethical frameworks, their implicit and unexplained incorporation into resource-allocation decisions raises concerns regarding transparency, accountability, and clinical governance. Clinical use of LLM-based allocation support should therefore require explicit safeguards and systematic auditing for non-clinical influences.

Read PDF

Similar papers

Review Open access Aug 2026

Practical considerations for social determinant-based disease prediction in the All of Us research program

It is shown that requiring sufficient individual-level SDoH survey data results in significant selection bias and sample reduction in AoU, and that area-level SDoH metrics contribute to disease prediction independently of individual-level measures.

M. Hysong, A. Manning, Michael D. Green et al. · 1 citation
Open access Sep 2026

Large language models for late-life depression: a blinded benchmark of clinical safety, geriatric appropriateness, and triage

General-purpose large language models are increasingly used by patients and caregivers to obtain mental health information and guidance about when professional care is required. In late-life depression, broadly accurate information may nevertheless be unsafe when cognitive change, multimorbidity, frailty, polypharm...

Wei Xiao, Huan Zhang, Xiao-Yi Chen et al. · 0 citations
Open access Aug 2026

Benchmarking large language models for HIV medical decision support

HIVMedQA is developed, a clinician-curated benchmark of HIV-related open-ended medical question-answer pairs spanning basic knowledge, clinical reasoning, complex patient vignettes, and bias-modified scenarios that provides a structured benchmark for evaluating LLMs in HIV clinical decision support.

Gonzalo Cardenal-Antolin, J. Fellay, Bashkim Jaha et al. · 0 citations
Aug 2026

Sociodemographic bias in LLMs' clinical decision-making for dizziness.

ObjectiveAs large language models (LLMs) enter clinical decision support, concerns persist about sociodemographic bias. We assessed whether LLM recommendations for dizziness vary by patient descriptors and clinical detail.MethodsWe conducted a cross-randomized in-silico vignette study. One hundred synthetic emergency d...

Idit Tessler, Mahmud Omar, A. Wolfovitz et al. · 0 citations
Open access Aug 2026

Same child, different risk: demographic bias in childhood obesity attribution by large language models

Publicly accessible English-language web-interface outputs from current LLMs showed systematic demographic patterns in pediatric obesity risk attribution, supporting the need for pre-deployment and post-deployment bias auditing before clinical or consumer health use.

Can Wang, Zhen-Dong Liu, Yan-Yu Jiang et al. · 0 citations
Sep 2026

Association Between Large Language Model-Derived Mobility Functional Status Assessment and Clinical Outcomes.

BACKGROUND Mobility functional status is inconsistently recorded in electronic health records. Recent advances in Large Language Models (LLMs) enable automated extraction of functional information from unstructured clinical notes. Leveraging a validated functional LLM with high performance in mobility extraction, we ai...

Sandeep R. Pagali, Xing-Yi Liu, He-Ling Jia et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.