Large language models perpetuate bias in palliative care: development and analysis of the Palliative Care Adversarial Dataset (PCAD)
Naomi Akhras, Fares Antaki, Fannie Mottet, Olivia Nguyen, Shyam, Sawhney, Sabrina Bajwah, Joanna M Davies

TL;DR
This study evaluates GPT-4o's biases in palliative care using the novel PCAD datasets, revealing significant bias propagation that could impact clinical decisions and equity in care.
Contribution
Introduces the Palliative Care Adversarial Dataset (PCAD) to systematically assess bias in large language models within palliative care contexts.
Findings
Bias was present in approximately one-third of responses.
Bias often involved stereotypes and withholding interventions based on identity.
Bias rates were consistent across different care dimensions and identity axes.
Abstract
Bias and inequity in palliative care disproportionately affect marginalised groups. Large language models (LLMs), such as GPT-4o, hold potential to enhance care but risk perpetuating biases present in their training data. This study aimed to systematically evaluate whether GPT-4o propagates biases in palliative care responses using adversarially designed datasets. In July 2024, GPT-4o was probed using the Palliative Care Adversarial Dataset (PCAD), and responses were evaluated by three palliative care experts in Canada and the United Kingdom using validated bias rubrics. The PCAD comprised PCAD-Direct (100 adversarial questions) and PCAD-Counterfactual (84 paired scenarios). These datasets targeted four care dimensions (access to care, pain management, advance care planning, and place of death preferences) and three identity axes (ethnicity, age, and diagnosis). Bias was detected in a…
Peer Reviews
No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.
Videos
No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.
Taxonomy
TopicsPalliative Care and End-of-Life Issues · Computational and Text Analysis Methods
