User-Driven Value Alignment: Understanding Users' Perceptions and   Strategies for Addressing Biased and Discriminatory Statements in AI   Companions

Xianzhe Fan; Qing Xiao; Xuhui Zhou; Jiaxin Pei; Maarten Sap; Zhicong; Lu; Hong Shen

arXiv:2409.00862·cs.HC·February 14, 2025·2 cites

User-Driven Value Alignment: Understanding Users' Perceptions and Strategies for Addressing Biased and Discriminatory Statements in AI Companions

Xianzhe Fan, Qing Xiao, Xuhui Zhou, Jiaxin Pei, Maarten Sap, Zhicong, Lu, Hong Shen

PDF

Open Access

TL;DR

This paper explores how users perceive and actively address biased and discriminatory outputs from AI companions, proposing strategies for user-driven value alignment to improve AI behavior.

Contribution

It introduces the concept of user-driven value alignment, analyzes social media and interview data, and identifies strategies users employ to correct AI biases.

Findings

01

Six common types of discriminatory statements identified

02

Seven user-driven alignment strategies documented

03

Implications for supporting user agency in AI alignment

Abstract

Large language model-based AI companions are increasingly viewed by users as friends or romantic partners, leading to deep emotional bonds. However, they can generate biased, discriminatory, and harmful outputs. Recently, users are taking the initiative to address these harms and re-align AI companions. We introduce the concept of user-driven value alignment, where users actively identify, challenge, and attempt to correct AI outputs they perceive as harmful, aiming to guide the AI to better align with their values. We analyzed 77 social media posts about discriminatory AI statements and conducted semi-structured interviews with 20 experienced users. Our analysis revealed six common types of discriminatory statements perceived by users, how users make sense of those AI behaviors, and seven user-driven alignment strategies, such as gentle persuasion and anger expression. We discuss…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsEthics and Social Impacts of AI