Low-Rank Contextual Reinforcement Learning from Heterogeneous Human Feedback

Seong Jin Lee; Will Wei Sun; Yufeng Liu

arXiv:2412.19436·stat.ML·March 5, 2026

Low-Rank Contextual Reinforcement Learning from Heterogeneous Human Feedback

Seong Jin Lee, Will Wei Sun, Yufeng Liu

PDF

Open Access

TL;DR

This paper introduces LoCo-RLHF, a novel framework for reinforcement learning from heterogeneous human feedback that leverages low-rank structures and pessimistic policies to improve personalization and robustness.

Contribution

The paper proposes a low-rank contextual model for RLHF and a new pessimistic policy to handle distributional shifts, advancing personalized alignment methods.

Findings

01

Outperforms existing methods in personalized RLHF tasks

02

Demonstrates robustness to distributional shifts in feedback

03

Achieves tighter sub-optimality bounds theoretically

Abstract

Reinforcement learning from human feedback (RLHF) has become a cornerstone for aligning large language models with human preferences. However, the heterogeneity of human feedback, driven by diverse individual contexts and preferences, poses significant challenges for reward learning. To address this, we propose a Low-rank Contextual RLHF (LoCo-RLHF) framework that integrates contextual information to better model heterogeneous feedback while maintaining computational efficiency. Our approach builds on a contextual preference model, leveraging the intrinsic low-rank structure of the interaction between user contexts and query-answer pairs to mitigate the high dimensionality of feature representations. Furthermore, we address the challenge of distributional shifts in feedback through our Pessimism in Reduced Subspace (PRS) policy, inspired by pessimistic offline reinforcement learning…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsNeural Networks and Applications · EEG and Brain-Computer Interfaces