Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs

Wenrui Zhou; Mohamed Hendy; Shu Yang; Qingsong Yang; Zikun Guo; Yuyu Luo; Lijie Hu; Di Wang

arXiv:2506.07180·cs.CL·May 1, 2026

Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs

Wenrui Zhou, Mohamed Hendy, Shu Yang, Qingsong Yang, Zikun Guo, Yuyu Luo, Lijie Hu, Di Wang

PDF

2 Repos

TL;DR

This paper introduces VISE, a comprehensive benchmark for evaluating sycophantic tendencies in Video-LLMs, and explores strategies to mitigate such biases without additional training.

Contribution

It presents the first systematic benchmark for Video-LLMs' sycophancy, analyzing its manifestations and proposing inference-time mitigation methods.

Findings

01

VISE enables detailed analysis of sycophantic behavior in Video-LLMs.

02

Two inference-time strategies can reduce sycophantic bias without retraining.

03

The benchmark covers diverse question formats and visual reasoning tasks.

Abstract

As video large language models (Video-LLMs) become increasingly integrated into real-world applications that demand grounded multimodal reasoning, ensuring their factual consistency and reliability is of critical importance. However, sycophancy, the tendency of these models to align with user input even when it contradicts the visual evidence, undermines their trustworthiness in such contexts. Current sycophancy research has largely overlooked its specific manifestations in the videolanguage domain, resulting in a notable absence of systematic benchmarks and targeted evaluations to understand how Video-LLMs respond under misleading user input. To fill this gap, we propose VISE(Video-LLM Sycophancy Benchmarking and Evaluation), the first benchmark designed to evaluate sycophantic behavior in state-of-the-art Video-LLMs across diverse question formats, prompt biases, and visual reasoning…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.