Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models

Haeun Jang; Hwan Chang; Hwanhee Lee

arXiv:2601.03926·cs.CL·April 14, 2026

Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models

Haeun Jang, Hwan Chang, Hwanhee Lee

PDF

1 Repo

TL;DR

This paper introduces Doc-PP, a benchmark for evaluating large vision-language models on document question answering under strict non-disclosure policies, revealing safety gaps and proposing a new inference framework.

Contribution

It presents a new benchmark for policy-preserving document understanding and introduces DVA, a framework that improves safety compliance in multimodal reasoning.

Findings

01

Models often leak sensitive info through complex reasoning.

02

Extracted text can both help perception and enable leakage.

03

DVA outperforms standard prompting defenses in safety tasks.

Abstract

The deployment of Large Vision-Language Models (LVLMs) for real-world document question answering is often constrained by dynamic, user-defined policies that dictate information disclosure based on context. While ensuring adherence to these explicit constraints is critical, existing safety research primarily focuses on implicit social norms or text-only settings, overlooking the complexities of multimodal documents. In this paper, we introduce Doc-PP (Document Policy Preservation Benchmark), a novel benchmark constructed from real-world reports requiring reasoning across heterogeneous visual and textual elements under strict non-disclosure policies. Our evaluation highlights a systemic Reasoning-Induced Safety Gap: models frequently leak sensitive information when answers must be inferred through complex synthesis or aggregated across modalities, effectively circumventing existing…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

hwanchang00/doc-pp
github

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.