Let's Measure the Elephant in the Room: Facilitating Personalized Automated Analysis of Privacy Policies at Scale

Rui Zhao; Vladyslav Melnychuk; Jun Zhao; Jesse Wright; Nigel Shadbolt

arXiv:2507.14214·cs.CL·July 22, 2025

Let's Measure the Elephant in the Room: Facilitating Personalized Automated Analysis of Privacy Policies at Scale

Rui Zhao, Vladyslav Melnychuk, Jun Zhao, Jesse Wright, Nigel Shadbolt

PDF

TL;DR

This paper presents PoliAnalyzer, a neuro-symbolic system that automates personalized privacy policy analysis at scale, helping users understand policy compliance with their preferences using NLP and formal reasoning.

Contribution

It introduces PoliAnalyzer, combining NLP and formal logic to analyze privacy policies against user preferences, enabling scalable, personalized privacy compliance assessments.

Findings

01

Achieved 90-100% F1-score in policy compliance detection.

02

Found 95.2% of policy segments align with user preferences.

03

Identified common policy violations like location data sharing.

Abstract

In modern times, people have numerous online accounts, but they rarely read the Terms of Service or Privacy Policy of those sites despite claiming otherwise. This paper introduces PoliAnalyzer, a neuro-symbolic system that assists users with personalized privacy policy analysis. PoliAnalyzer uses Natural Language Processing (NLP) to extract formal representations of data usage practices from policy texts. In favor of deterministic, logical inference is applied to compare user preferences with the formal privacy policy representation and produce a compliance report. To achieve this, we extend an existing formal Data Terms of Use policy language to model privacy policies as app policies and user preferences as data policies. In our evaluation using our enriched PolicyIE dataset curated by legal experts, PoliAnalyzer demonstrated high accuracy in identifying relevant data usage practices,…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.