Global Contentious Politics Database (GLOCON) Annotation Manuals
F{\i}rat Duru\c{s}an, Ali H\"urriyeto\u{g}lu, Erdem Y\"or\"uk, Osman, Mutlu, \c{C}a\u{g}r{\i} Yoltar, Burak G\"urel, Alvaro Comin

TL;DR
This paper introduces the GLOCON annotation manuals, which standardize the coding of protest events in news articles to support automated detection and extraction for social science research.
Contribution
It provides detailed annotation guidelines adapted from linguistic standards for coding protest events, enhancing the quality and consistency of the GLOCON Gold Standard Corpus.
Findings
Manual ensures high annotation accuracy and consistency.
Guidelines are adapted from established linguistic annotation standards.
The corpus supports automated protest event detection.
Abstract
The database creation utilized automated text processing tools that detect if a news article contains a protest event, locate protest information within the article, and extract pieces of information regarding the detected protest events. The basis of training and testing the automated tools is the GLOCON Gold Standard Corpus (GSC), which contains news articles from multiple sources from each focus country. The articles in the GSC were manually coded by skilled annotators in both classification and extraction tasks with the utmost accuracy and consistency that automated tool development demands. In order to assure these, the annotation manuals in this document lay out the rules according to which annotators code the news articles. Annotators refer to the manuals at all times for all annotation tasks and apply the rules that they contain. The content of the annotation manual is built on…
Peer Reviews
No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.
Videos
No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.
Taxonomy
TopicsSocial Media and Politics
