CLMIR: A Textual Dataset for Rumor Identification and Marking

Bin Ma; Yifei Zhang; Yongjin Xian; Qi Li; Linna Zhou; Gongxun Miao

arXiv:2508.11138·cs.CY·August 18, 2025

CLMIR: A Textual Dataset for Rumor Identification and Marking

Bin Ma, Yifei Zhang, Yongjin Xian, Qi Li, Linna Zhou, Gongxun Miao

PDF

TL;DR

This paper introduces CLMIR, a new dataset for rumor detection that not only identifies rumors but also marks the specific content within posts that constitutes the rumor, enhancing interpretability.

Contribution

The creation of CLMIR, a fine-grained rumor dataset that enables content marking, facilitating more interpretable and precise rumor detection algorithms.

Findings

01

Enables training of rumor detection models with content marking

02

Improves interpretability and reasoning in rumor detection systems

03

Supports practical applications like rumor tracing and moderation

Abstract

With the rise of social media, rumor detection has drawn increasing attention. Although numerous methods have been proposed with the development of rumor classification datasets, they focus on identifying whether a post is a rumor, lacking the ability to mark the specific rumor content. This limitation largely stems from the lack of fine-grained marks in existing datasets. Constructing a rumor dataset with rumor content information marking is of great importance for fine-grained rumor identification. Such a dataset can facilitate practical applications, including rumor tracing, content moderation, and emergency response. Beyond being utilized for overall performance evaluation, this dataset enables the training of rumor detection algorithms to learn content marking, and thus improves their interpretability and reasoning ability, enabling systems to effectively address specific rumor…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.