ReadNet: A Hierarchical Transformer Framework for Web Article   Readability Analysis

Changping Meng; Muhao Chen; Jie Mao; Jennifer Neville

arXiv:2103.04083·cs.IR·March 9, 2021

ReadNet: A Hierarchical Transformer Framework for Web Article Readability Analysis

Changping Meng, Muhao Chen, Jie Mao, Jennifer Neville

PDF

1 Repo

TL;DR

ReadNet introduces a hierarchical transformer framework that effectively captures sentence difficulty, semantic content, and article structure to improve web article readability assessment, outperforming existing methods.

Contribution

This paper presents a novel hierarchical self-attention model that integrates sentence difficulty, semantics, and structure for enhanced readability analysis.

Findings

01

Achieves state-of-the-art performance on benchmark datasets.

02

Effectively models complex article structures and semantics.

03

Outperforms strong baseline approaches.

Abstract

Analyzing the readability of articles has been an important sociolinguistic task. Addressing this task is necessary to the automatic recommendation of appropriate articles to readers with different comprehension abilities, and it further benefits education systems, web information systems, and digital libraries. Current methods for assessing readability employ empirical measures or statistical learning techniques that are limited by their ability to characterize complex patterns such as article structures and semantic meanings of sentences. In this paper, we propose a new and comprehensive framework which uses a hierarchical self-attention model to analyze document readability. In this model, measurements of sentence-level difficulty are captured along with the semantic meanings of each sentence. Additionally, the sentence-level features are incorporated to characterize the overall…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

vdefont/readnet
pytorch

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.