# Measuring Similarity: Computationally Reproducing the Scholar's   Interests

**Authors:** Ashley Lee, Jo Guldi, Andras Zsom

arXiv: 1812.05984 · 2018-12-17

## TL;DR

This paper explores how computational text classification methods can be made transparent and understandable to humanists, enabling critique and improvement of personalized document grouping algorithms.

## Contribution

It introduces a framework for translating computational classification procedures into human-readable terms for scholarly critique.

## Key findings

- Proposes methods for translating algorithms into human-understandable language
- Highlights opportunities for expert critique of automated classification
- Suggests improvements for transparency in personalized search algorithms

## Abstract

Computerized document classification already orders the news articles that Apple's "News" app or Google's "personalized search" feature groups together to match a reader's interests. The invisible and therefore illegible decisions that go into these tailored searches have been the subject of a critique by scholars who emphasize that our intelligence about documents is only as good as our ability to understand the criteria of search. This article will attempt to unpack the procedures used in computational classification of texts, translating them into term legible to humanists, and examining opportunities to render the computational text classification process subject to expert critique and improvement.

---
Source: https://tomesphere.com/paper/1812.05984