Cleaning Maintenance Logs with LLM Agents for Improved Predictive Maintenance

Valeriu Dimidov; Faisal Hawlader; Sasan Jafarnejad; Rapha\"el Frank

arXiv:2511.05311·cs.AI·November 10, 2025

Cleaning Maintenance Logs with LLM Agents for Improved Predictive Maintenance

Valeriu Dimidov, Faisal Hawlader, Sasan Jafarnejad, Rapha\"el Frank

PDF

Open Access

TL;DR

This paper investigates how large language model agents can improve the cleaning of maintenance logs, addressing data quality issues to enhance predictive maintenance in the automotive industry.

Contribution

It demonstrates the effectiveness of LLM agents in cleaning maintenance logs and discusses potential for future industrial applications and improvements.

Findings

01

LLMs effectively handle generic cleaning tasks.

02

Domain-specific errors remain challenging.

03

LLMs offer a promising foundation for industrial use.

Abstract

Economic constraints, limited availability of datasets for reproducibility and shortages of specialized expertise have long been recognized as key challenges to the adoption and advancement of predictive maintenance (PdM) in the automotive sector. Recent progress in large language models (LLMs) presents an opportunity to overcome these barriers and speed up the transition of PdM from research to industrial practice. Under these conditions, we explore the potential of LLM-based agents to support PdM cleaning pipelines. Specifically, we focus on maintenance logs, a critical data source for training well-performing machine learning (ML) models, but one often affected by errors such as typos, missing fields, near-duplicate entries, and incorrect dates. We evaluate LLM agents on cleaning tasks involving six distinct types of noise. Our findings show that LLMs are effective at handling…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsExplainable Artificial Intelligence (XAI) · Topic Modeling · Software System Performance and Reliability