Self-training Large Language Models through Knowledge Detection

Wei Jie Yeo; Teddy Ferdinan; Przemyslaw Kazienko; Ranjan Satapathy,; Erik Cambria

arXiv:2406.11275·cs.CL·November 13, 2024·1 cites

Self-training Large Language Models through Knowledge Detection

Wei Jie Yeo, Teddy Ferdinan, Przemyslaw Kazienko, Ranjan Satapathy,, Erik Cambria

PDF

Open Access 1 Repo

TL;DR

This paper introduces a self-training method for large language models that autonomously labels data and improves generation accuracy, reducing hallucinations and catastrophic forgetting without relying heavily on labeled datasets.

Contribution

It presents a novel reference-free consistency-based self-training framework that enhances LLM performance and robustness in a scalable, cost-effective manner.

Findings

01

Reduces hallucination in generated outputs

02

Mitigates catastrophic forgetting in OOD benchmarks

03

Decreases dependence on labeled datasets

Abstract

Large language models (LLMs) often necessitate extensive labeled datasets and training compute to achieve impressive performance across downstream tasks. This paper explores a self-training paradigm, where the LLM autonomously curates its own labels and selectively trains on unknown data samples identified through a reference-free consistency method. Empirical evaluations demonstrate significant improvements in reducing hallucination in generation across multiple subjects. Furthermore, the selective training framework mitigates catastrophic forgetting in out-of-distribution benchmarks, addressing a critical limitation in training LLMs. Our findings suggest that such an approach can substantially reduce the dependency on large labeled datasets, paving the way for more scalable and cost-effective language model training.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

wj210/Self-Training-LLM
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsNatural Language Processing Techniques