A Scalable Data-Driven Framework for Systematic Analysis of SEC 10-K   Filings Using Large Language Models

Syed Affan Daimi; Asma Iqbal

arXiv:2409.17581·cs.AI·September 27, 2024

A Scalable Data-Driven Framework for Systematic Analysis of SEC 10-K Filings Using Large Language Models

Syed Affan Daimi, Asma Iqbal

PDF

Open Access 1 Repo

TL;DR

This paper presents a scalable, automated framework utilizing large language models to systematically analyze SEC 10-K filings, providing quantitative performance ratings and visual insights for a large number of companies efficiently.

Contribution

It introduces a novel data-driven system that automates extraction, analysis, and visualization of 10-K filings using LLMs, enabling comprehensive and efficient corporate performance assessment.

Findings

01

Effective extraction and segmentation of 10-K sections.

02

Generation of quantitative performance ratings.

03

Interactive GUI for visualization and comparison.

Abstract

The number of companies listed on the NYSE has been growing exponentially, creating a significant challenge for market analysts, traders, and stockholders who must monitor and assess the performance and strategic shifts of a large number of companies regularly. There is an increasing need for a fast, cost-effective, and comprehensive method to evaluate the performance and detect and compare many companies' strategy changes efficiently. We propose a novel data-driven approach that leverages large language models (LLMs) to systematically analyze and rate the performance of companies based on their SEC 10-K filings. These filings, which provide detailed annual reports on a company's financial performance and strategic direction, serve as a rich source of data for evaluating various aspects of corporate health, including confidence, environmental sustainability, innovation, and workforce…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

sulphatet/10-K-Filings-Rating-System
noneOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsMathematics, Computing, and Information Processing