LegalSearchLM: Rethinking Legal Case Retrieval as Legal Elements Generation

Chaeeun Kim; Jinu Lee; Wonseok Hwang

arXiv:2505.23832·cs.CL·October 7, 2025

LegalSearchLM: Rethinking Legal Case Retrieval as Legal Elements Generation

Chaeeun Kim, Jinu Lee, Wonseok Hwang

PDF

Open Access 1 Video

TL;DR

LegalSearchLM introduces a novel legal case retrieval approach that generates legal elements directly, supported by a large-scale Korean benchmark, significantly improving retrieval accuracy and generalization over existing methods.

Contribution

The paper presents LEGAR BENCH, a large-scale Korean legal case retrieval benchmark, and LegalSearchLM, a generative model that improves legal case retrieval by reasoning over legal elements.

Findings

01

LegalSearchLM outperforms baselines by 6-20% on LEGAR BENCH.

02

LegalSearchLM demonstrates 15% better generalization to out-of-domain cases.

03

LEGAR BENCH covers 411 crime types with 1.2M cases.

Abstract

Legal Case Retrieval (LCR), which retrieves relevant cases from a query case, is a fundamental task for legal professionals in research and decision-making. However, existing studies on LCR face two major limitations. First, they are evaluated on relatively small-scale retrieval corpora (e.g., 100-55K cases) and use a narrow range of criminal query types, which cannot sufficiently reflect the complexity of real-world legal retrieval scenarios. Second, their reliance on embedding-based or lexical matching methods often results in limited representations and legally irrelevant matches. To address these issues, we present: (1) LEGAR BENCH, the first large-scale Korean LCR benchmark, covering 411 diverse crime types in queries over 1.2M candidate cases; and (2) LegalSearchLM, a retrieval model that performs legal element reasoning over the query case and directly generates content…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

LegalSearchLM: Rethinking Legal Case Retrieval as Legal Elements Generation· underline

Taxonomy

TopicsArtificial Intelligence in Law · Legal Education and Practice Innovations · Comparative and International Law Studies