Reliable Text-to-SQL with Adaptive Abstention

Kaiwen Chen; Yueting Chen; Xiaohui Yu; Nick Koudas

arXiv:2501.10858·cs.DB·January 22, 2025

Reliable Text-to-SQL with Adaptive Abstention

Kaiwen Chen, Yueting Chen, Xiaohui Yu, Nick Koudas

PDF

Open Access

TL;DR

This paper introduces RTS, a framework that improves the reliability of text-to-SQL systems by detecting errors, abstaining, and involving humans, especially focusing on schema linking with probabilistic guarantees.

Contribution

RTS is the first to incorporate adaptive abstention and human-in-the-loop mechanisms with probabilistic schema linking guarantees in text-to-SQL models.

Findings

01

Achieves near-perfect schema linking accuracy on BIRD benchmark.

02

Significantly improves robustness and reliability of text-to-SQL conversion.

03

Small models with RTS nearly match state-of-the-art larger models.

Abstract

Large language models (LLMs) have revolutionized natural language interfaces for databases, particularly in text-to-SQL conversion. However, current approaches often generate unreliable outputs when faced with ambiguity or insufficient context. We present Reliable Text-to-SQL (RTS), a novel framework that enhances query generation reliability by incorporating abstention and human-in-the-loop mechanisms. RTS focuses on the critical schema linking phase, which aims to identify the key database elements needed for generating SQL queries. It autonomously detects potential errors during the answer generation process and responds by either abstaining or engaging in user interaction. A vital component of RTS is the Branching Point Prediction (BPP) which utilizes statistical conformal techniques on the hidden layers of the LLM model for schema linking, providing probabilistic guarantees on…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsDistributed and Parallel Computing Systems