SepFormer: Coarse-to-fine Separator Regression Network for Table Structure Recognition

Nam Quan Nguyen; Xuan Phong Pham; Tuan-Anh Tran

arXiv:2506.21920·cs.CV·June 30, 2025

SepFormer: Coarse-to-fine Separator Regression Network for Table Structure Recognition

Nam Quan Nguyen, Xuan Phong Pham, Tuan-Anh Tran

PDF

TL;DR

SepFormer is a novel transformer-based model that efficiently recognizes table structures from images by predicting separators in a coarse-to-fine manner, achieving high speed and competitive accuracy.

Contribution

It introduces a unified, coarse-to-fine separator regression approach with a DETR-style architecture for improved table structure recognition.

Findings

01

Runs at 25.6 FPS on average

02

Achieves comparable performance with state-of-the-art methods

03

Effective on multiple benchmark datasets

Abstract

The automated reconstruction of the logical arrangement of tables from image data, termed Table Structure Recognition (TSR), is fundamental for semantic data extraction. Recently, researchers have explored a wide range of techniques to tackle this problem, demonstrating significant progress. Each table is a set of vertical and horizontal separators. Following this realization, we present SepFormer, which integrates the split-and-merge paradigm into a single step through separator regression with a DETR-style architecture, improving speed and robustness. SepFormer is a coarse-to-fine approach that predicts table separators from single-line to line-strip separators with a stack of two transformer decoders. In the coarse-grained stage, the model learns to gradually refine single-line segments through decoder layers with additional angle loss. At the end of the fine-grained stage, the model…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.