Seg2Act: Global Context-aware Action Generation for Document Logical Structuring

Zichao Li, Shaojie He, Meng Liao, Xuanang Chen, Yaojie Lu, Hongyu Lin, Yanxiong Lu, Xianpei Han, Le Sun


Abstract
Document logical structuring aims to extract the underlying hierarchical structure of documents, which is crucial for document intelligence. Traditional approaches often fall short in handling the complexity and the variability of lengthy documents. To address these issues, we introduce Seg2Act, an end-to-end, generation-based method for document logical structuring, revisiting logical structure extraction as an action generation task. Specifically, given the text segments of a document, Seg2Act iteratively generates the action sequence via a global context-aware generative model, and simultaneously updates its global context and current logical structure based on the generated actions. Experiments on ChCatExt and HierDoc datasets demonstrate the superior performance of Seg2Act in both supervised and transfer learning settings.
Anthology ID:
2024.emnlp-main.1003
Volume:
Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing
Month:
November
Year:
2024
Address:
Miami, Florida, USA
Editors:
Yaser Al-Onaizan, Mohit Bansal, Yun-Nung Chen
Venue:
EMNLP
SIG:
Publisher:
Association for Computational Linguistics
Note:
Pages:
18077–18088
Language:
URL:
https://preview.aclanthology.org/icon-24-ingestion/2024.emnlp-main.1003/
DOI:
10.18653/v1/2024.emnlp-main.1003
Bibkey:
Cite (ACL):
Zichao Li, Shaojie He, Meng Liao, Xuanang Chen, Yaojie Lu, Hongyu Lin, Yanxiong Lu, Xianpei Han, and Le Sun. 2024. Seg2Act: Global Context-aware Action Generation for Document Logical Structuring. In Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing, pages 18077–18088, Miami, Florida, USA. Association for Computational Linguistics.
Cite (Informal):
Seg2Act: Global Context-aware Action Generation for Document Logical Structuring (Li et al., EMNLP 2024)
Copy Citation:
PDF:
https://preview.aclanthology.org/icon-24-ingestion/2024.emnlp-main.1003.pdf
Software:
 2024.emnlp-main.1003.software.zip