HiPool: Modeling Long Documents Using Graph Neural Networks

Li, Irene; Feng, Aosong; Radev, Dragomir; Ying, Rex

Computer Science > Computation and Language

arXiv:2305.03319 (cs)

[Submitted on 5 May 2023 (v1), last revised 15 May 2023 (this version, v2)]

Title:HiPool: Modeling Long Documents Using Graph Neural Networks

Authors:Irene Li, Aosong Feng, Dragomir Radev, Rex Ying

View PDF

Abstract:Encoding long sequences in Natural Language Processing (NLP) is a challenging problem. Though recent pretraining language models achieve satisfying performances in many NLP tasks, they are still restricted by a pre-defined maximum length, making them challenging to be extended to longer sequences. So some recent works utilize hierarchies to model long sequences. However, most of them apply sequential models for upper hierarchies, suffering from long dependency issues. In this paper, we alleviate these issues through a graph-based method. We first chunk the sequence with a fixed length to model the sentence-level information. We then leverage graphs to model intra- and cross-sentence correlations with a new attention mechanism. Additionally, due to limited standard benchmarks for long document classification (LDC), we propose a new challenging benchmark, totaling six datasets with up to 53k samples and 4034 average tokens' length. Evaluation shows our model surpasses competitive baselines by 2.6% in F1 score, and 4.8% on the longest sequence dataset. Our method is shown to outperform hierarchical sequential models with better performance and scalability, especially for longer sequences.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2305.03319 [cs.CL]
	(or arXiv:2305.03319v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2305.03319
Journal reference:	ACL 2023 main proceedings

Submission history

From: Irene Li [view email]
[v1] Fri, 5 May 2023 06:58:24 UTC (233 KB)
[v2] Mon, 15 May 2023 03:48:36 UTC (4,057 KB)

Computer Science > Computation and Language

Title:HiPool: Modeling Long Documents Using Graph Neural Networks

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:HiPool: Modeling Long Documents Using Graph Neural Networks

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators