Can Perplexity Reflect Large Language Model's Ability in Long Text Understanding?

Hu, Yutong; Huang, Quzhe; Tao, Mingxu; Zhang, Chen; Feng, Yansong

Computer Science > Computation and Language

arXiv:2405.06105 (cs)

[Submitted on 9 May 2024]

Title:Can Perplexity Reflect Large Language Model's Ability in Long Text Understanding?

Authors:Yutong Hu, Quzhe Huang, Mingxu Tao, Chen Zhang, Yansong Feng

View PDF HTML (experimental)

Abstract:Recent studies have shown that Large Language Models (LLMs) have the potential to process extremely long text. Many works only evaluate LLMs' long-text processing ability on the language modeling task, with perplexity (PPL) as the evaluation metric. However, in our study, we find that there is no correlation between PPL and LLMs' long-text understanding ability. Besides, PPL may only reflect the model's ability to model local information instead of catching long-range dependency. Therefore, only using PPL to prove the model could process long text is inappropriate. The local focus feature of PPL could also explain some existing phenomena, such as the great extrapolation ability of the position method ALiBi. When evaluating a model's ability in long text, we might pay more attention to PPL's limitation and avoid overly relying on it.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2405.06105 [cs.CL]
	(or arXiv:2405.06105v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2405.06105

Submission history

From: Yutong Hu [view email]
[v1] Thu, 9 May 2024 21:15:49 UTC (40 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CL

< prev | next >

new | recent | 2024-05

Change to browse by:

References & Citations

export BibTeX citation

Computer Science > Computation and Language

Title:Can Perplexity Reflect Large Language Model's Ability in Long Text Understanding?

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Can Perplexity Reflect Large Language Model's Ability in Long Text Understanding?

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators