The Detection of Distributional Discrepancy for Text Generation

Chen, Xingyuan; Cai, Ping; Jin, Peng; Du, Haokun; Wang, Hongjun; Dai, Xingyu; Chen, Jiajun

Computer Science > Computer Vision and Pattern Recognition

arXiv:1910.04859 (cs)

[Submitted on 28 Sep 2019 (v1), last revised 24 Nov 2019 (this version, v2)]

Title:The Detection of Distributional Discrepancy for Text Generation

Authors:Xingyuan Chen, Ping Cai, Peng Jin, Haokun Du, Hongjun Wang, Xingyu Dai, Jiajun Chen

View PDF

Abstract:The text generated by neural language models is not as good as the real text. This means that their distributions are different. Generative Adversarial Nets (GAN) are used to alleviate it. However, some researchers argue that GAN variants do not work at all. When both sample quality (such as Bleu) and sample diversity (such as self-Bleu) are taken into account, the GAN variants even are worse than a well-adjusted language model. But, Bleu and self-Bleu can not precisely measure this distributional discrepancy. In fact, how to measure the distributional discrepancy between real text and generated text is still an open problem. In this paper, we theoretically propose two metric functions to measure the distributional difference between real text and generated text. Besides that, a method is put forward to estimate them. First, we evaluate language model with these two functions and find the difference is huge. Then, we try several methods to use the detected discrepancy signal to improve the generator. However the difference becomes even bigger than before. Experimenting on two existing language GANs, the distributional discrepancy between real text and generated text increases with more adversarial learning rounds. It demonstrates both of these language GANs fail.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1910.04859 [cs.CV]
	(or arXiv:1910.04859v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1910.04859

Submission history

From: Peng Jin [view email]
[v1] Sat, 28 Sep 2019 07:12:34 UTC (6,636 KB)
[v2] Sun, 24 Nov 2019 06:24:04 UTC (6,630 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:The Detection of Distributional Discrepancy for Text Generation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:The Detection of Distributional Discrepancy for Text Generation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators