Cell-aware Stacked LSTMs for Modeling Sentences

Choi, Jihun; Kim, Taeuk; Lee, Sang-goo

Computer Science > Computation and Language

arXiv:1809.02279 (cs)

[Submitted on 7 Sep 2018 (v1), last revised 1 Nov 2019 (this version, v2)]

Title:Cell-aware Stacked LSTMs for Modeling Sentences

Authors:Jihun Choi, Taeuk Kim, Sang-goo Lee

View PDF

Abstract:We propose a method of stacking multiple long short-term memory (LSTM) layers for modeling sentences. In contrast to the conventional stacked LSTMs where only hidden states are fed as input to the next layer, the suggested architecture accepts both hidden and memory cell states of the preceding layer and fuses information from the left and the lower context using the soft gating mechanism of LSTMs. Thus the architecture modulates the amount of information to be delivered not only in horizontal recurrence but also in vertical connections, from which useful features extracted from lower layers are effectively conveyed to upper layers. We dub this architecture Cell-aware Stacked LSTM (CAS-LSTM) and show from experiments that our models bring significant performance gain over the standard LSTMs on benchmark datasets for natural language inference, paraphrase detection, sentiment classification, and machine translation. We also conduct extensive qualitative analysis to understand the internal behavior of the suggested approach.

Comments:	ACML 2019
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:1809.02279 [cs.CL]
	(or arXiv:1809.02279v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.1809.02279

Submission history

From: Jihun Choi [view email]
[v1] Fri, 7 Sep 2018 02:17:23 UTC (319 KB)
[v2] Fri, 1 Nov 2019 07:23:42 UTC (319 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CL

< prev | next >

new | recent | 2018-09

Change to browse by:

References & Citations

DBLP - CS Bibliography

listing | bibtex

Jihun Choi
Taeuk Kim
Sang-goo Lee

export BibTeX citation

Computer Science > Computation and Language

Title:Cell-aware Stacked LSTMs for Modeling Sentences

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Cell-aware Stacked LSTMs for Modeling Sentences

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators