Structural Guidance for Transformer Language Models

Qian, Peng; Naseem, Tahira; Levy, Roger; Astudillo, Ramón Fernandez

Computer Science > Computation and Language

arXiv:2108.00104 (cs)

[Submitted on 30 Jul 2021]

Title:Structural Guidance for Transformer Language Models

Authors:Peng Qian, Tahira Naseem, Roger Levy, Ramón Fernandez Astudillo

View PDF

Abstract:Transformer-based language models pre-trained on large amounts of text data have proven remarkably successful in learning generic transferable linguistic representations. Here we study whether structural guidance leads to more human-like systematic linguistic generalization in Transformer language models without resorting to pre-training on very large amounts of data. We explore two general ideas. The "Generative Parsing" idea jointly models the incremental parse and word sequence as part of the same sequence modeling task. The "Structural Scaffold" idea guides the language model's representation via additional structure loss that separately predicts the incremental constituency parse. We train the proposed models along with a vanilla Transformer language model baseline on a 14 million-token and a 46 million-token subset of the BLLIP dataset, and evaluate models' syntactic generalization performances on SG Test Suites and sized BLiMP. Experiment results across two benchmarks suggest converging evidence that generative structural supervisions can induce more robust and humanlike linguistic generalization in Transformer language models without the need for data intensive pre-training.

Comments:	To be issued as paper revision for ACL 2021
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2108.00104 [cs.CL]
	(or arXiv:2108.00104v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2108.00104

Submission history

From: Peng Qian [view email]
[v1] Fri, 30 Jul 2021 23:14:51 UTC (251 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CL

< prev | next >

new | recent | 2021-08

Change to browse by:

References & Citations

DBLP - CS Bibliography

listing | bibtex

Peng Qian
Tahira Naseem
Roger Levy
Ramón Fernandez Astudillo

export BibTeX citation

Computer Science > Computation and Language

Title:Structural Guidance for Transformer Language Models

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Structural Guidance for Transformer Language Models

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators