GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding

Explore this paper's citation graph

Summary

A benchmark of nine diverse NLU tasks, an auxiliary dataset for probing models for understanding of specific linguistic phenomena, and an online platform for evaluating and comparing models, which favors models that can represent linguistic knowledge in a way that facilitates sample-efficient learning and effective knowledge-transfer across tasks.

Type
preprint
Published
2018-04-20
Cited by
8,891
References
77
Access
Open access

Keywords

Computer science, Natural language understanding, Benchmark (surveying), Task (project management), Suite

References

Cited by

Related papers