Skip to main content
Datasetbenchmarksv1.3

SuperGLUE

by New York University · open-source · Last verified 2026-03-17

SuperGLUE is a benchmark suite of 8 challenging NLU tasks including question answering, coreference resolution, causal reasoning, and word-sense disambiguation, designed as a harder successor to GLUE. It includes human baselines and has driven significant progress in pre-trained language model capabilities.

https://super.gluebenchmark.com
B+
B+Good
Adoption: AQuality: A+Freshness: BCitations: A+Engagement: F

Specifications

License
Various (task-dependent)
Pricing
open-source
Capabilities
nlu-evaluation, multi-task-benchmark, language-understanding
Integrations
huggingface-datasets, lm-eval-harness
Use Cases
model-evaluation, nlu-benchmarking, pre-training-assessment
API Available
No
Tags
benchmark, nlp-benchmark, natural-language-understanding, multi-task, glue-successor
Added
2026-03-17
Completeness
100%

Index Score

74.5
Adoption
82
Quality
91
Freshness
65
Citations
94
Engagement
0

Put AI to work for your business

Deploy this dataset alongside autonomous AaaS agents that handle tasks end-to-end — no babysitting required.

Explore the full AI ecosystem on Agents as a Service