WinoGrande what it measures, how it is scored and who leads it
Reasoning · Allen Institute for AI and University of Washington · introduced Jul 24, 2019 · Accuracy · saturated
About this benchmark
- Task
- Resolve which of two options a pronoun-like blank refers to, using commonsense.
- Dataset
- A large set of Winograd-style problems built with adversarial filtering.
- Method
- Two-option fill-in-the-blank.
- Metric
- Accuracy
- Organization
- Allen Institute for AI and University of Washington
- Introduced
- Jul 24, 2019
Versions
Newest first. New versions are added, never rewritten.
| Version | Date | |
|---|---|---|
| 1Released with the paper. | Jul 24, 20197 years ago | Released with the paper. |
Sources: each benchmark's paper (arXiv) and its official site or leaderboard. Scores are the boards captured here every day, exactly as published.