bigscience-workshop / bigscience-workshop/evaluation
Add MNLI to Full Benchmark
Open
few_shot
- Dominant language
- Python
- Stars
- 42
- Forks
- 24
- PR merge metrics
- No merged PRs in 30d
Description
coordinate with whoever is working on SuperGLUE, we only need to include MNLI once. But NLI will be held-out from model training (whereas the other SuperGLUE tasks will not) so interpreting MNLI results is different from other superglue tasks.
use to test generalization to unseen task; maybe use FLEX?
Contributor guide
Assessment
This issue has not been assessed yet.