bigscience-workshop / bigscience-workshop/evaluation

Add MNLI to Full Benchmark

Open
#33 3 comments 0 reactions 0 assignees View on GitHub
few_shot
Dominant language
Python
Stars
42
Forks
24
PR merge metrics
No merged PRs in 30d

Description

coordinate with whoever is working on SuperGLUE, we only need to include MNLI once. But NLI will be held-out from model training (whereas the other SuperGLUE tasks will not) so interpreting MNLI results is different from other superglue tasks.

use to test generalization to unseen task; maybe use FLEX?

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.