bigscience-workshop / bigscience-workshop/biomedical

Create dataset loader for Abbrev Dataset

Open
#249 9 comments 0 reactions 0 assignees View on GitHub
English New Dataset
Dominant language
Python
Stars
505
Forks
117
PR merge metrics
No merged PRs in 30d

Description

## Adding a Dataset
- **Name:** Abbrev Dataset
- **Description:** The Abbrev dataset is made available by Stevenson, et al. (2009). It consists of the acronyms and long-forms from Medline abstracts that were intially prsented by Liu, et al. (2001). The dataset is automatically re-created by identifying the acronyms long froms in Medline and replacing it with it's acronym. The dataset consists of three subsets containing 100, 200 and 300 instances respectively
- **Task:** SPAN_CLASS
- **Paper:** https://aclanthology.org/W09-1309
- **Data:** https://nlp.cs.vcu.edu/data.html
- **License:** ?

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.