bigscience-workshop / bigscience-workshop/biomedical

Create dataset loader for CafeteriaSA

Open
#868 0 comments 0 reactions 0 assignees View on GitHub
New Dataset
Dominant language
Python
Stars
505
Forks
117
PR merge metrics
No merged PRs in 30d

Description

## Adding a Dataset
- **Name:** *CafeteriaSA*
- **Description:** *Annotated corpus of 500 scientific abstracts from PubMed that consists of 6407 annotated food entities*
- **Task:** *NER,NED*
- **Paper:** *[https://doi.org/10.1093/database/baac107](https://doi.org/10.1093/database/baac107)*
- **Data:** *[Zenodo-Link](https://zenodo.org/record/6683798#.Y9fHx-zML0q)*
- **License:** *CC-4.0-International*
- **Motivation:** *Interesting entity type (food entries) for which only few corpora exist*

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.