bigscience-workshop / bigscience-workshop/lam

Add dataset: multilingual_named_entity_recognition_for_medieval_charters

Open
#81 0 comments 0 reactions 0 assignees View on GitHub
dataset
Dominant language
No language data
Stars
91
Forks
8
PR merge metrics
No merged PRs in 30d

Description

### A URL for this dataset

https://doi.org/10.5281/zenodo.6463699

### Dataset description

Annotated dataset for training named entities recognition models for medieval charters in Latin, French and Spanish.

### Dataset modality

Text

### Dataset licence

Creative Commons Attribution 4.0 International

### Other licence

_No response_

### How can you access this data

As a download from a repository/website

### size of dataset

>10GB

### Confirm the dataset has an open licence

- [X] To the best of my knowledge, this dataset is accessible via an open licence

### Contact details for data custodian

_No response_

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.