bigscience-workshop / bigscience-workshop/biomedical

Proposal to add the CLEAR dataset

Open
#360 0 comments 0 reactions 0 assignees View on GitHub
French New Dataset
Dominant language
Python
Stars
505
Forks
117
PR merge metrics
No merged PRs in 30d

Description

## Adding a Dataset
- **Name:** *CLEAR*
- **Description:** *The dataset contains three corpora of documents with comparable contents. Each corpus provides technical and simple/simplified texts on a given topic in French.*
- **Task:** *Text Pairs (Text Simplification)*
- **Paper:** *https://aclanthology.org/W18-7002/*
- **Data:** *http://natalia.grabar.free.fr/resources.php*
- **License:** *Unknown*
- **Motivation:** *The largest Text Simplification corpus in French*

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.