baidu / baidu/DuReader

what should I do if I want to use my data?

Open
#50 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
1.2k
Forks
306
PR merge metrics
No merged PRs in 30d

Description

I want to know what the various keys of the json data set represent. For example, ‘is_selected’, ‘answer_spans’, and ‘match_scores’. And I see that there are no such keys in the raw data.

Contributor guide

No contributing guide indexed for this repository

Research direction

The issue mentions the processed JSON dataset and the raw data, but no filenames or entry points. First compare the available JSON keys with the raw dataset and identify where the dataset schema is documented; done means the meanings of is_selected, answer_spans, and match_scores, plus the difference from raw data, are documented.

Written by the indexing model from the issue text.

Assessment

Domain
data, documentation
Issue type
Documentation
Difficulty
3/5
Estimated time
1-2 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.