bigscience-workshop / bigscience-workshop/data_tooling
Create dataset global_voices_spanish
- Dominant language
- HTML
- Stars
- 91
- Forks
- 47
- PR merge metrics
- No merged PRs in 30d
Description
- uid: global_voices_spanish
- type: primary
- description:
- name: Global Voices Spanish
- description: Global Voices pages in Spanish
- homepage: https://es.globalvoices.org/
- validated: True
- languages:
- language_names:
- Spanish
- language_comments:
- language_locations:
- Americas
- Europe
- validated: False
- custodian:
- name:
- in_catalogue: global_voices
- type:
- location:
- contact_name:
- contact_email:
- contact_submitter: False
- additional:
- validated: False
- availability:
- procurement:
- for_download: Yes - it has a direct download link or links
- download_url: https://es.globalvoices.org/
- download_email:
- licensing:
- has_licenses: Yes
- license_text: https://globalvoices.org/about/global-voices-attribution-policy/
- license_properties:
- open license
- license_list:
- cc-by-3.0: Creative Commons Attribution 3.0 Unported
- pii:
- has_pii: Yes
- generic_pii_likely: very likely
- generic_pii_list:
- names
- email addresses
- numeric_pii_likely: somewhat likely
- numeric_pii_list:
- telephone numbers
- sensitive_pii_likely: very likely
- sensitive_pii_list:
- racial or ethnic origin
- political opinions
- religious or philosophical beliefs
- no_pii_justification_class:
- no_pii_justification_text:
- validated: False
- source_category:
- category_type: website
- category_web: news or magazine website
- category_media:
- validated: False
- media:
- category:
- text
- text_format:
- .HTML
- audiovisual_format:
- image_format:
- database_format:
- text_is_transcribed: No
- instance_type: article
- instance_count: 10K
Contributor guide
No contributing guide indexed for this repository
Research direction
Start with the supplied dataset record and add it as global_voices_spanish.json in the repository’s dataset catalog. Done means the file is present with the listed metadata and matches the catalog’s expected structure.
Written by the indexing model from the issue text.
Assessment
- Domain
- data
- Issue type
- Feature
- Difficulty
- 1/5
- Estimated time
- Under an hour
- Activity status
- Stale
- Clarity
- Clearly specified
- Newbie friendliness
- 45/100