tensorflow / tensorflow/datasets

[data request] IBM Diversity in Faces Dataset (DiF)

Open
#299 27 comments 34 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

dataset request
Dominant language
Python
Stars
4.6k
Forks
1.6k
Avg merge
3h 54m
Merged PRs (30d)
1

Description

  • Name of dataset: IBM Diversity in Faces (DiF)

  • URL of dataset: https://www.research.ibm.com/artificial-intelligence/trusted-ai/diversity-in-faces/

  • License of dataset: Terms of Use

  • Short description of dataset and use case(s): "The Diversity in Faces (DiF) is a large and diverse dataset that seeks to advance the study of fairness and accuracy in facial recognition technology. The first of its kind available to the global research community,DiF provides a dataset of annotations of 1 million human facial images."

Folks who would also like to see this dataset in tensorflow/datasets, please thumbs-up so the developers can know which requests to prioritize.

Contributor guide

Open the contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

Start with the IBM Diversity in Faces dataset URL and its linked Terms of Use, then review tensorflow/datasets to identify the entry point for adding a dataset. Done would require confirming that the terms permit inclusion and integrating the dataset into TFDS; this issue names no files or tests.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, tensorflow
Domain
data, machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.