Call-for-Code-for-Racial-Justice / Call-for-Code-for-Racial-Justice/TakeTwo-DataScience

Implement Machine Learning component V3 (dsmvp-v3)

未关闭
#10 1 条评论 0 个 reaction 已指派 0 人 在 GitHub 查看
Machine Learning
主要语言
Jupyter Notebook
星标
8
派生
8
PR 合并指标
30 天内没有已合并 PR

描述

As part of the progression of machine learning components with increasing levels of sophistication, implement version 3 ("dsmvp-v3") with the following characteristics:

Explainable Model: A machine learning model that can learn to detect racially biased expressions in context based on input labeled data without explicit division of "expression" and "context”, i.e. labeled data consisting of pairs, and the trained model is to output sub-expression(s) of a new test input text identified to be biased expressions in context.

This may need to make use of an AIX (Explainable AI model/method) on text data, which can learn to classify an entire text, and at the same time, point to portions of the text that are likely most responsible for the classification judgement.
This may have to be invented, or further literature search may be required.

At minimum, a method akin to those AIX methods targeting tabular data (e.g. contrastive explanation method in AIX 360) can be applied with relatively straightforward modifications. (Reference: https://arxiv.org/abs/1802.07623)

Coding of dsmvp-v3 should be similar to and share many aspects of how dsmvp-v3 in the repository is implemented, using Jupyter notebook and accessing the database via taketwo-webapi, etc.

贡献指南

打开贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。