daisybio / daisybio/drevalpy

Make docs better for new data

Open
#449 0 comments 0 reactions 0 assignees View on GitHub
enhancement
Dominant language
Python
Stars
44
Forks
12
Avg merge
7d 19h
Merged PRs (30d)
10

Description

It's not documented well how to run a model with your own dataset, especially the part about the drug input. E.g., it doesn't state clearly how to generate the drug features and where the fingerprints should be located and how the file should look like (drug ids = column names). It also doesn't state clearly that you should adapt the meta/landmark_genes_reduced.csv, especially for data coming from a different species.

Contributor guide

Open the contributing guide

Research direction

Start by locating the documentation for running a model with a custom dataset and review how drug features, fingerprints, and drug IDs are described. Clarify how to generate the features, where fingerprints belong and how their file should be structured, and explain when to adapt meta/landmark_genes_reduced.csv, including for data from another species.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
data, documentation, machine-learning
Issue type
Documentation
Difficulty
3/5
Estimated time
1-2 days
Activity status
Quiet
Clarity
Mostly clear
Newbie friendliness
55/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.