INRIA / INRIA/scikit-learn-mooc

Add a subsection in predictive modeling pipeline regarding missing values handling

Open
#300 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
1.4k
Forks
600
Avg merge
6d 20h
Merged PRs (30d)
2

Description

We should add a small section to show how:

- how to handle missing data
- illustrate the usage of an imputer in a pipeline
- illustrate the usage of a pipeline within a column transformer

+ exercise

Contributor guide

Open the contributing guide

Research direction

Locate the predictive modeling pipeline material in the course notebooks and read the surrounding lesson structure and exercises first. Add a subsection covering missing-data handling, an imputer in a pipeline, and a pipeline inside a column transformer, with an exercise that demonstrates these concepts.

Written by the indexing model from the issue text.

Assessment

Tech stack
jupyter-notebook, python
Domain
content, documentation, machine-learning
Issue type
Documentation
Difficulty
3/5
Estimated time
1-2 days
Activity status
Stale
Clarity
Mostly clear
Newbie friendliness
42/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.