SparkClient: Examples using Kubeflow pipeline
Open
area/spark
- Dominant language
- Python
- Stars
- 148
- Forks
- 262
- Avg merge
- 1d 2h
- Merged PRs (30d)
- 1
Description
We would like to make sure SparkClient works seemlessly in kubeflow ecosystem and hence running spark jobs in kubeflow pipeline and validate it.
Contributor guide
Research direction
Start by locating the SparkClient entry point and any existing Kubeflow pipeline examples in the repository. Run a Spark job through a Kubeflow pipeline and document or add an example that validates SparkClient works in that environment.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- kubernetes, python, spark
- Domain
- distributed-systems, machine-learning
- Issue type
- Documentation
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Quiet
- Clarity
- Needs clarification
- Newbie friendliness
- 45/100