AbsaOSS / AbsaOSS/enceladus

Performance monitoring and parameter tuning for Enceladus runs

Open
#736 0 comments 0 reactions 1 assignee Claimed by @DzMakatun View on GitHub
Epic priority: medium
Dominant language
Scala
Stars
33
Forks
16
PR merge metrics
No merged PRs in 30d

Description

Spark jobs take the following parameters: number of executors, memory/cores per executor, driver memory/cores. The performance of spark jobs can be evaluated using various metrics: e.g. job duration, CPU time, resource usage etc.

- Define relevant performance measures for Enceladus jobs
- Investigate influence of job parameters on the performance
- Investigate other relevant factors: input format/size/schema, conformance rules number/types, cluster configuration/load, etc.
- Propose optimization methods for setting the job parameters

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.