AbsaOSS / AbsaOSS/enceladus

Further partitioning of Publish data based on other fields in the Dataset

Open
#1,704 0 comments 1 reaction 0 assignees View on GitHub
Epic priority: undecided under discussion
Dominant language
Scala
Stars
33
Forks
16
PR merge metrics
No merged PRs in 30d

Description

## Background
It might be useful to further partition the data after the conformance phase, based on the fields of the dataset. Done per user setup.

## Goal
Enable sub-partitoning of data at the end of the Conformance phase based on dataset configuration

## Expected Task List
A list of expected issues that will be needed to achieve this Epic
1. New optional set-up for dataset specifying the sub-partioning based on the dataset fields
2. The actual partitioning in _SparkJobs_
3. According UI changes to allow the set up.

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.