alibaba / alibaba/feathub

Support VALUE_COUNTS and COLLECT_LIST in SparkProcessor

Open
#232 0 comments 0 reactions 0 assignees View on GitHub
type:feature
Dominant language
Python
Stars
350
Forks
60
PR merge metrics
No merged PRs in 30d

Description

This issue has no description.

Contributor guide

No contributing guide indexed for this repository

Research direction

Locate the SparkProcessor implementation and inspect how existing aggregation functions are represented and tested. Add support for VALUE_COUNTS and COLLECT_LIST, then run the relevant SparkProcessor tests to verify both aggregations.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, spark
Domain
data-engineering
Issue type
Feature
Difficulty
3/5
Estimated time
1-2 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
35/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.