apache / apache/beam

[Task]: Align Spark/Hadoop versions in Spark multiple-versions unit/VR tests

Open
#27,562 3 comments 0 reactions 0 assignees View on GitHub
awaiting triage java P3 spark stale task
Dominant language
Java
Stars
8.7k
Forks
4.7k
Avg merge
1d 20h
Merged PRs (30d)
196

Description

### What needs to happen?

Currently, we run Spark runner tests against multiple Spark/Hadoop versions. We need to make sure that these versions are properly aligned and officially compatible among them. Also, we need to make sure that only the tested version of Hadoop/Spark classes is sitting on classpath.

### Issue Priority

Priority: 3 (nice-to-have improvement)

### Issue Components

- [ ] Component: Python SDK
- [X] Component: Java SDK
- [ ] Component: Go SDK
- [ ] Component: Typescript SDK
- [ ] Component: IO connector
- [ ] Component: Beam examples
- [ ] Component: Beam playground
- [ ] Component: Beam katas
- [ ] Component: Website
- [X] Component: Spark Runner
- [ ] Component: Flink Runner
- [ ] Component: Samza Runner
- [ ] Component: Twister2 Runner
- [ ] Component: Hazelcast Jet Runner
- [ ] Component: Google Cloud Dataflow Runner

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.