apache / apache/gravitino

[FEATURE] Support Fluss catalog

Open
#6,578 11 comments 5 reactions 0 assignees View on GitHub
feature
Dominant language
Java
Stars
3.2k
Forks
935
Avg merge
1d 16h
Merged PRs (30d)
298

Description

### Describe the feature

Support Fluss catalog intergation.

### Motivation

[Fluss](https://github.com/alibaba/fluss) is a streaming storage built for real-time analytics which can serve as the real-time data layer for Lakehouse architectures. It bridges the gap between data streaming and data Lakehouse by enabling low-latency, high-throughput data ingestion and processing while seamlessly integrating with popular compute engines like Apache Flink, while Apache Spark, and StarRocks are coming soon. It's recommended to support streaming storage including Fluss.

### Describe the solution

Introduce streaming storage module, including support Fluss catalog.

### Additional context

_No response_

Contributor guide

Open the contributing guide

Research direction

Start by reviewing the proposed streaming storage module and the Fluss catalog integration described in the issue, along with the linked Fluss project. Determine the project entry points and integration boundaries before implementation. Done means Gravitino supports a Fluss catalog as part of its streaming storage support.

Written by the indexing model from the issue text.

Assessment

Tech stack
java
Domain
data-engineering
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Quiet
Clarity
Needs clarification
Newbie friendliness
35/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.