apache / apache/paimon

[Feature] Support for Kafka CDC parsing `debezium-avro-confluent` format data

Open
#1,979 0 comments 0 reactions 0 assignees View on GitHub
enhancement
Dominant language
Java
Stars
3.4k
Forks
1.4k
Avg merge
1d 11h
Merged PRs (30d)
396

Description

### Search before asking

- [X] I searched in the [issues](https://github.com/apache/incubator-paimon/issues) and found nothing similar.

### Motivation

In our production environment, Kafka is mostly used to store data in Avro format, in conjunction with the `Confluentinc/Schema-Registry`.

### Solution

_No response_

### Anything else?

_No response_

### Are you willing to submit a PR?

- [X] I'm willing to submit a PR!

Contributor guide

No contributing guide indexed for this repository

Research direction

The issue does not name any files, tests, or entry points, and provides no implementation details beyond Kafka CDC data in the debezium-avro-confluent format. Start by locating the existing Kafka CDC parsing paths and format handlers; done should mean that this format is parsed successfully with coverage for its expected data.

Written by the indexing model from the issue text.

Assessment

Tech stack
java, kafka
Domain
data-engineering, stream-processing
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.