mapstruct / mapstruct/mapstruct-examples

Convert Spark Dataframe to Dataset with MapStruct

未关闭
#128 0 条评论 1 个 reaction 已指派 0 人 在 GitHub 查看

还没有人认领这个 Issue。

主要语言
Java
星标
1.4k
派生
504
PR 合并指标
30 天内没有已合并 PR

描述

When using Apache Spark with Java there is a pretty common use case of converting Spark's Dataframes to POJO-based Datasets. The thing is that many times your Dataframe is imported from a database in which the column namings and types are different from your POJO.

Example for this can be found on the following Stackoverflow question (which has been solved programmatically with a map and rename functions). I was wondering if that example could be solved with MapStruct (and maybe would be added to this examples repo).

贡献指南

这个仓库没有索引到贡献指南

从这里开始

  1. 先读完整个 Issue,再读项目的贡献指南。
  2. 在 Issue 下留言说明你要接手 —— 这能避免两个人做同样的事。
  3. Fork 仓库,在一个分支上完成修改。
  4. 提交 Pull Request,并在描述里引用这个 Issue 编号。

调研方向

首先查看 mapstruct-examples 仓库和链接的 Stack Overflow 问题,以了解所请求的 Spark DataFrame 到 Dataset 的转换。确定适合放置 Spark 和 Java 示例的位置,然后确定 MapStruct 是否能够处理不同的数据库列名和 POJO 字段类型。当有一个涵盖该转换的、文档齐全且可运行的示例时,即视为完成。

由索引模型根据 Issue 内容生成。

评估

技术栈
java, spark
领域
data-engineering
Issue 类型
功能
难度
4/5
预计耗时
3-5 天
活跃度
停滞
描述清晰度
基本清楚
新手友好度
25/100

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。