AbsaOSS / AbsaOSS/spline-spark-agent

Add Databricks environment details to execution plan

Open
#409 10 comments 0 reactions 0 assignees View on GitHub
dependency: Databricks feature
Dominant language
Scala
Stars
210
Forks
102
Avg merge
1d 1h
Merged PRs (30d)
1

Description

@wajda @cerveada , after taking some inspiration from Openlineage installtion steps, I was able to do codeless spline installation on Databricks (please see attached document for detailed steps). What I am missing now is to get some additional details from Databricks cluster like notebook name, user etc.. is it possible to build this functionality in spline jar , so we don't need run extra code snippet?

This has been implemented in [openlineage ](https://github.com/OpenLineage/OpenLineage/blob/main/integration/spark/src/main/common/java/io/openlineage/spark/agent/facets/builder/DatabricksEnvironmentFacetBuilder.java) using a databricks environment facet builder. I hope this can be done in spline as well.

[Install_Spline_Codeless.docx](https://github.com/AbsaOSS/spline-spark-agent/files/8161698/Install_Spline_Codeless.docx)

[Shell_Scripts.zip](https://github.com/AbsaOSS/spline-spark-agent/files/8161717/Shell_Scripts.zip)

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.