BONSAMURAIS / BONSAMURAIS/reproducibility

Attain competency in `DataPackages`

Open
#6 3 comments 2 reactions 0 assignees View on GitHub
During Hackathon
Dominant language
No language data
Stars
0
Forks
0
PR merge metrics
No merged PRs in 30d

Description

`DataPackages` from [frictionless data](https://frictionlessdata.io/docs/using-data-packages-in-python/) are the preferred format for sharing structured data generated in the hackathon. One role of the reproducibility group will be to ensure that the work products of the other groups can be easily packaged in this manner for distribution to other users.

The first task in this objective is to become familiar with how datapackages work, determine whether we need to create a custom schema or use the existing ones, and understand clearly how to package the work output from another group into this format.

This issue will be closed when the first output object from another working group is rendered as a `DataPackage` using an automated workflow.

Contributor guide

No contributing guide indexed for this repository

Research direction

Start with the linked Frictionless Data documentation for using Data Packages in Python and determine whether an existing schema fits the hackathon outputs. Then identify an output object from another working group and define the automated workflow needed to render it as a DataPackage. Done means that one such output is packaged automatically.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
data-engineering
Issue type
Feature
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.