apache / apache/arrow-julia

Do some advanced benchmarking on arrow data + common operations

Open
#188 10 comments 0 reactions 0 assignees View on GitHub
Dominant language
Julia
Stars
312
Forks
78
PR merge metrics
No merged PRs in 30d

Description

Chatted briefly with @bkamins on this; we'd like to take some common data processing workflows and run benchmarks comparing arrow data vs. regular in-memory Julia data.

Contributor guide

No contributing guide indexed for this repository

Research direction

No files, tests, or entry points are named. Review the issue and its discussion with @bkamins to define the common data-processing workflows and benchmark setup, then compare Arrow data with regular in-memory Julia data and document the results.

Written by the indexing model from the issue text.

Assessment

Tech stack
julia
Domain
data, performance
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.