pydata / pydata/sparse

Tracking issue for benchmarking improvements.

Open
#889 0 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
Python
Stars
668
Forks
141
Avg merge
2d 8h
Merged PRs (30d)
4

Description

  • Delete asv benchmarks if any.
  • Make pytest-codspeed benchmarks deterministic.
  • Use pytest fixtures wherever possible instead of generating them inline.
  • The benchmark suite needs to be reduced, perhaps by using hypothesis to sample from the overall set of benchmarks; and skipping the small datasets.
  • Expand by benchmarking more functions (function's list to be added later).

Contributor guide

Open the contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

Start by reading the existing pytest-codspeed benchmark suite and its fixtures. Define how to reduce the benchmark set, including treatment of small datasets, and identify the additional functions to benchmark; done means the suite reflects the selected scope and covers the added functions.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
performance, testing
Issue type
Refactor
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
28/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.