pydata / pydata/xarray

compose weighted with groupby, coarsen, resample, rolling etc.

Open
#3,937 7 comments 4 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

API design topic-groupby
Dominant language
Python
Stars
4.2k
Forks
1.4k
Avg merge
2d 15h
Merged PRs (30d)
14

Description

It would be nice to make weighted work with groupby - e.g. #3935 (comment)

However, it is not entirely clear to me how that should be done. One way would be to do:

da.groupby(...).weighted(weights).mean()

this would require that the groupby operation is applied over the weights (how would this be done?)
Or should it be

da.weighted(weights).groupby(...).mean()

but this seems less intuitive to me.

Or

da.groupby(..., weights=weights).mean()

Contributor guide

Open the contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

Start by reading the referenced #3935 comment and comparing the proposed weighted/groupby forms in this issue. Determine the intended API and how weighted operations should compose with groupby, coarsen, resample, and rolling; done means the chosen behavior is implemented and covered by appropriate tests.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
data
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.