NVIDIA / NVIDIA/cudf

[FEA] Support more dtypes in JIT GroupBy `apply`

Open
#12,608 6 comments 0 reactions 0 assignees View on GitHub
feature request good first issue numba Python
Dominant language
C++
Stars
9.8k
Forks
1.1k
Avg merge
3d 6m
Merged PRs (30d)
278

Description

**Is your feature request related to a problem? Please describe.**
When https://github.com/rapidsai/cudf/pull/11452 lands, we'll get JIT `Groupby.apply` for a subset of UDFs and importantly, dtypes. However over the summer we only got as far as writing overloads for `float64` and `int64` dtypes in the users source data. It'd be nice if we could support more dtypes, starting at least with the rest of the numeric types.

**Describe the solution you'd like**
Extend the existing `groupby.apply`, `engine='jit'` framework to support the following additional dtypes:
- `float32`
- `int32`
- `int16`
- `int8`
- `uint64`
- `uint32`
- `uint16`
- `uint8`
- `bool`

A lot of the machinery in the original PR is fairly general and should make adding many of these easy- however there will undoubtedly be edge cases. As such it makes for a pretty good first issue for anyone jumping into the numba extension piece of the codebase.

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.