NVIDIA / NVIDIA/cccl

[CUB]: Gather requirements and prepare high level design for device resident problem size for cub::DeviceRadixSort

Open
#10,758 0 comments 0 reactions 1 assignee Claimed by @NaderAlAwar View on GitHub
Dominant language
C++
Stars
2.5k
Forks
486
Avg merge
2d 6h
Merged PRs (30d)
295

Description

This issue can be closed with a list of design alternatives we considered with pros and cons for each, e.g., performance, temporary storage size, interface changes.

Note: this applies to the onesweep kernel implementation only. We can have older architectures use the onesweep in this case.

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.