[GraphBolt][CUDA] Optimized kernel for neighbor sampling general path
Open
Work Item
- Dominant language
- Python
- Stars
- 14.3k
- Forks
- 3.1k
- PR merge metrics
- No merged PRs in 30d
Description
We use segmented sort. When https://github.com/NVIDIA/cccl/issues/931 is resolved, we can use the newly added algorithm to get a performance increase.
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.