dmlc / dmlc/dgl

[GraphBolt][CUDA] Optimized kernel for neighbor sampling general path

Open
#7,272 0 comments 0 reactions 1 assignee Claimed by @mfbalin View on GitHub
Work Item
Dominant language
Python
Stars
14.3k
Forks
3.1k
PR merge metrics
No merged PRs in 30d

Description

We use segmented sort. When https://github.com/NVIDIA/cccl/issues/931 is resolved, we can use the newly added algorithm to get a performance increase.

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.