NVIDIA / NVIDIA/cccl

[FEA]: Improve performance of atomic_ref

Open
#2,583 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
C++
Stars
2.5k
Forks
486
Avg merge
2d 6h
Merged PRs (30d)
295

Description

This will generate release/acquire/acq_rel/seq_cst fences within the inner loop, but it'd suffice to issue the right fences before/after the loop.

_Originally posted by @gonzalobg in https://github.com/NVIDIA/cccl/pull/2255#discussion_r1779740262_

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.