deepseek-ai / deepseek-ai/DeepEP
About num_qps_per_rank
Open
- Dominant language
- Cuda
- Stars
- 10.1k
- Forks
- 1.4k
- Avg merge
- 4d 1h
- Merged PRs (30d)
- 2
Description
I notice that when Buffer is initialized, all modes (internode, intranode, low-latency) pass the parameter 'num_qps_per_rank', but it seems that only low-latency mode will use it in __init__. I wondered why?
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.