what if scheduler/server/worker failed?
Open
- Dominant language
- C++
- Stars
- 1.6k
- Forks
- 540
- PR merge metrics
- No merged PRs in 30d
Description
Can we restart this node and the training job still works
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.