dask / dask/distributed

Resilience

Open
#1,072 20 comments 5 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
1.7k
Forks
778
Avg merge
2h 50m
Merged PRs (30d)
3

Description

Are there any plans for a highly available distributed cluster?

How much effort would it take (or is it possible at all) to:
- periodically persist the state of the scheduler for failover?
- or use a persistent backend instead of memory?
- have standby schedulers (via leader election etc.)?

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.