cds-snc / cds-snc/notification-api

Refactor alerts to distinguish between infrastructure failure, capacity, or service limits

Open
#1,458 5 comments 0 reactions 0 assignees View on GitHub
Tech Debt
Dominant language
Python
Stars
59
Forks
19
Avg merge
2d 21h
Merged PRs (30d)
32

Description

Candidates for refactoring:

- [ ] logs-10-celery-error-1-minute-critical: This is currently tracking `"?\"ERROR/Worker\" ?\"ERROR/ForkPoolWorker\" ?\"WorkerLostError\""` found in cloudwatch `eks-cluster/application` logs. This is too generic. Identify a way to distinguish between intentional thrown errors (message limits) and legitimate celery failures.

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.