envoyproxy / envoyproxy/gateway
Best Practices for Load Shedding in Envoy Gateway
- Lenguaje dominante
- Go
- Estrellas
- 3k
- Forks
- 864
- Merge medio
- 1 d 22 h
- PR fusionados (30 d)
- 148
Descripción
Hello,
I'm looking for guidance on implementing load shedding in Envoy Gateway to protect backend services under high load.
In networking, mechanisms such as Random Early Detection (RED) and traffic policing proactively prevent congestion by dropping or limiting traffic before the network becomes saturated. I'm looking for the equivalent approach for HTTP/API traffic using Envoy Gateway.
Specifically, I'd like to know:
Does Envoy Gateway support proactive load shedding based on resource pressure (e.g., request latency, concurrency, queue depth, CPU utilization, or other overload signals)?
Is Envoy's adaptive concurrency filter currently supported and configurable through Envoy Gateway?
What is the recommended way to reject excess requests before backend services become overloaded?
Are there best practices for combining:
Local or global rate limiting
Circuit breakers
Adaptive concurrency
Overload Manager
Kubernetes HPA
If some of these capabilities are not yet exposed by Envoy Gateway, what is the recommended production approach today?
Our goal is not only to enforce rate limits, but to gracefully shed load when the platform approaches its safe operating capacity, ensuring that the system remains responsive instead of allowing latency to grow until services become unavailable.
If there are existing examples, documentation, or recommended configuration patterns for this use case, I would greatly appreciate being pointed to them.
Thank you!
Guía de contribución
No hay ninguna guía de contribución indexada para este repositorio
Línea de trabajo
No file, test, or entry point is identified in the issue. Start by reviewing Envoy Gateway's existing documentation and configuration support for rate limiting, circuit breakers, adaptive concurrency, Overload Manager, and Kubernetes HPA; done means documenting which capabilities are supported and the recommended production load-shedding approach.
Escrito por el modelo de indexación a partir del texto del issue.
Evaluación
- Stack tecnológico
- kubernetes
- Área
- backend-api-design, networking
- Tipo de issue
- Documentación
- Dificultad
- 5/5
- Tiempo estimado
- Más de una semana
- Estado de actividad
- Activo
- Claridad
- Necesita aclaración
- Aptitud para principiantes
- 35/100