Enhanced monitoring support in AKS MCP server
- Dominant language
- Go
- Stars
- 140
- Forks
- 45
- Avg merge
- 4d 3h
- Merged PRs (30d)
- 2
Description
**WHAT**
We would like a richer integration with the telemetry in AKS
- Can I use the AKS platform metrics
- How can I check if managed prometheus is enabled, if enabled can I get specific metrics for the cluster
- How can I check if Container Insights is enabled, if enabled can I get logs from pods/nodes
- How can I check if Diagnostic settings is installed, can I get the audit logs or autoscaler or specific component logs from the cluster
- How can I check if Application Insights is connected to the namespace, can I get specific app traces or logs
- How can I get the Azure Monitor alerts for the cluster
- How can I get the Resource Health alerts for the cluster
**WHY**
This will help with richer troubleshooting for customers especially for historical incidents
This might need a refactoring of the current tools to better design for the use cases
Contributor guide
Research direction
Start by reviewing the AKS MCP server's current monitoring tools and how they access cluster telemetry. Map the requested AKS platform metrics, managed Prometheus, Container Insights, diagnostic settings, Application Insights, Azure Monitor alerts, and Resource Health alerts, then determine the refactoring and scope needed to support these troubleshooting use cases.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- azure, go, kubernetes
- Domain
- backend-api-design, cloud, observability
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100