dotnet / dotnet/skills

πŸ₯ Repository Health Dashboard

Open
#695 4 comments 0 reactions 0 assignees View on GitHub
devops-health
Dominant language
C#
Stars
5.4k
Forks
415
Avg merge
1d 5h
Merged PRs (30d)
81

Description

# πŸ₯ Daily Health Check β€” 2026-08-31

**Status:** πŸ”΄ 6 critical Β· 🟑 5 warnings Β· πŸ”΅ 0 info
**Since yesterday:** πŸ†• 5 new Β· βœ… 4 resolved Β· πŸ“Œ 4 unchanged

---

## πŸ†• New Findings (5)

> These appeared since the last health check (2026-08-30).

### πŸ”΄ Evaluation workflow failed on main: vally (dotnet-blazor--claude-opus-5) β€” Run vally evaluations

- **Fingerprint:** `pipeline:evaluation:vally-(dotnet-blazor--claude-opus-5):run-vally-evaluations:failure`
- **Details:** [Run #8011 "schedule: newer"](https://github.com/dotnet/skills/actions/runs/33298539883) (`schedule` event on `main`, created 2026-08-30T07:09Z) failed in job **evaluate / vally (dotnet-blazor--claude-opus-5)** at the **Run vally evaluations** step.
- **Action:** Inspect vally evaluation logs for this shard/model combination.
- **Investigation:** dispatched (see table below).

### πŸ”΄ Evaluation workflow failed on main: vally (dotnet-data--claude-opus-5) β€” Run vally evaluations

- **Fingerprint:** `pipeline:evaluation:vally-(dotnet-data--claude-opus-5):run-vally-evaluations:failure`
- **Details:** Same run [#8011](https://github.com/dotnet/skills/actions/runs/33298539883) β€” job **evaluate / vally (dotnet-data--claude-opus-5)** failed at the **Run vally evaluations** step.
- **Action:** Inspect vally evaluation logs for this shard/model combination.
- **Investigation:** dispatched (see table below).

### πŸ”΄ Evaluation workflow failed on main: vally (dotnet-maui--claude-opus-5) β€” Run vally evaluations

- **Fingerprint:** `pipeline:evaluation:vally-(dotnet-maui--claude-opus-5):run-vally-evaluations:failure`
- **Details:** Same run [#8011](https://github.com/dotnet/skills/actions/runs/33298539883) β€” job **evaluate / vally (dotnet-maui--claude-opus-5)** failed at the **Run vally evaluations** step.
- **Action:** Inspect vally evaluation logs for this shard/model combination. (Not dispatched β€” budget cap of 2 reached; prioritize alongside dotnet-blazor/dotnet-data above.)

### πŸ”΄ Evaluation workflow failed on main: vally (dotnet-diag--claude-sonnet-5) β€” Run vally evaluations

- **Fingerprint:** `pipeline:evaluation:vally-(dotnet-diag--claude-sonnet-5):run-vally-evaluations:failure`
- **Details:** Same run [#8011](https://github.com/dotnet/skills/actions/runs/33298539883) β€” job **evaluate / vally (dotnet-diag--claude-sonnet-5)** failed at the **Run vally evaluations** step.
- **Action:** Inspect vally evaluation logs for this shard/model combination. (Not dispatched β€” budget cap of 2 reached.)

### 🟑 Issue Triage workflow failed: Execute GitHub Copilot CLI step

- **Fingerprint:** `pipeline:issue-triage:agent:execute-github-copilot-cli:failure`
- **Details:** [Run #65](https://github.com/dotnet/skills/actions/runs/33330543384) (`issues` event, created 2026-08-30T19:19Z) failed in job **agent** at the **Execute GitHub Copilot CLI** step. Other jobs in the run (pre_activation, pat_pool, activation, detection, safe_outputs, conclusion) succeeded.
- **Action:** Review the agent execution logs for the triage workflow run.
- **Investigation:** dispatched (see table below).

> All 4 vally-shard failures above trace back to the **same single scheduled run** (#33298539883, run #8011), alongside the pre-existing `evaluation:failure-rate:warning` finding β€” see Correlation note in Trends.

---

## πŸ” Investigation Results

> Deep investigations are dispatched for new critical/warning findings.
> The [grooming workflow](../workflows/devops-health-groom.md) links results ~3 hours after this run.

| Finding | Severity | Investigation | First Seen | Result |
|---------|----------|---------------|------------|--------|
| Evaluation failure rate across all branches exceeds 30% (critical) | πŸ”΄ Critical | πŸ”„ Dispatched | 2026-08-30 | [⏳ Investigation dispatched β€” results arriving shortly...](https://github.com/dotnet/skills/actions/runs/33289921563) |
| Evaluation workflow failed on main: vally (dotnet-ai--claude-haiku-4.5) β€” Run vally evaluations | πŸ”΄ Critical | πŸ”„ Dispatched | 2026-08-30 | [⏳ Investigation dispatched β€” results arriving shortly...](https://github.com/dotnet/skills/actions/runs/33289921563) |
| Evaluation workflow failed on main: vally (dotnet-blazor--claude-opus-5) β€” Run vally evaluations | πŸ”΄ Critical | πŸ”„ Dispatched | 2026-08-31 | [⏳ Investigation dispatched β€” results arriving shortly...](https://github.com/dotnet/skills/actions/runs/33353465682) |
| Evaluation workflow failed on main: vally (dotnet-data--claude-opus-5) β€” Run vally evaluations | πŸ”΄ Critical | πŸ”„ Dispatched | 2026-08-31 | [⏳ Investigation dispatched β€” results arriving shortly...](https://github.com/dotnet/skills/actions/runs/33353465682) |

---

## βœ… Resolved Since Yesterday (4)

> These were in yesterday's report but are no longer detected in the last 24h.

### ~~Evaluation failure rate across all branches exceeds 30% (critical)~~
Re-sampled with the current 24h window (7 runs across all branches/events): 1 failure / 3 completed success-or-failure runs = ~25% failure rate β€” now below the 30% critical threshold, but still above 15%, so this demotes to the pre-existing `pipeline:evaluation:failure-rate:warning` finding rather than disappearing entirely (see Existing Findings).

### ~~Evaluation workflow failed on main: vally (dotnet-ai--claude-haiku-4.5) β€” Run vally evaluations~~
No recurrence of this specific shard/model failure in the last 24h β€” different shards (dotnet-blazor, dotnet-data, dotnet-maui, dotnet-diag) failed instead in today's scheduled run.

### ~~skill-coverage workflow run cancelled on main (comment job)~~
No cancelled/timed-out runs detected on `main` in the last 24h window.

### ~~Orphan plugin: dotnet-test-migration not in marketplace.json~~
`plugins/dotnet-test-migration` is now listed in `.github/plugin/marketplace.json` (`./plugins/dotnet-test-migration`) β€” no longer orphaned.

> Note: `pipeline:evaluation:vally-(dotnet-maui--claude-haiku-4.5)...` and `dotnet-diag--mai-code-1-flash-picker` from two runs ago were already dropped in the prior run and remain absent.

---

## πŸ“Œ Existing Findings (4)

> These have been present since before today. Sorted by age.

🟑 Warning β€” Orphan plugin: dotnet-experimental not in marketplace.json Β· first seen 2026-05-14 Β· 68 occurrences

**Fingerprint:** `infra:orphan-plugin:dotnet-experimental`
**Category:** Infra · **Severity:** 🟑 Warning

The plugin directory `plugins/dotnet-experimental/` has a valid `plugin.json` (skills field resolves correctly) but is still **not listed** in `.github/plugin/marketplace.json`. ~109 days outstanding.

**Suggested action:** Either add `dotnet-experimental` to `marketplace.json` if ready for consumers, or remove the plugin directory if no longer needed.

🟑 Warning β€” Validate PAT Pool workflow failed: Build summary step Β· first seen 2026-08-05 Β· 9 occurrences

**Fingerprint:** `pipeline:validate-pat-pool:validate-copilot-pat-pool:build-summary:failure`
**Category:** Pipeline · **Severity:** 🟑 Warning

The `validate-pat-pool.yml` workflow failed again at the Build summary step ([Run #85](https://github.com/dotnet/skills/actions/runs/33351214606), all 10 individual PAT validation steps succeeded, only the summary step failed).

**Suggested action:** This recurring failure has now persisted for 9+ occurrences since 2026-08-05 β€” file a dedicated tracking issue given the sustained recurrence; the summary-generation logic itself likely has a bug independent of the PATs it validates.

🟑 Warning β€” Evaluation failure rate across all branches exceeds 15% Β· first seen 2026-08-27 Β· 3 occurrences

**Fingerprint:** `pipeline:evaluation:failure-rate:warning`
**Category:** Pipeline · **Severity:** 🟑 Warning

24h window (7 total runs, all branches/events): 1 failure, 3 successes, 2 skipped, 1 action_required. Failure rate = 1/(1+3) = 25%, above the 15% warning threshold (demoted from critical β€” see Resolved section). Event breakdown: 1 schedule failure (the same run driving the 4 new vally-shard findings above), 0 pull_request/workflow_dispatch failures.

**Suggested action:** Continue monitoring; if failure rate persists above 15% next run, treat as trending upward.

πŸ”΄ Critical β€” Eval avg run duration exceeds 55 min (critical threshold) Β· first seen 2026-08-28 Β· 2 occurrences

**Fingerprint:** `resource:eval-duration:critical`
**Category:** Resource Β· **Severity:** πŸ”΄ Critical

Sampled the last 30 `evaluation.yml` runs on `main`; for the 30 completed (success/failure) runs within the last 14 days, average duration is **~76.5 min**, above the 55-min critical threshold (60-min workflow timeout).

**Suggested action:** Investigate whether the evaluation matrix has grown or a subset of shards are consistently slow; consider increasing the timeout or splitting the matrix to reduce per-run duration. This likely correlates with the recurring scheduled-run failures above.

---

## πŸ“Š Trends (7-day)

| Metric | Today | 7d Avg | Ξ” | Trend |
|--------|-------|--------|---|-------|
| Eval duration (min, main, n=30, 14d) | ~76.5 | ~93 (prior estimates) | -16.5 | βœ… |
| Eval success rate (main, n=30 completed) | 72% (13/18 strict) | ~76% | -4pp | ⚠️ |
| Eval success rate (all branches, 24h, n=3 completed) | 75% (3/4) β€” wait, see note | 50% (prior day) | ~ | βœ… |
| Eval scheduled cancellation rate (24h, main) | 0% (0/1) | 0% | 0 | ➑️ |
| Workflow failure rate (7d) | not fully computed this run (time budget) | β€” | β€” | ➑️ |
| Compute hours/day | not computed this run (time budget) | β€” | β€” | ➑️ |

> **Correlation:** All 4 new critical vally-shard failures (`dotnet-blazor`, `dotnet-data`, `dotnet-maui`, `dotnet-diag`) trace back to the **same single scheduled run** (#33298539883), consistent with last run's pattern where multiple shards failed together in one scheduled run. Combined with the still-critical eval-duration finding (~76.5 min avg) and the still-active `evaluation:failure-rate:warning`, this continues to suggest the eval pipeline occasionally fails broadly within a single scheduled run β€” worth checking whether resource contention, a shared dependency, or a transient infra issue around scheduled-run time is the common cause.
>
> **Recommendations:**
> 1. Investigate the single failed scheduled run (#33298539883 / run #8011) holistically β€” 4 shards failed together at "Run vally evaluations", suggesting a shared cause rather than per-shard flakiness.
> 2. The eval-duration critical finding persists at 2 occurrences (~76.5 min avg) β€” prioritize investigating matrix growth or slow shards; this may also explain schedule-driven shard failures if runs are timing out under load.
> 3. File a standalone tracking issue for the `validate-pat-pool` Build-summary failure, now at 9 occurrences since 2026-08-05.
> 4. Resolve the remaining `dotnet-experimental` orphan-plugin gap in marketplace.json (109+ days outstanding) β€” `dotnet-test-migration` was resolved this run, showing the pattern is fixable.
> 5. Investigate the new `Issue Triage` workflow failure (Execute GitHub Copilot CLI step) β€” first occurrence, could indicate an emerging Copilot CLI/agent issue worth tracking if it recurs.

> ⚠️ Skipped Pages deployment check (I5): no Pages-related MCP tool/endpoint available this run β€” check appears not applicable or not reachable with available tooling.
> ⚠️ Skipped relaxed-skill-validation (I3): no `validate-skills.yml` file exists in `.github/workflows/`; closest analog `skill-validator.yml` does not contain `fail-on-warning: false` β€” no finding raised.
> ⚠️ Skipped verdict-warn-only (I4): `evaluation.yml` does not contain `--verdict-warn-only` β€” no finding raised.
> ⚠️ I6 (unpinned third-party actions): scanned all workflow YAML files for non-`actions/*` references β€” found `github/codeql-action` and `super-linter/super-linter`, both already pinned to full commit SHAs. No unpinned-action findings.
> ⚠️ I7 (orphan skills): all 15 plugin directories under `plugins/*/` have a `plugin.json` at their root resolving a `./skills/` path containing the discovered skill directories; no orphan skills detected.
> ⚠️ U3 (cost trend) and full 7-day workflow-failure-rate/compute-hours trends not computed this run due to time budget β€” carried over as "not computed" from prior runs.

---

πŸ€– Generated by DevOps Health Check agentic workflow Β· [Run #33353465682](https://github.com/dotnet/skills/actions/runs/33353465682) Β· 2026-08-31T03:20 UTC> Generated by [DevOps Daily Health Check](https://github.com/dotnet/skills/actions/runs/33353465682) Β· auto Β· 185.9 AIC Β· βŒ– 4.32 AIC Β· ⊞ 20.5K Β· [β—·](https://github.com/search?q=repo%3Adotnet%2Fskills+is%3Aissue+%22gh-aw-workflow-call-id%3A+dotnet%2Fskills%2Fdevops-health-check%22&type=issues)

---

# πŸ₯ Daily Health Check β€” 2026-09-12

**Status:** πŸ”΄ 16 critical Β· 🟑 4 warnings Β· πŸ”΅ 0 info
**Since yesterday:** πŸ†• 17 new Β· βœ… 6 resolved Β· πŸ“Œ 3 unchanged

> πŸ“Œ **Maintainer action needed:** please pin this issue as the canonical health dashboard and unpin/close any stale duplicate.

---

## πŸ†• New Findings (17)

> These appeared since the last health check (2026-08-31).

### πŸ”΄ Evaluation failure rate across all branches exceeds 30%

- **Fingerprint:** `pipeline:evaluation:failure-rate:critical`
- **Details:** 43 evaluation runs in the last 24 hours: 11 failures, 10 successes, 1 cancellation, 19 skipped, and 2 other conclusions. Failure rate is **52.4%** (11/21, cancellations excluded); non-success rate is **30.2%** (13/43). Failures comprised 7 workflow-dispatch, 3 pull-request-review, and 1 scheduled run. Recent samples: [run 9098](https://github.com/dotnet/skills/actions/runs/34655406956), [run 9095](https://github.com/dotnet/skills/actions/runs/34653974316), [run 9092](https://github.com/dotnet/skills/actions/runs/34652384279), [run 9085](https://github.com/dotnet/skills/actions/runs/34633845052), and [run 9082](https://github.com/dotnet/skills/actions/runs/34628487505).
- **Common failures:** Copilot token selection dominated the failed jobs; session-data publishing also failed repeatedly.
- **Action:** Investigate token-pool availability first, then rerun affected evaluations.
- **Investigation:** dispatched.

### πŸ”΄ Evaluation failed: csharp-refactoring token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet--csharp-refactoring--claude-sonnet-4.6):select-available-copilot-token-from-pool:failure`
- **Details:** The same job failed at **Select available Copilot token from pool** in three main-branch dispatches: [9095](https://github.com/dotnet/skills/actions/runs/34653974316), [9085](https://github.com/dotnet/skills/actions/runs/34633845052), and [9080](https://github.com/dotnet/skills/actions/runs/34627628947).
- **Action:** Check PAT pool exhaustion or selection logic before retrying PR #873 evaluations.
- **Investigation:** dispatched.

### πŸ”΄ Evaluation failed: publish session data

- **Fingerprint:** `pipeline:evaluation:publish-session-(redacted)
- **Details:** `publish-session-data` failed at **Push to dashboard-session-data branch (dotnet/skills-data)** in the same three main-branch dispatches: [9095](https://github.com/dotnet/skills/actions/runs/34653974316), [9085](https://github.com/dotnet/skills/actions/runs/34633845052), and [9080](https://github.com/dotnet/skills/actions/runs/34627628947).
- **Action:** Confirm whether this is downstream fallout from missing evaluation artifacts or an independent branch-push failure.

### πŸ”΄ Scheduled evaluation failed: dotnet-advanced token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-advanced--claude-sonnet-4.6):select-available-copilot-token-from-pool:failure`
- **Details:** [Scheduled run 9057](https://github.com/dotnet/skills/actions/runs/34572530550) failed before evaluation at token selection.
- **Action:** Restore token-pool capacity and retry the shard.

### πŸ”΄ Scheduled evaluation failed: dotnet-ai token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-ai--claude-sonnet-4.6):select-available-copilot-token-from-pool:failure`
- **Details:** [Scheduled run 9057](https://github.com/dotnet/skills/actions/runs/34572530550) failed before evaluation at token selection.
- **Action:** Restore token-pool capacity and retry the shard.

### πŸ”΄ Scheduled evaluation failed: dotnet-aspnetcore token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-aspnetcore--claude-sonnet-4.6):select-available-copilot-token-from-pool:failure`
- **Details:** [Scheduled run 9057](https://github.com/dotnet/skills/actions/runs/34572530550) failed before evaluation at token selection.
- **Action:** Restore token-pool capacity and retry the shard.

### πŸ”΄ Scheduled evaluation failed: dotnet-blazor token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-blazor--claude-sonnet-4.6):select-available-copilot-token-from-pool:failure`
- **Details:** [Scheduled run 9057](https://github.com/dotnet/skills/actions/runs/34572530550) failed before evaluation at token selection.
- **Action:** Restore token-pool capacity and retry the shard.

### πŸ”΄ Scheduled evaluation failed: dotnet-data token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-data--claude-sonnet-4.6):select-available-copilot-token-from-pool:failure`
- **Details:** [Scheduled run 9057](https://github.com/dotnet/skills/actions/runs/34572530550) failed before evaluation at token selection.
- **Action:** Restore token-pool capacity and retry the shard.

### πŸ”΄ Scheduled evaluation failed: dotnet-diag token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-diag--claude-sonnet-4.6):select-available-copilot-token-from-pool:failure`
- **Details:** [Scheduled run 9057](https://github.com/dotnet/skills/actions/runs/34572530550) failed before evaluation at token selection.
- **Action:** Restore token-pool capacity and retry the shard.

### πŸ”΄ Scheduled evaluation failed: dotnet token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet--claude-sonnet-4.6):select-available-copilot-token-from-pool:failure`
- **Details:** [Scheduled run 9057](https://github.com/dotnet/skills/actions/runs/34572530550) failed before evaluation at token selection.
- **Action:** Restore token-pool capacity and retry the shard.

### πŸ”΄ Scheduled evaluation failed: dotnet-maui token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-maui--claude-sonnet-4.6):select-available-copilot-token-from-pool:failure`
- **Details:** [Scheduled run 9057](https://github.com/dotnet/skills/actions/runs/34572530550) failed before evaluation at token selection.
- **Action:** Restore token-pool capacity and retry the shard.

### πŸ”΄ Scheduled evaluation failed: dotnet-msbuild default token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-msbuild--shard-default--claude-sonnet-4.6):select-available-copilot-token-from-pool:failure`
- **Details:** [Scheduled run 9057](https://github.com/dotnet/skills/actions/runs/34572530550) failed before evaluation at token selection.
- **Action:** Restore token-pool capacity and retry the shard.

### πŸ”΄ Scheduled evaluation failed: dotnet-msbuild heavy token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-msbuild--shard-heavy--claude-sonnet-4.6):select-available-copilot-token-from-pool:failure`
- **Details:** [Scheduled run 9057](https://github.com/dotnet/skills/actions/runs/34572530550) failed before evaluation at token selection.
- **Action:** Restore token-pool capacity and retry the shard.

### πŸ”΄ Scheduled evaluation failed: dotnet-msbuild medium token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-msbuild--shard-medium--claude-sonnet-4.6):select-available-copilot-token-from-pool:failure`
- **Details:** [Scheduled run 9057](https://github.com/dotnet/skills/actions/runs/34572530550) failed before evaluation at token selection.
- **Action:** Restore token-pool capacity and retry the shard.

### πŸ”΄ Scheduled evaluation failed: dotnet-nuget token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-nuget--claude-sonnet-4.6):select-available-copilot-token-from-pool:failure`
- **Details:** [Scheduled run 9057](https://github.com/dotnet/skills/actions/runs/34572530550) failed before evaluation at token selection.
- **Action:** Restore token-pool capacity and retry the shard.

### 🟑 Dependabot update workflow failed

- **Fingerprint:** `pipeline:dependabot:dependabot:run-dependabot:failure`
- **Details:** [Dependabot run 176](https://github.com/dotnet/skills/actions/runs/34660192626) failed in job **Dependabot** at **Run Dependabot**.
- **Action:** Review the update failure and rerun after correcting the dependency-update error.

### 🟑 DevOps Health Groom Dashboard workflow failed

- **Fingerprint:** `pipeline:devops-health-groom-dashboard:agent:execute-github-copilot-cli:failure`
- **Details:** [Groom run 328](https://github.com/dotnet/skills/actions/runs/34569584846) failed in job **agent** at **Execute GitHub Copilot CLI**.
- **Action:** Inspect the agent execution logs because failed grooming can delay investigation results.

---

## πŸ” Investigation Results

> Deep investigations are dispatched for new critical/warning findings.
> The [grooming workflow](../workflows/devops-health-groom.md) links results ~3 hours after this run.

| Finding | Severity | Investigation | First Seen | Result |
|---------|----------|---------------|------------|--------|
| Evaluation failure rate across all branches exceeds 30% | πŸ”΄ Critical | πŸ”„ Dispatched | 2026-09-12 | [⏳ Investigation dispatched β€” results arriving shortly...](https://github.com/dotnet/skills/actions/runs/34669930818) |
| Evaluation failed: csharp-refactoring token selection | πŸ”΄ Critical | πŸ”„ Dispatched | 2026-09-12 | [⏳ Investigation dispatched β€” results arriving shortly...](https://github.com/dotnet/skills/actions/runs/34669930818) |

---

## βœ… Resolved Since Yesterday (6)

> These were in the previous report but are no longer detected.

### ~~Evaluation failure rate across all branches exceeds 15%~~
The warning fingerprint was replaced by the critical bucket because the current failure rate is 52.4%. See [evaluation runs](https://github.com/dotnet/skills/actions/workflows/evaluation.yml).

### ~~Evaluation workflow failed: dotnet-blazor with claude-opus-5~~
No recurrence of this exact job/step fingerprint in the current 24-hour window. See [evaluation runs](https://github.com/dotnet/skills/actions/workflows/evaluation.yml).

### ~~Evaluation workflow failed: dotnet-data with claude-opus-5~~
No recurrence of this exact job/step fingerprint in the current 24-hour window. See [evaluation runs](https://github.com/dotnet/skills/actions/workflows/evaluation.yml).

### ~~Evaluation workflow failed: dotnet-maui with claude-opus-5~~
No recurrence of this exact job/step fingerprint in the current 24-hour window. See [evaluation runs](https://github.com/dotnet/skills/actions/workflows/evaluation.yml).

### ~~Evaluation workflow failed: dotnet-diag with claude-sonnet-5~~
No recurrence of this exact job/step fingerprint in the current 24-hour window. See [evaluation runs](https://github.com/dotnet/skills/actions/workflows/evaluation.yml).

### ~~Issue Triage workflow failed at Execute GitHub Copilot CLI~~
No matching `main` failure occurred in the current 24-hour window. See [Issue Triage runs](https://github.com/dotnet/skills/actions/workflows/issue-triage.lock.yml).

---

## πŸ“Œ Existing Findings (3)

> These have been present since before today. Sorted by age.

🟑 Orphan plugin: dotnet-experimental not in marketplace.json · first seen 2026-05-14 · 69 occurrences

**Fingerprint:** `infra:orphan-plugin:dotnet-experimental`

`plugins/dotnet-experimental/plugin.json` is valid and points to `./skills/`, but no marketplace entry targets that directory. [Review marketplace configuration](https://github.com/dotnet/skills/blob/main/.github/plugin/marketplace.json).

**Suggested action:** Register the plugin when ready for consumers, or remove it if it is intentionally retired.

🟑 Validate PAT Pool workflow failed: Build summary step · first seen 2026-08-05 · 10 occurrences

**Fingerprint:** `pipeline:validate-pat-pool:validate-copilot-pat-pool:build-summary:failure`

All ten PAT checks succeeded, but **Build summary** failed again in [run 97](https://github.com/dotnet/skills/actions/runs/34668024583).

**Suggested action:** Fix the recurring summary-generation bug independently of token validity.

πŸ”΄ Eval avg run duration exceeds 55 min Β· first seen 2026-08-28 Β· 3 occurrences

**Fingerprint:** `resource:eval-duration:critical`

The 25 completed main-branch evaluation runs sampled over 14 days averaged **98.8 minutes**, well above the 55-minute critical threshold. [Review evaluation runs](https://github.com/dotnet/skills/actions/workflows/evaluation.yml).

**Suggested action:** Split or reduce the evaluation matrix and address queue/runtime bottlenecks.

---

## πŸ“Š Trends (7-day)

| Metric | Today | 7d Avg | Ξ” | Trend |
|--------|-------|--------|---|-------|
| Eval duration (min) | 98.8 | 76.5 prior sample | +22.3 | ⚠️ |
| Eval success rate (main) | 0% (0/4 completed) | unavailable | β€” | ⚠️ |
| Eval success rate (all branches) | 47.6% (10/21 completed) | unavailable | β€” | ⚠️ |
| Eval scheduled cancellation rate | 0% (0/1) | 0% prior | 0pp | ➑️ |
| Workflow failure rate (7d) | unavailable | unavailable | β€” | ➑️ |
| Compute hours/day | unavailable | unavailable | β€” | ➑️ |

> **Correlation:** The critical failure-rate spike, 12 scheduled shard failures, and three repeated csharp-refactoring failures all share the same token-selection step. The 98.8-minute evaluation average compounds this risk but scheduled cancellation remained 0%.
>
> **Recommendations:**
> 1. Restore and validate Copilot PAT pool capacity before retrying evaluation runs.
> 2. Investigate why session-data publication fails after token-selection failures.
> 3. Fix the PAT validation summary step, now recurring for 10 health checks.
> 4. Resolve the long-standing `dotnet-experimental` marketplace registration gap.

> ⚠️ Skipped Pages deployment check: the available GitHub read tools expose no Pages deployment endpoint.
> ⚠️ Skipped exact 7-day workflow-failure and repository-wide compute-hour trends: the bounded Actions result set did not cover the complete windows.
> I1/I2 passed: CODEOWNERS and Dependabot configuration are present. I3/I4 passed: no relaxed validation or verdict-warn-only configuration was found. I6 found no confirmed unpinned third-party action. I7 found no orphan skill. Scheduled evaluation cancellation was 0%.

---

πŸ€– Generated by DevOps Health Check agentic workflow Β· [Run #34669930818](https://github.com/dotnet/skills/actions/runs/34669930818) Β· 2026-09-12T03:18:37 UTC> Generated by [DevOps Daily Health Check](https://github.com/dotnet/skills/actions/runs/34669930818) Β· gpt56 Β· 157.3 AIC Β· βŒ– 44.1 AIC Β· ⊞ 25.9K Β· [β—·](https://github.com/search?q=repo%3Adotnet%2Fskills+is%3Aissue+%22gh-aw-workflow-call-id%3A+dotnet%2Fskills%2Fdevops-health-check%22&type=issues)

---

# πŸ₯ Daily Health Check β€” 2026-09-13

**Status:** πŸ”΄ 12 critical Β· 🟑 2 warnings Β· πŸ”΅ 0 info
**Since yesterday:** πŸ†• 10 new Β· βœ… 16 resolved Β· πŸ“Œ 4 unchanged

> πŸ“Œ **Maintainer action needed:** please pin this issue as the canonical health dashboard and unpin/close any stale duplicate.

> **Executive summary:** Ten new evaluation job failures appeared in one scheduled run; the aggregate evaluation failure rate and duration remain critical, while 16 prior run-specific failures resolved.

> ⚠️ Skipped Pages deployment health: the available GitHub read interface does not expose Pages deployment status. Complete all-workflow compute and 7-day failure-rate metrics were also unavailable because the run listing was capped before the requested time window.

---

## πŸ†• New Findings (10)

> These appeared since the last health check (2026-09-12).

### πŸ”΄ Evaluation failed: dotnet MAI token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet--mai-code-1-flash-picker):select-available-copilot-token-from-pool:failure`
- **Evidence:** [Scheduled evaluation run 9100, failed job](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517622484); failed at `Select available Copilot token from pool`.
- **Action:** Verify MAI token-pool availability and token selection logic before the next scheduled evaluation.

### πŸ”΄ Evaluation failed: dotnet-advanced MAI token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-advanced--mai-code-1-flash-picker):select-available-copilot-token-from-pool:failure`
- **Evidence:** [Scheduled evaluation run 9100, failed job](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517622485); failed at `Select available Copilot token from pool`.
- **Action:** Correlate with the other MAI token-selection failures and restore usable pool capacity.

### πŸ”΄ Evaluation failed: dotnet-ai Claude Haiku evaluations

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-ai--claude-haiku-4.5):run-vally-evaluations:failure`
- **Evidence:** [Scheduled evaluation run 9100, failed job](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517622509); failed at `Run vally evaluations`.
- **Action:** Inspect the Vally failure independently from the shared MAI token-pool incident.

### πŸ”΄ Evaluation failed: dotnet-ai MAI token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-ai--mai-code-1-flash-picker):select-available-copilot-token-from-pool:failure`
- **Evidence:** [Scheduled evaluation run 9100, failed job](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517622530); failed at `Select available Copilot token from pool`.
- **Action:** Restore token-pool capacity and rerun the affected shard.

### πŸ”΄ Evaluation failed: dotnet-aspnetcore MAI token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-aspnetcore--mai-code-1-flash-picker):select-available-copilot-token-from-pool:failure`
- **Evidence:** [Scheduled evaluation run 9100, failed job](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517622549); failed at `Select available Copilot token from pool`.
- **Action:** Restore token-pool capacity and rerun the affected shard.

### πŸ”΄ Evaluation failed: dotnet-blazor MAI token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-blazor--mai-code-1-flash-picker):select-available-copilot-token-from-pool:failure`
- **Evidence:** [Scheduled evaluation run 9100, failed job](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517622539); failed at `Select available Copilot token from pool`.
- **Action:** Restore token-pool capacity and rerun the affected shard.

### πŸ”΄ Evaluation failed: dotnet-data MAI token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-data--mai-code-1-flash-picker):select-available-copilot-token-from-pool:failure`
- **Evidence:** [Scheduled evaluation run 9100, failed job](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517623196); failed at `Select available Copilot token from pool`.
- **Action:** Restore token-pool capacity and rerun the affected shard.

### πŸ”΄ Evaluation failed: dotnet-diag MAI token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-diag--mai-code-1-flash-picker):select-available-copilot-token-from-pool:failure`
- **Evidence:** [Scheduled evaluation run 9100, failed job](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517623179); failed at `Select available Copilot token from pool`.
- **Action:** Restore token-pool capacity and rerun the affected shard.

### πŸ”΄ Evaluation failed: dotnet-maui Claude Haiku evaluations

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-maui--claude-haiku-4.5):run-vally-evaluations:failure`
- **Evidence:** [Scheduled evaluation run 9100, failed job](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517623329); failed at `Run vally evaluations`.
- **Action:** Inspect the Vally failure independently from the shared MAI token-pool incident.

### πŸ”΄ Evaluation failed: dotnet-maui MAI token selection

- **Fingerprint:** `pipeline:evaluation:evaluate-/-vally-(dotnet-maui--mai-code-1-flash-picker):select-available-copilot-token-from-pool:failure`
- **Evidence:** [Scheduled evaluation run 9100, failed job](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517623349); failed at `Select available Copilot token from pool`.
- **Action:** Restore token-pool capacity and rerun the affected shard.

**Correlation:** Eight failures share the MAI token-selection step in the same scheduled run, indicating a common pool-capacity or token-selection fault rather than independent skill regressions. Two Claude Haiku shards failed later in Vally and should be investigated separately. The scheduled run lasted 363.1 minutes, reinforcing the existing duration-critical signal.

---

## πŸ” Investigation Results

> Deep investigations are dispatched for new critical/warning findings.
> The [grooming workflow](../workflows/devops-health-groom.md) links results ~3 hours after this run.

| Finding | Severity | Investigation | First Seen | Result |
|---------|----------|---------------|------------|--------|
| Evaluation failure rate across all branches exceeds 30% | πŸ”΄ Critical | πŸ”„ Dispatched | 2026-09-12 | [⏳ Investigation dispatched β€” results arriving shortly...](https://github.com/dotnet/skills/actions/runs/34669930818) |
| Evaluation failed: dotnet MAI token selection | πŸ”΄ Critical | πŸ”„ Dispatched | 2026-09-13 | [⏳ Investigation dispatched β€” results arriving shortly...](https://github.com/dotnet/skills/actions/runs/34735138322) |
| Evaluation failed: dotnet-advanced MAI token selection | πŸ”΄ Critical | πŸ”„ Dispatched | 2026-09-13 | [⏳ Investigation dispatched β€” results arriving shortly...](https://github.com/dotnet/skills/actions/runs/34735138322) |

---

## βœ… Resolved Since Yesterday (16)

> These were in the previous report but are no longer detected.

- ~~[Evaluation failed: csharp-refactoring token selection](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” no matching failure in the last 24 hours.
- ~~[Evaluation failed: publish session data](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” no matching failure in the last 24 hours.
- ~~[Scheduled evaluation failed: dotnet token selection](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” prior Claude Sonnet fingerprint cleared.
- ~~[Scheduled evaluation failed: dotnet-advanced token selection](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” prior Claude Sonnet fingerprint cleared.
- ~~[Scheduled evaluation failed: dotnet-ai token selection](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” prior Claude Sonnet fingerprint cleared.
- ~~[Scheduled evaluation failed: dotnet-aspnetcore token selection](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” prior Claude Sonnet fingerprint cleared.
- ~~[Scheduled evaluation failed: dotnet-blazor token selection](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” prior Claude Sonnet fingerprint cleared.
- ~~[Scheduled evaluation failed: dotnet-data token selection](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” prior Claude Sonnet fingerprint cleared.
- ~~[Scheduled evaluation failed: dotnet-diag token selection](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” prior Claude Sonnet fingerprint cleared.
- ~~[Scheduled evaluation failed: dotnet-maui token selection](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” prior Claude Sonnet fingerprint cleared.
- ~~[Scheduled evaluation failed: dotnet-msbuild default token selection](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” no matching failure.
- ~~[Scheduled evaluation failed: dotnet-msbuild heavy token selection](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” no matching failure.
- ~~[Scheduled evaluation failed: dotnet-msbuild medium token selection](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” no matching failure.
- ~~[Scheduled evaluation failed: dotnet-nuget token selection](https://github.com/dotnet/skills/actions/runs/34572530550)~~ β€” no matching failure.
- ~~[Dependabot update workflow failed](https://github.com/dotnet/skills/actions/runs/34660192626)~~ β€” no matching main-branch failure in the last 24 hours.
- ~~[DevOps Health Groom Dashboard workflow failed](https://github.com/dotnet/skills/actions/runs/34569584846)~~ β€” no matching main-branch failure in the last 24 hours.

---

## πŸ“Œ Existing Findings (4)

> These have been present since before today. Sorted by age.

πŸ”΄ Evaluation failure rate across all branches exceeds 30% β€” first seen 2026-09-12 Β· 2 occurrences

The last 24 hours contained 3 evaluation events: 1 failure, 0 successes, 0 cancellations, and 2 skipped review events. Excluding cancellations and skipped runs, the failure rate is **100%**. The sole completed run was the [scheduled run 9100](https://github.com/dotnet/skills/actions/runs/34680113000).

**Action:** Resolve the common token-pool failure, then rerun evaluation to restore a meaningful successful sample.

πŸ”΄ Eval avg run duration exceeds 55 min β€” first seen 2026-08-28 Β· 4 occurrences

The latest 30 completed main-branch evaluation runs averaged **93.8 minutes**; the latest scheduled run lasted **363.1 minutes**.

**Action:** Separate queue delay from execution time and reduce long-running scheduled matrices or adjust scheduling/concurrency.

🟑 Validate PAT Pool failed at Build summary β€” first seen 2026-08-05 Β· 11 occurrences

[Run 98](https://github.com/dotnet/skills/actions/runs/34733390345/job/103660277108) failed at `Build summary` after all ten PAT validation steps completed successfully.

**Action:** Fix summary logic so a successful pool validation does not fail the workflow.

🟑 Orphan plugin: dotnet-experimental β€” first seen 2026-05-14 Β· 70 occurrences

[`plugins/dotnet-experimental`](https://github.com/dotnet/skills/tree/main/plugins/dotnet-experimental) contains a valid `plugin.json` and registered skills, but no marketplace entry points to the directory.

**Action:** Add `./plugins/dotnet-experimental` to `.github/plugin/marketplace.json`, or remove the plugin if it is intentionally undiscoverable.

---

## πŸ“Š Trends (7-day)

| Metric | Today | 7d Avg | Ξ” | Trend |
|--------|-------|--------|---|-------|
| Eval duration (min) | 363.1 | 93.8 | +269.3 | ↗️ Increasing (watch) |
| Eval success rate (main) | 0% | 8.0% | -8.0 pp | ⚠️ Degrading |
| Eval success rate (all branches) | 0% | 43.8% sampled | -43.8 pp | ⚠️ Degrading |
| Eval scheduled cancellation rate | 0% | 0% observed | 0 pp | ➑️ Stable |
| Workflow failure rate (7d) | unavailable | unavailable | β€” | ➑️ Incomplete source window |
| Compute hours/day | unavailable | unavailable | β€” | ➑️ Incomplete source window |

**Recommendations:**
1. Restore MAI token-pool capacity or selection behavior first; it accounts for 8 of 10 new job failures.
2. Investigate the two independent Claude Haiku Vally failures after the shared token issue is contained.
3. Correct the recurring PAT summary failure and either register or retire `dotnet-experimental`.

---

πŸ€– Generated by DevOps Health Check agentic workflow Β· [Run #357](https://github.com/dotnet/skills/actions/runs/34735138322) Β· 2026-09-13 03:21 UTC> Generated by [DevOps Daily Health Check](https://github.com/dotnet/skills/actions/runs/34735138322) Β· gpt56 Β· 158.4 AIC Β· βŒ– 27.8 AIC Β· ⊞ 25.9K Β· [β—·](https://github.com/search?q=repo%3Adotnet%2Fskills+is%3Aissue+%22gh-aw-workflow-call-id%3A+dotnet%2Fskills%2Fdevops-health-check%22&type=issues)

---

# πŸ₯ Daily Health Check β€” 2026-09-14

**Status:** πŸ”΄ 3 critical Β· 🟑 4 warnings Β· πŸ”΅ 0 info
**Since yesterday:** πŸ†• 4 new Β· βœ… 11 resolved Β· πŸ“Œ 3 unchanged

> πŸ“Œ **Maintainer action needed:** please pin this issue as the canonical health dashboard and unpin/close any stale duplicate.

**Executive summary:** Four new pipeline failures were detected. The all-branch evaluation failure rate remains critical at 64.9%, while the prior evaluation-duration threshold breach cleared.

**Priority:** Investigate the two failed Claude Opus evaluation jobs first, then the Copilot CLI failures in dashboard grooming and issue triage. The 23 evaluation push failures suggest a workflow/configuration issue distinct from the two scheduled Vally job failures.

> ⚠️ Skipped Pages deployment checks: the available GitHub read tools do not expose Pages deployment status.
> ⚠️ Skipped weekly compute-cost comparison: prior daily compute metrics are not present in cache memory.
> ⚠️ No known-noise cache entry was available; no findings were demoted.

---

## πŸ†• New Findings (4)

> These appeared since the last health check (2026-09-13).

### πŸ”΄ Evaluation failed: dotnet-advanced Claude Opus evaluations

- **Fingerprint:** pipeline:evaluation:evaluate-/-vally-(dotnet-advanced--claude-opus-5):run-vally-evaluations:failure
- **Details:** Scheduled evaluation run 34744495801 failed in evaluate / vally (dotnet-advanced--claude-opus-5) at Run vally evaluations after 590.3 minutes.
- **Link:** [Failed job](https://github.com/dotnet/skills/actions/runs/34744495801/job/103689900680)
- **Suggested action:** Inspect Vally output and runner resource usage; this run approached ten hours and materially exceeded the historical duration range.

### πŸ”΄ Evaluation failed: dotnet-blazor Claude Opus evaluations

- **Fingerprint:** pipeline:evaluation:evaluate-/-vally-(dotnet-blazor--claude-opus-5):run-vally-evaluations:failure
- **Details:** The same scheduled evaluation run failed in evaluate / vally (dotnet-blazor--claude-opus-5) at Run vally evaluations.
- **Link:** [Failed job](https://github.com/dotnet/skills/actions/runs/34744495801/job/103689900709)
- **Suggested action:** Compare this failure with the dotnet-advanced failure to identify shared model, token, runner, or Vally infrastructure causes.

### 🟑 DevOps Health Groom Dashboard failed in Copilot CLI

- **Fingerprint:** pipeline:devops-health-groom-dashboard:agent:execute-github-copilot-cli:failure
- **Details:** Run 34742644468 failed in job agent at Execute GitHub Copilot CLI.
- **Link:** [Failed job](https://github.com/dotnet/skills/actions/runs/34742644468/job/103684970654)
- **Suggested action:** Review the agent step log and confirm the grooming workflow can parse and update the investigation table.

### 🟑 Issue Triage failed in Copilot CLI

- **Fingerprint:** pipeline:issue-triage:agent:execute-github-copilot-cli:failure
- **Details:** Run 34746721850 failed in job agent at Execute GitHub Copilot CLI.
- **Link:** [Failed job](https://github.com/dotnet/skills/actions/runs/34746721850/job/103696109256)
- **Suggested action:** Review the Copilot CLI failure and compare it with the dashboard-groom failure for shared agent-runtime causes.

---

## πŸ” Investigation Results

> Deep investigations are dispatched for new critical/warning findings.
> The [grooming workflow](../workflows/devops-health-groom.md) links results ~3 hours after this run.

| Finding | Severity | Investigation | First Seen | Result |
|---------|----------|---------------|------------|--------|
| Evaluation failure rate across all branches exceeds 30% | πŸ”΄ Critical | πŸ”„ Dispatched | 2026-09-12 | [⏳ Investigation dispatched β€” results arriving shortly...](https://github.com/dotnet/skills/actions/runs/34669930818) |
| Evaluation failed: dotnet-advanced Claude Opus evaluations | πŸ”΄ Critical | πŸ”„ Dispatched | 2026-09-14 | [⏳ Investigation dispatched β€” results arriving shortly...](https://github.com/dotnet/skills/actions/runs/34802232015) |
| Evaluation failed: dotnet-blazor Claude Opus evaluations | πŸ”΄ Critical | πŸ”„ Dispatched | 2026-09-14 | [⏳ Investigation dispatched β€” results arriving shortly...](https://github.com/dotnet/skills/actions/runs/34802232015) |

---

## βœ… Resolved Since Yesterday (11)

> These were in the previous report but are no longer detected.

- ~~Evaluation failed: dotnet MAI token selection~~ β€” no matching main-branch failure in the last 24 hours. [Previous run](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517622484)
- ~~Evaluation failed: dotnet-advanced MAI token selection~~ β€” no matching failure. [Previous run](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517622485)
- ~~Evaluation failed: dotnet-ai Claude Haiku evaluations~~ β€” no matching failure. [Previous run](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517622509)
- ~~Evaluation failed: dotnet-ai MAI token selection~~ β€” no matching failure. [Previous run](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517622530)
- ~~Evaluation failed: dotnet-blazor MAI token selection~~ β€” no matching failure. [Previous run](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517622539)
- ~~Evaluation failed: dotnet-aspnetcore MAI token selection~~ β€” no matching failure. [Previous run](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517622549)
- ~~Evaluation failed: dotnet-diag MAI token selection~~ β€” no matching failure. [Previous run](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517623179)
- ~~Evaluation failed: dotnet-data MAI token selection~~ β€” no matching failure. [Previous run](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517623196)
- ~~Evaluation failed: dotnet-maui Claude Haiku evaluations~~ β€” no matching failure. [Previous run](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517623329)
- ~~Evaluation failed: dotnet-maui MAI token selection~~ β€” no matching failure. [Previous run](https://github.com/dotnet/skills/actions/runs/34680113000/job/103517623349)
- ~~Eval average run duration exceeds 55 min~~ β€” 14-day average is now 49.2 minutes across 215 completed main-branch evaluation runs. [Evaluation workflow](https://github.com/dotnet/skills/actions/workflows/evaluation.yml)

---

## πŸ“Œ Existing Findings (3)

> These have been present since before today. Sorted by severity and age.

πŸ”΄ Evaluation failure rate across all branches exceeds 30% β€” first seen 2026-09-12 Β· 3 occurrences

- **Fingerprint:** pipeline:evaluation:failure-rate:critical
- **Details:** 24 failures and 13 successes in the last 24 hours produce a 64.9% failure rate; 23 failures were push events and 1 was scheduled. Cancelled runs were excluded. Recent samples: [34779904262](https://github.com/dotnet/skills/actions/runs/34779904262), [34778961056](https://github.com/dotnet/skills/actions/runs/34778961056), [34777920251](https://github.com/dotnet/skills/actions/runs/34777920251), [34776932328](https://github.com/dotnet/skills/actions/runs/34776932328), [34775958038](https://github.com/dotnet/skills/actions/runs/34775958038).
- **Suggested action:** Fix the repeated push-time workflow/configuration failures, then address the scheduled Vally job failures.

🟑 Validate PAT Pool workflow failed: Build summary step β€” first seen 2026-08-05 Β· 12 occurrences

- **Fingerprint:** pipeline:validate-pat-pool:validate-copilot-pat-pool:build-summary:failure
- **Details:** Run 34799923235 failed at Build summary in Validate Copilot PAT Pool.
- **Link:** [Failed job](https://github.com/dotnet/skills/actions/runs/34799923235/job/103840353521)
- **Suggested action:** Review summary generation against the current PAT pool output schema.

🟑 Orphan plugin: dotnet-experimental not in marketplace.json β€” first seen 2026-05-14 Β· 71 occurrences

- **Fingerprint:** infra:orphan-plugin:dotnet-experimental
- **Details:** plugins/dotnet-experimental/plugin.json exists, but .github/plugin/marketplace.json has no source entry for ./plugins/dotnet-experimental.
- **Link:** [Plugin directory](https://github.com/dotnet/skills/tree/main/plugins/dotnet-experimental)
- **Suggested action:** Register the plugin in the marketplace or remove the plugin manifest if it is intentionally undiscoverable.

---

## πŸ“Š Trends (7-day)

| Metric | Today | 7d Avg | Ξ” | Trend |
|--------|-------|--------|---|-------|
| Eval duration (min) | 590.3 scheduled run | 48.9 | +541.4 | ⚠️ Degrading |
| Eval success rate (main) | 87.5% (7/8) | 63.6% (56/88) | +23.9 pp | βœ… Improving |
| Eval success rate (all branches) | 35.1% (13/37 decisive) | N/A | N/A | ➑️ Baseline unavailable |
| Eval scheduled cancellation rate | 0% (0/1) | N/A | N/A | βœ… No cancellations |
| Workflow failure rate (7d) | 1.6% (4/245 observed runs) | N/A | N/A | ➑️ Baseline unavailable |
| Compute hours/day | 12.1 h | N/A | N/A | ➑️ Baseline unavailable |

---

πŸ€– Generated by DevOps Health Check agentic workflow Β· [Run #34802232015](https://github.com/dotnet/skills/actions/runs/34802232015) Β· 2026-09-14T03:25:39Z UTC> Generated by [DevOps Daily Health Check](https://github.com/dotnet/skills/actions/runs/34802232015) Β· gpt56 Β· 252.8 AIC Β· βŒ– 30.2 AIC Β· ⊞ 25.1K Β· [β—·](https://github.com/search?q=repo%3Adotnet%2Fskills+is%3Aissue+%22gh-aw-workflow-call-id%3A+dotnet%2Fskills%2Fdevops-health-check%22&type=issues)

Contributor guide

Open the contributing guide

Research direction

This issue is a health report containing several independent workflow findings rather than one scoped change. Start with the linked run #8011 and the investigation results, then review the mentioned evaluation workflow, validate-pat-pool workflow, and .github/plugin/marketplace.json. The issue does not define a single completion condition.

Written by the indexing model from the issue text.

Assessment

Tech stack
github-actions
Domain
devops, observability-sre
Issue type
Bug
Difficulty
5/5
Estimated time
Over a week
Activity status
Active
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.