github / github/gh-aw-threat-detection

[detection-stats] Detection stats for github/gh-aw - 2026-09-06

Open
#1,038 0 comments 0 reactions 0 assignees View on GitHub
automation detection-stats
Dominant language
Go
Stars
13
Forks
7
Avg merge
9h 52m
Merged PRs (30d)
25

Description

578 external-detector runs analysed, detection-job error rate 0%, soft failures 69, threat rate 0%.

## Summary

# gh-aw detection statistics - 2026-09-06 (UTC)

Repository: `github/gh-aw`
Window: `2026-09-06T00:00:00Z` .. `2026-09-06T23:59:59Z`
API requests: 1470, rate-limit pauses: 1
Data complete: yes

## Totals

| Metric | Count |
|---|---|
| Workflow runs in window | 1339 |
| Agentic runs (`*.lock.yml`) | 909 |
| Runs with a `detection` job | 579 |
| ... using the external detector | 578 |
| ... using the built-in detector | 1 |
| ... detector could not be determined | 0 |
| Agentic runs without a `detection` job | 330 |

All rates below are over the **external detector** population (578 runs). Since gh-aw #54111 the external detector is the compile-time default; a run counts as external when its `detection` job showed the `Install threat-detect binary` step, when another run of the same workflow did that day, or when the workflow is `.lock.yml` and no run showed the built-in shape (a completed detection job with steps but no marker). Runs on a workflow that opted out with `features: gh-aw-detection: false` count as built-in.

## Detection job outcomes

| Outcome | Count | Rate |
|---|---|---|
| `success` | 306 | 52.94% |
| `skipped` | 272 | 47.06% |

**Error rate (failure/timed_out/action_required): 0%**

## Verdict availability

| State | Meaning | Count |
|---|---|---|
| `present` | detection artifact downloaded and parsed | 231 |
| `absent` | detection job ran but published no artifact (soft failure) | 69 |
| `skipped` | detection job was skipped or was still running (nothing to fetch) | 272 |
| `unreadable` | artifact zip could not be unpacked | 6 |

Green detection jobs that published no verdict: **69** (detection steps are `continue-on-error`, so a missing verdict artifact is the only reliable signal for these).

## Detection results

| Result | Count |
|---|---|
| Runs with a parsed verdict | 231 |
| Clean (no threat) | 231 |
| Any threat | 0 |
| `prompt_injection` | 0 |
| `secret_leak` | 0 |
| `malicious_patch` | 0 |

**Threat rate (of runs with a verdict): 0%**

## By workflow

| Workflow | Runs | Failed | Cancelled | Skipped | No verdict | Threats |
|---|---|---|---|---|---|---|
| PR Sous Chef | 86 | 0 | 0 | 0 | 1 | 0 |
| Q | 63 | 0 | 0 | 63 | 0 | 0 |
| Daily Trajectory Grader Implementer | 47 | 0 | 0 | 47 | 0 | 0 |
| Issue Monster | 47 | 0 | 0 | 4 | 14 | 0 |
| [aw] Failure Investigator (6h) | 47 | 0 | 0 | 43 | 0 | 0 |
| Avenger | 24 | 0 | 0 | 20 | 4 | 0 |
| Deployment Incident Monitor | 13 | 0 | 0 | 13 | 0 | 0 |
| Daily Go Test Parallelizer | 12 | 0 | 0 | 2 | 2 | 0 |
| Auto-Triage Issues | 9 | 0 | 0 | 3 | 4 | 0 |
| Contribution Check | 6 | 0 | 0 | 0 | 0 | 0 |
| Code Scanning Fixer | 4 | 0 | 0 | 1 | 0 | 0 |
| Deep Report | 4 | 0 | 0 | 0 | 0 | 0 |
| Design Decision Gate 🏗️ | 4 | 0 | 0 | 0 | 0 | 0 |
| Impeccable Skills Reviewer | 4 | 0 | 0 | 0 | 0 | 0 |
| Matt Pocock Skills Reviewer | 4 | 0 | 0 | 0 | 0 | 0 |
| PR Code Quality Reviewer | 4 | 0 | 0 | 0 | 0 | 0 |
| PR Triage Agent | 4 | 0 | 0 | 0 | 0 | 0 |
| Ponytail Reviewer | 4 | 0 | 0 | 0 | 0 | 0 |
| Test Quality Sentinel | 4 | 0 | 0 | 0 | 0 | 0 |
| Workflow Generator | 4 | 0 | 0 | 4 | 0 | 0 |
| Daily Credit Limit Test | 2 | 0 | 0 | 1 | 0 | 0 |
| Agent Job Health Monitor | 1 | 0 | 0 | 0 | 1 | 0 |

_171 further workflows omitted; see `stats.json`._

## Change since 2026-09-05

| Metric | 2026-09-05 | 2026-09-06 | Δ | 7-day mean |
|---|---|---|---|---|
| External detector runs | 692 | 578 | -114 | 757.6 |
| Error rate | 0% | 0% | +0 pp | 0% |
| Soft failures | 63 | 69 | +6 | 74.6 |
| Runs with verdict | 252 | 231 | -21 | 254.7 |
| Any threat | 0 | 0 | +0 | 0.14 |
| Threat rate | 0% | 0% | +0 pp | 0.05% |

## Watch list

- Issue Monster — 0 failed, 14 without a verdict, out of 47 runs
- Avenger — 0 failed, 4 without a verdict, out of 24 runs
- Auto-Triage Issues — 0 failed, 4 without a verdict, out of 9 runs
- Daily Go Test Parallelizer — 0 failed, 2 without a verdict, out of 12 runs
- PR Sous Chef — 0 failed, 1 without a verdict, out of 86 runs

Collected by: https://github.com/github/gh-aw-threat-detection/actions/runs/34081296608
Full data: the detection-stats-34081296608 artifact on that run.

> Generated by [Detection Stats Daily](https://github.com/github/gh-aw-threat-detection/actions/runs/34081296608) · copilot · auto · 29.1 AIC · ⌖ 14.3 AIC · ⊞ 11.3K · [◷](https://github.com/search?q=repo%3Agithub%2Fgh-aw-threat-detection+is%3Aissue+%22gh-aw-workflow-call-id%3A+github%2Fgh-aw-threat-detection%2Fdetection-stats-daily%22&type=issues)

Contributor guide

Open the contributing guide

Research direction

This issue is an automated daily detection-statistics report, with results collected by the Detection Stats Daily workflow. Start by reviewing the linked workflow run and its detection-stats-34081296608 artifact; no requested code change, file, test, or completion condition is identified.

Written by the indexing model from the issue text.

Assessment

Tech stack
github-actions
Domain
security
Issue type
Documentation
Difficulty
5/5
Estimated time
Over a week
Activity status
Active
Clarity
Needs clarification
Newbie friendliness
15/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.