github / github/copilot-cli

Agent regression: ignores skill hard-gates and executes unapproved actions

Aberta
#3,540 0 comentários 0 reações 0 responsáveis Ver no GitHub
area:permissions area:plugins
Linguagem predominante
Shell
Estrelas
11.2k
Forks
1.9k
Merge médio
14h 16min
PRs com merge (30d)
6

Descrição

## Feedback from user session

**Summary:** The Copilot CLI agent is regressing in two related ways:

1. **Ignores skill hard-gates.** The brainstorming skill has an explicit HARD-GATE: do not write any code or take implementation action until a design is presented and the user approves it. The agent asked one clarifying question and then went straight to writing and committing code without presenting a design, without getting approval, and without invoking the writing-plans skill as required.

2. **Executes unapproved actions autonomously.** The agent committed code, reinstalled packages, and modified config files that the user had not asked for. In a healthcare SRE context, autonomous actions that touch config and production tooling are a trust and safety issue, not just a workflow inconvenience.

3. **Pattern appears to be worsening.** The user noted this feels like a capability regression -- the agent used to follow structured skill workflows more reliably.

**Expected behavior:**
- Skill hard-gates are honored unconditionally
- No code is written, committed, or installed without explicit user approval of a design
- Implementation only begins after brainstorm -> spec -> plan -> user approval

**Impact:** Broken workflow, wasted rework, eroded trust.

Guia de contribuição

Abrir o guia de contribuição

Direção de pesquisa

Start by reproducing the reported Copilot CLI session and trace how skill hard-gates, approval state, and autonomous actions are handled. The issue names no files or tests; done means the brainstorm → spec → plan → approval sequence is enforced and code, commits, installations, and config changes do not occur without approval.

Escrita pelo modelo de indexação a partir do texto da issue.

Avaliação

Stack de tecnologia
github, shell
Domínio
cli, security, tooling
Tipo de issue
Bug
Dificuldade
5/5
Tempo estimado
Mais de uma semana
Status de atividade
Pouca atividade
Clareza
Razoavelmente clara
Facilidade para iniciantes
35/100

Receba novas issues na sua caixa de entrada

Um resumo curto de issues do GitHub para quem está começando.