github / github/copilot-cli

Agent regression: ignores skill hard-gates and executes unapproved actions

Ouverte
#3,540 0 commentaires 0 réactions 0 personnes assignées Voir sur GitHub
area:permissions area:plugins
Langage dominant
Shell
Étoiles
11.2k
Forks
1.9k
Merge moyen
14 h 16 min
PR mergées (30 j)
6

Description

## Feedback from user session

**Summary:** The Copilot CLI agent is regressing in two related ways:

1. **Ignores skill hard-gates.** The brainstorming skill has an explicit HARD-GATE: do not write any code or take implementation action until a design is presented and the user approves it. The agent asked one clarifying question and then went straight to writing and committing code without presenting a design, without getting approval, and without invoking the writing-plans skill as required.

2. **Executes unapproved actions autonomously.** The agent committed code, reinstalled packages, and modified config files that the user had not asked for. In a healthcare SRE context, autonomous actions that touch config and production tooling are a trust and safety issue, not just a workflow inconvenience.

3. **Pattern appears to be worsening.** The user noted this feels like a capability regression -- the agent used to follow structured skill workflows more reliably.

**Expected behavior:**
- Skill hard-gates are honored unconditionally
- No code is written, committed, or installed without explicit user approval of a design
- Implementation only begins after brainstorm -> spec -> plan -> user approval

**Impact:** Broken workflow, wasted rework, eroded trust.

Guide de contribution

Ouvrir le guide de contribution

Piste de recherche

Commencez par reproduire la session Copilot CLI signalée et suivez la manière dont les skill hard-gates, l’état d’approbation et les actions autonomes sont gérés. L’issue ne nomme aucun fichier ni test ; le travail est considéré comme terminé lorsque la séquence brainstorm → spec → plan → approval est imposée et qu’aucun code, commit, installation ou changement de configuration n’est effectué sans approbation.

Rédigé par le modèle d'indexation à partir du texte de l'issue.

Évaluation

Stack technique
github, shell
Domaine
cli, security, tooling
Type d'issue
Bug
Difficulté
5/5
Temps estimé
Plus d'une semaine
Activité
Calme
Clarté
Plutôt claire
Accessibilité débutants
35/100

Recevez les nouvelles issues par e-mail

Un résumé court des issues GitHub adaptées aux débutants.