Auto mode using a safety classifier for tool calls
Personne n'a encore pris cette issue.
Évaluation
- Difficulté
- 5/5
- Temps estimé
- Plus d'une semaine
- Accessibilité débutants
- 30/100
Piste de recherche
Aucun fichier, test ou point d’entrée concret n’est indiqué. Commencez par localiser les modes edit et plan existants ainsi que la fonctionnalité qui permet d’exécuter automatiquement les commandes approuvées. Le travail serait considéré comme terminé avec la définition d’un Auto Mode qui évalue les appels d’outils à l’aide d’un classificateur de sécurité configurable par l’utilisateur et évite les approbations manuelles inutiles.
Rédigé par le modèle d'indexation à partir du texte de l'issue.
Description
Feature Description
Instead of just having edit and plan modes, introduce a new mode called "Auto Mode". This would use an automatic safety classifier for things like bash commands and such, similar to what Claude Code has, so that users don't have to constantly babysit the harness when it's running commands. It would likely be powered by a cheap open source model, like DeepSeek V4 Flash or Qwen 3.8 Flash to keep costs low, but the classifier model should be able to be changed by the user at any point.
Use Case
Allows for users who are working to step back while the agent does its own thing, instead of having to manually accept bash commands and the like. There is already functionality for allowing the agent to run specific commands on its own, after it's been allowed, but I find that often doesn't work very well for one reason or another. In workflows where the user may be doing multiple things at once, it can get annoying to have to go back and allow things in order to stop the agent from hanging on a single bash command request for potentially minutes on end while the user isn't aware it's even there.
Additional Context
No response
How important is this to you?
Important for my workflow
- Langage dominant
- Aucune donnée de langage
- Étoiles
- 4k
- Forks
- 350
- Métriques de merge des PR
- Aucune PR mergée en 30 j
Guide de contribution
Aucun guide de contribution indexé pour ce dépôt
Par où commencer
- Lisez l'issue en entier, puis le guide de contribution du projet.
- Signalez en commentaire que vous la prenez — cela évite que deux personnes fassent le même travail.
- Forkez le dépôt et travaillez sur une branche.
- Ouvrez une pull request qui référence le numéro de l'issue.
Autres issues de CommandCodeAI/command-code
-
Difficulté 2/5 1-3 heures Accessibilité débutants 68/100
CommandCodeAI/command-code#855 ·
-
Difficulté 2/5 1-3 heures Accessibilité débutants 78/100
CommandCodeAI/command-code#841 · 1 commentaire ·
-
Difficulté 2/5 1-3 heures Accessibilité débutants 68/100
CommandCodeAI/command-code#655 · 1 commentaire ·
-
Difficulté 2/5 1-3 heures Accessibilité débutants 68/100
CommandCodeAI/command-code#608 ·
-
Difficulté 3/5 1-2 jours Accessibilité débutants 70/100
CommandCodeAI/command-code#893 ·
Toutes les issues de CommandCodeAI/command-code
Issues similaires
-
bug-unconfirmed
Difficulté 2/5 1-3 heures Accessibilité débutants 76/100
-
enhancement
Difficulté 2/5 1-3 heures Accessibilité débutants 68/100
JuliusBrussee/caveman#1102 · 1 commentaire ·
-
Difficulté 2/5 1-3 heures Accessibilité débutants 78/100
use-agent-os/agent-os#3263 ·
-
[Bug]: context-limit error parsing has no pattern for llama.cpp's "context size (N tokens)" phrasing Ouvertearea/compression area/local-models area/sessions comp/agent duplicate P2 sweeper:risk-session-state type/bug
Difficulté 2/5 1-3 heures Accessibilité débutants 82/100
NousResearch/hermes-agent#117793 · 1 commentaire ·
-
possible bug
Difficulté 2/5 1-3 heures Accessibilité débutants 88/100
Mintplex-Labs/anything-llm#6415 · 1 commentaire ·