aws-samples / aws-samples/sample-autonomous-cloud-coding-agents
RFC: Workflow graph syntax compiled to step-runner IR
- Vorherrschende Sprache
- TypeScript
- Sterne
- 143
- Forks
- 46
- Ø Merge
- 3 T. 9 Std.
- Gemergte PRs (30 T.)
- 20
Beschreibung
## Primary area
Agent (Python runtime)
## Related issue or feature request
- [WORKFLOWS.md](https://github.com/aws-samples/sample-autonomous-cloud-coding-agents/blob/main/docs/design/WORKFLOWS.md) — linear YAML steps today
- [ADR-014](https://github.com/aws-samples/sample-autonomous-cloud-coding-agents/blob/main/docs/decisions/ADR-014-workflow-driven-tasks.md)
- #457 (fix loops / `retry_target` on verify steps)
- #230 (event-driven governance — sync checkpoints for plan-before-code)
## Summary
When linear YAML steps become limiting, introduce a **workflow graph authoring format** that compiles to the existing step-runner intermediate representation (ordered steps with explicit jump metadata). The runner execution engine stays unchanged in v1; the compiler is the new component.
## Use case and motivation
Current workflows are intentionally linear with one `run_agent` ([WORKFLOWS.md](https://github.com/aws-samples/sample-autonomous-cloud-coding-agents/blob/main/docs/design/WORKFLOWS.md)). Fix loops and human gates add control flow via fields on steps, but complex flows (parallel review, plan-revise loops, conditional skip) become awkward as YAML lists.
A graph syntax makes branching, loops, and parallelism **visible in review** and diffable. Compilation to IR keeps the runtime simple and testable.
## Proposal
### Authoring
- New workflow file with format TBD instead of `steps:` in YAML.
- Nodes map to step kinds; edges carry conditions (`outcome=succeeded`, human choice keys).
- `model` / stylesheet attributes on graph or nodes compile to per-step overrides.
### Compilation
- CDK synth-time and `agent` loader invoke compiler → normalized `steps` + `control_flow` IR validated by existing JSON Schema where possible.
- Golden fixtures in `contracts/workflow-graph/` (graph input → expected IR).
### Execution (phased)
- **v1:** Compiler emits linearized steps with `jump_to` metadata; runner gains minimal jump support.
- **v2:** Parallel fan-out (explicitly out of scope for #248 / #99 today).
## Out of scope
- Visual editor / web UI for graphs
- Meta-agents that generate graphs at runtime
- Replacing YAML for simple workflows — graph is opt-in
- Multi-`run_agent` without explicit RFC approval
## Potential challenges
- **Validator parity** — Graph cross-field rules must match YAML rules (cedar-parity lesson).
- **Resume/checkpoint** — Jump targets must serialize in `workflow_state.json`.
- **Authoring burden** — learning curve; mitigate with examples in `agent/workflows/`.
## Dependencies and integrations
- `agent/src/workflow/compiler/` (new)
- `agent/workflows/schema/`
- `cdk/` synth-time validation
- `contracts/workflow-validation/` corpus extension
## Alternative solutions
1. **YAML-only control flow** — `retry_target`, human gates via #230; sufficient for medium term.
2. **Python DSL** — More expressive but not reviewable by non-Python authors.
3. **Registry-stored graphs** — Phase 4 #246; local compiler still needed for dev.
---
**Note:** Non-triaged RFCs may not get timely review. PRs on non-triaged issues might not be accepted.
Beitragsleitfaden
Rechercherichtung
Lies zuerst docs/design/WORKFLOWS.md und ADR-014, und untersuche dann die vorhandenen Strukturen in agent/workflows/schema/ und contracts/workflow-validation/. Die vorgeschlagene Arbeit würde agent/src/workflow/compiler/ und Golden-Fixtures unter contracts/workflow-graph/ hinzufügen; sie gilt als abgeschlossen, wenn das Graphformat und das Kompilierungsverhalten ausreichend genau spezifiziert sind, um validierte step-runner-IR zu erzeugen, ohne die Ausführungs-Engine über den angegebenen v1-Umfang hinaus zu ändern.
Vom Indexierungsmodell aus dem Issue-Text verfasst.
Bewertung
- Tech-Stack
- python, typescript
- Bereich
- build-system, compilers, tooling
- Issue-Typ
- Feature
- Schwierigkeit
- 5/5
- Geschätzter Aufwand
- Über eine Woche
- Aktivitätsstatus
- Ruhig
- Klarheit
- Muss geklärt werden
- Anfängerfreundlichkeit
- 30/100