Repository navigation
Conversation
|
Hey Bastian, here are some of my thoughts after reviewing the PR: Control/LifecycleI think we should flip the order of control around. From what I can tell, in this stack, the playbook spawns agents and creates new subagents to execute tasks. That puts the playbook above the agent in terms of lifecycle and control. I think it should be the opposite, an agent can run multiple playbooks sequentially through its lifespan. Or, it can be spawned specifically to run a playbook by itself. That means the playbook lifecycle is <= the agent lifecycle. I think having a playbook manage the lifecycle of agents is an interesting idea, but I think we should play with playbooks in the other form for a while before we expand their responsibility. A more controlling/complex feature tends to encode more opinions and be less flexible, so starting with the simpler one lets us evolve it as we discover more. Code Steps & TransitionsThis goes along with the simplicity argument I made above, but I think we should start out playbooks as a simpler tool that allows agents to execute complex workflows, but doesn't try to be a workflow engine. I think, with where we are in the new cycle of agent capabilities that's come out with Astra and Fable, we are at a point where people are better served by just using the top models again (and not trying to bring in the older models that need more steering and coordination). Therefore, I think we all should focus on UX and tooling that helps the most capable agents when we're building J5. From a UX standpoint for Playbooks, I think that means:
From the agent tools perspective, I think that means:
SimplicityThis isn't just a normal coding best practice, I think given how fast T3 is moving and how they seem to be going into some of the areas we are building, keeping things simple is going to make it easier for us to integrate as they evolve. And we 100% want to be able to do this so we can spend more time on features and less time on the stuff they've already solved. |
The external Aurora Playbooks need review feedback and approval-controlled publication loops.
Extend YAML execution and publication adapters to collect PR feedback, revisit review and rework phases, and preserve explicit publication approval. Keep diagnostic persona execution read-only and document the external Playbook source.
External definitions are proposed separately in First-horizon/agent-skills#1: https://github.com/First-horizon/agent-skills/pull/1.
Stack 9/9. Depends on #163. Merge in order, retargeting to
j5/mainas predecessors land. Full-stack testing branch:codex/playbooks-09-aurora-loops.Validation at the integrated stack tip: 179 server/shared tests and 104 web tests passed; server and web typechecks passed. Each regenerated runtime manifest was checked. The adapter slice independently passed 92 tests, including the parser regressions.
Review order:
Model: GPT-5 · Harness: Codex