2.8 KiB
Priorities
The ranked work queue for the autonomous improvement loop. The architecture-review pass (the architect) owns this file: each run it turns the roadmap plus an internal scan (gaps in the services → agents → workflows lifecycle, API coherence, drift, tech debt, test and DX friction) into a single ordered list — highest-value first — and links each item to a tracking issue. The hourly continuous-improvement pass works the top item whose issue is still open. So the architect decides what, and the increment loop builds it.
Reading / editing. An item is done when its linked issue closes (the increment
that builds it adds Closes #<issue>). Roadmap phase (Now → Next → Later) is the
primary ordering; internal findings are interleaved by value, not kept in a
separate list. The human can reorder this list — or the issues — at any time to
redirect the loop; direction always wins.
Off-limits to the loop (the architect proposes these as notes, never as queue items the loop can auto-merge): brand/positioning copy, breaking public-API changes, architectural rewrites. Those go to the human.
Work queue (ranked)
- Surface the first-agent and 0→hero example paths in the CLI (#3983) — #3986 closed the live plan/delegate notify regression, and there are no open
codexPRs in flight, so the queue should return to the current adoption goal. The README, website docs, maintained examples, and 0→hero harness now describe the on-ramp, but users still need to discover those paths directly after install; a CI-verifiable CLI wayfinding task is the highest-value next step for scaffold → run → chat → inspect. - Broaden provider streaming and keep chat/A2A streaming end to end (#3903) — streaming remains the highest developer-visible Next-phase seam after the current wayfinding gap. Real chat and long-running A2A tasks need token streaming to stay coherent from provider →
ai.Stream→micro chat→ A2Amessage/stream, with mock/default CI coverage plus key-gated live provider checks and safe fallback for non-streaming providers. - Trace agent runs as OpenTelemetry spans (#3908) — the blog/README/roadmap story promises an operable harness, and the developer on-ramp now includes chat, inspect, and run-history checkpoints. The next observability gap is production-grade trace correlation for
RunInfo: steps, tool calls, delegation, status, durations, and failures should be visible as spans while defaulting to no-op when tracing is not configured. Seeded by Claude Code from the roadmap + open issues; thereafter maintained by the architecture-review pass.