Chat with an orchestrator and an army of specialists — research, planning, design, dev, test, security, docs. Say the word, and it comes back with a PR, a demo video, a live link, and a plain-English list of anything that needs your call.
Local, OpenRouter, or your existing Claude / ChatGPT subscription or API key. Nothing is locked in.
Memory that actually forgets
Decisions persist; superseded facts get scored on recency, source, and trustworthiness — then tombstoned instead of piling up.
A real audit trail
Every decision, human or agent, traces back to its source. Claims need evidence. Conflicts get surfaced, not buried.
Spend control to the token
Budgets per project, user, model, or time window — with a live dashboard that drills into individual conversations.
Mission control
A live runtime console, session trees for parent/child runs, intent approvals, run status, and proof of what got produced.
Cross-model contracts
Define handoff rules so Codex and Claude work in tandem under terms you set — and review each other's work.
Bring your own tracker
Linear, Shortcut, Azure DevOps, Jira, or the built-in board. Every work item carries agent, model, effort, and priority.
API / CLI / MCP
For people who'd rather prompt it than watch a dashboard all day — or just check progress from the beach.
Spend$86.40 of $120 · July
Live budget drill-down, per conversation.
Work board — proof, not promises
Agent studio — durable agent identities
The part that changed how I work
Turns out AI is lazy. So assay checks its own work.
Anything risky or non-trivial runs through a self-critique loop: five specialist agents have to hit 97/100 before a second model gives a second opinion — and it sends work back for revisions about 6 out of 10 times. Ten to twenty agents doing this in parallel burns tokens, but a lot less than the hours you'd spend finding the bug yourself.
And they do it all while you sleep.
run/7f3a · pass 2running
Critique score97 / 100
Sent back
6 / 10
revision rate
Second opinion
GPT ⇄ Claude
cross-model review
early access
Stop supervising. Start delegating.
Bring your own model and your messiest backlog. You'll get notified when the work is done — or when the only thing blocking it is you.