the toolkit · open-source · MIT
The tools your coding agentcan’t build alone.
Your harness already plans, swarms, and ships. wicked-garden refuses to wrestle it and fills only the gaps a planner-executor can’t fill alone — starting with the one that matters most: done is re-derived from evidence, never asserted.
load out garden and its peers — one command
interactive · picks your products across CLIs · direct install below
- 94 skills
- 12 domains
- 10 work-shapes
- MIT · local-first
01 / the toolbox
Six gaps your agent can’t close alone.
Your harness plans, swarms, and ships. These are the six things a planner-executor genuinely can’t do on its own — and they’re a sample: the full kit runs to 94 skills across 12 domains. It plays itself; click any tool to pin it.
Re-runs the proof behind the claim. A false pass is REJECTED; a missing backend FAILS CLOSED. Never a vacuous green.
02 / the gate
Play the lying agent. Watch the gate refuse.
Garden’s one non-negotiable, made drivable. It plays itself — breaking the evidence and re-stamping the verdict. Flip a switch or pull the lever to take control: the gate re-derives the claim instead of taking your word for it.
The agent stamped this “done.” The gate re-runs the verifier, re-hashes the recording, and checks the vault is even there. Break a condition and prove it can’t be fooled.
03 / the whole toolkit
Six tools were the sample. Here’s the rest.
The toolbox shows the signature gap-fillers. Underneath sits the full surface — 94 skills across 12 domains, 34 of them fork/worker skills that run in isolated subagent contexts — all reading the same evidence-first discipline. Everything below is real and in the repo today.
Vague ask → SMART criteria, UX & a11y review, mockups, visual direction, user-signal synthesis.
The rubric on demand — security, compliance, incident, infra, distributed traces, CI.
Graph-driven change — rename & patch across files, debug from a trace, architecture, docs.
Review the agentic codebase itself — topology, framework detection, trust & safety.
Orchestration workers — implement, research, review; swarm, worktrees, workflow runners.
Pull-model context — briefing, intent, propose-skills, classify a request, ground it in the repo.
A second opinion that isn’t self-grading — an independent external-model panel and facilitator.
Evidence, re-derived — prove re-runs the proof; a semantic reviewer checks intent, not just diffs.
ETL, data-quality, ontology, and ML workflows under the same evidence discipline.
Run any task under a named behavioral profile — a reusable review cast on demand.
Impact & lineage over the real graph — the injected edges grep and a static call-graph can’t see.
Ten shapes read off each prompt — a typo to a cutover; deliberate challenges the ask first.
One install bundles the family. Every peer is an opt-in layer you adopt when you want it — the kit works without any of them. The evidence backend the gate re-derives against rides inside wicked-testing, not a separate install.
04 / the bench
One command. The whole loadout.
The family installer is the fastest way in — one interactive command that picks your products across every CLI and installs the shared wicked CLI. Prefer just this plugin? The direct path is right below.
interactive · picks your products across CLIs · ships the wicked CLI
or install just wicked-garden directly
MIT · open-source · v12.27.0 · local-first — nothing leaves your machine.