Plan traceable work bullets that connect intent, evidence, acceptance, and proof gaps. Use when an idea or reframe needs decomposition into traceable work units before tracker or…
Use when the user asks about Katie Roberts's talk \"Stop Maintaining, Start Evolving: Applying AI-Native Engineering in Brownfield Codebases\" (AI Native DevCon, June 2026) —…
Turn Slack #code-fixes, CodeRabbit, Codex Review, CI, and check-status noise into a repo-and-PR action queue.
Analyze, design, or triage LLM evaluation workflows. Use when the user asks for evaluator design, error analysis, judge prompts, RAG evals, synthetic data, or review tooling.
Review the completed slice for correctness, risk, evidence gaps, and next-stage blockers. Use when the user asks for review, PR readiness, risk assessment, blocker identification,…
Use when the user asks about Lieven Scheire's talk \"Artificial Intelligence\" (a Belgian physicist/comedian's keynote on AI for a developer audience) — including questions about…
Plan tracker-ready slices with dependencies, labels, acceptance fields, and validation placeholders. Use when trace work needs Linear-ready or tracker-ready issue structure…
Define problem scope, requirements, and decision options before spec or plan stages. Use when the user has ambiguity in what to build, why it matters, or which direction to choose.
Analyze recent Codex session evidence for repeated manual workflows and route them to skills, subagents, validators, or no artifact when Jamie asks what he keeps doing manually.
Use when a project or Codex runtime needs environment TOML created, triaged, or updated with safe setup, actions, exec-server providers, and validation evidence.
Explains Kevin Groetzinger's Skills Everywhere talk and helps teams operationalize reusable skills: trigger design, ownership, discoverability, maintenance, quality review, and…
Analyze and validate compound Harness Engineering run state, blockers, validation status, and Linear context.
Use when the user asks about Dave Kerr's AI Native DevCon talk on bipolar disorder, dysregulation, and AI, including responsible interpretation of personal and clinical themes…
Use when the user asks about Peter Wilson and Davide Eynard's AI Native DevCon talk on cq, a Stack Overflow-like knowledge commons for agents, local/team/public knowledge sharing,…
Analyze options and trade-offs before traceable work is ready, without pretending exploration is a plan.
Routes old phase-heartbeat requests into HE phase work: check the approved plan, reuse or create a 10-minute continuation only with authority, run phase gates, delegate the active…
Explains the Product Brain talk and helps design curated product-memory systems for AI-assisted product work: knowledge structure, provenance, synthesis cadence, ownership, and…
Assists with questions about a practitioner panel talk titled 'From Pipelines to Prompts: Surviving the Shift to AI' featuring Stephane Jourdan, Simon (Saxo Bank), and Samantha.
Use when the user asks about Daniel Jones (Deejay) and Tomasz's talk \"More software, faster — Odevo's AI Native transformation\" — including questions about how Odevo (Sweden's…
Review diffs, PRs, specs, plans, or review-feedback items and return severity-ranked engineering findings with exact locations.
Analyze, compare, and recommend a Codex build primitive. Use when the user is packaging or automating a workflow and the right primitive is unclear.
Generate closure-grade HE eval and drift proof for one execution slice. Use when Linear, milestone, or source-prompt closure needs validation evidence.
Use when the user asks about Dana Lawson's talk \"Built for Humans. Now Agents Are Here.\" (Netlify CTO, 2026) — including questions about Agent Experience (AX), the AX paradox,…
Use when the user asks about Macey Baker and Baruch Sadogursky's workshop-style AI Native DevCon session on turning repeated agent work into skills, rules, scripts, hooks, and…
Review services, APIs, and multi-component systems for reliability risks including failure modes, cascading failures, resilience gaps, and SLO readiness.
Answers questions about Brian Douglas's talk on training AI on your own code. Use when a user asks about Brian Douglas's pipeline for capturing agent sessions, extracting skills…
Use when the user asks about Amit Kushwaha's AI Native DevCon talk on benchmarking agent-era systems, measuring performance beyond single LLM calls, inference, workflow…
Diagnose, fix, and validate mise runtime failures. Use when commands fail from mise config, missing runtimes, stale pins, trust prompts, or shell activation drift.
Use when creating, installing, validating, folding, or troubleshooting Codex custom subagent role TOML and discoverability config.
Create, install, validate, and orchestrate Codex custom subagents as standalone TOMLs with canonical global defaults (`~/dev/configs/codex/agents/{name}/{name}.toml`,…
Explains Robert Overweg's One Brain, No Filtering talk and helps design safe knowledge-memory systems: context maps, retrieval rules, provenance labels, local knowledge-store…
Remove AI slop and corporate jargon from text without applying a personal voice. Use when the user asks to "unslopify", "remove AI slop", "deslopify", "clean up AI writing",…
Check if a repository or agent-facing product surface is ready for AI coding agents. Use when you need to audit repo agent compatibility, review AGENTS.md, find missing test/build…
Explains, summarizes, and turns Matthias Luebken's talk on embedding Pi-style coding agents into safe product-design artifacts: tool-contract sketches, guardrail checklists,…
Create or refactor AGENTS.md and linked instruction docs using progressive disclosure. Use when the user wants repo-specific agent guidance organized, deduplicated, or routed…
Deepen an existing system or UI spec so boundaries, lifecycle rules, failure handling, and validation are strong enough for planning.
Plan how to implement one approved slice, including file targets, validation gates, rollback, and proof lanes.
Audit, validate, and troubleshoot Agentation integrations in frontend apps. Use when annotations, MCP registration, endpoint sync, webhook delivery, or watch mode readiness are…
Analyze stale or conflicting lifecycle state by separating local, PR, CI, review, tracker, artifact, and session truth.
Explains Jack Wotherspoon's Humans vs Slop talk and helps create quality gates for AI-heavy software work: review-cost analysis, slop detection heuristics, durable-value metrics,…
Build only the approved slice while preserving scope, unrelated worktree changes, and validation evidence.
Generate, validate, and refresh @brainwav/diagram architecture artifacts when repo diagrams, context packs, PR impact, or CI drift evidence is needed.
Use when the user asks about Maximiliano Firtman's (\"Maxi\") AI Native DevCon talk on Web MCP and the agentic web — including questions about how Web MCP differs from MCP,…
Write Harness Engineering specs before planning. Use when a feature, QA report, Linear issue, or UI source needs a clear WHAT contract.
Answers questions about, summarizes, and applies May Walter's AI Native DevCon talk \"From Blind Spots to Merged PRs\" on runtime intelligence for coding agents.
Create one selected trace or tracker slice spec with acceptance criteria, constraints, validation, and exit conditions.
Selects the correct Harness Engineering lifecycle stage and compatibility alias route. Use when a request is ambiguous, mixes brainstorm/spec/plan/work/review intent, references…
Assists with questions about Guy Podjarny's talk \"Skills are the new Code\". Use when the user wants to understand, apply, audit, or explore frameworks from this keynote —…
Deepen an existing implementation plan so sequencing, verification, and risk treatment are strong enough for execution.
Explains and applies Liran Tal's AI Native DevCon talk on skills security. Use for questions about skill supply-chain risk, AI tool vetting, third-party plugin risks, provenance…
Restore broken behavior by reproducing failures, identifying root cause, and delivering verified fixes.
Scan Codex session history for skill failures, usage patterns, and coverage gaps. Use when the user wants daily skill-health monitoring or evidence-backed recommendations about…
Answers questions about Lamis's (Anthropic) AI Native DevCon talk on context engineering, agent memory systems, and dreaming — an asynchronous, out-of-band memory-curation…
Use when the user asks about Ian Thomas's talk \"AI Native Engineering\" (Meta / Reality Labs / Horizon Experiences) — including questions about Meta's AI4P (AI For Productivity)…
Plan a stuck, stale, or unsafe frame into a smaller viable direction with rollback and decision boundaries.
Create a durable guardrail, validator, doc, eval, or skill-reference update from a verified failure or learning.
Generate closure proof with exact commands, outcomes, artifact paths, untested lanes, and residual risk.
Use when the user asks about Shachar Azriel's AI Native DevCon talk on executable specs, verification layers for agentic coding, planner/verifier separation, requirement mapping,…
Route skill lifecycle requests to a Skill Factory lane. Use when users ask to create, harden, install, audit, or skillify skills.
Review PRs, branches, diffs, and workflow artifacts for package-level go/no-go readiness with severity-ranked synthesis.