Analyze session work and automatically convert reusable patterns into Claude Code skills. Use when: "세션을 스킬로", "스킬 만들어", "이거 스킬로", "skill factory", "이 작업 자동화해", "스킬 추출", "make…
Use when the user asks about Shaun Smith's AI Native DevCon talk on MCP transports, remote Streamable HTTP servers, stateless protocol direction, Hugging Face MCP adoption, and…
Capture a completed Codex workflow as a reusable SKILL.md package by analyzing session context plus optional session-collector evidence, interviewing the user with structured…
Compress HE cognition artifacts into evidence-backed strategy. Use when intent, review, triage, ADR, core, or source-prompt comparison evidence needs durable direction.
WHAT: Route plugin-factory requests to the right lane. WHEN: Use when plugin creation, building, installation, review, or routing is broad, mixed, or under-specified.
Design, review, and validate Codex app automations when recurring background workflows need safe scheduling, scope, preflight, and consolidation.
Improve existing Harness Engineering implementations or workflows with evidence-backed changes. Use when users ask for targeted enhancement of shipped or drafted work.
Convert approved HE cognition into small live-ready Linear execution tracking. Use when strategy, reframe, plan, bug, or source-prompt evidence needs scoped issue, milestone, or…
Build and audit polished interaction refinements for existing React or Tauri UI when motion, accessibility, reduced-motion, and browser-verified behavior need focused improvement.
Analyze broad frontend design requests and route them to the correct local UI skill after classifying intent and maturity.
Review, create, and validate Bash scripts when shell work needs strict mode, quoting safety, portability, or interpreter-compatible behavior.
Assists with questions about Hannah Foxwell's talk 'The Reinvention of the Dev Team'. Use when a user asks about Foxwell's arguments on agentic software development, engineering…
Run, audit, and design authorized Recon Workbench workflows when scoped target interrogation needs evidence artifacts, redaction, validation, and safe reporting.
Analyze broad, mixed, or unclear Plugin Factory follow-up requests and select the correct plugin lane. Use when plugin intent lacks a clear lane owner.
Use when the user asks about Luke Marsden's talk \"Giving Every Agent Its Own Desktop: Lessons from Dogfooding HelixML\" — including questions about HelixML, giving each agent its…
Answers questions about, applies frameworks from, and generates artifacts based on Tammuz Dubnov's talk \"When Our PM Started Writing Code: What Merge Rate Taught Us About AI…
Diagnose Codex Desktop or CLI local-state bloat and safe recovery options. Use when sessions, archived history, logs, worktrees, or stale Codex config may be making Codex feel…
Answers questions about, summarises key insights from, and helps apply concepts from Simon Martinelli's talk \"Lessons from Spec-driven Development\" — providing verbatim-grounded…
Generate, review, and refine high-retention technical YouTube hooks, outlines, and scripts. Use when the user wants video scripting tailored to a topic, audience, runtime, and…
Use when the user asks about Don Syme's AI Native DevCon talk on Continuous AI, GitHub agentic workflows, repository automation, and safe developer-controlled automation loops.
Summarizes Simon Maple's AI Native DevCon welcome and explains the conference framing: context window, latent space, tool pool, hallway track, attendee goals, and practical…
Build behavior-safe code changes with TDD and RED/GREEN evidence. Use when he-plan or he-work requires TDD for a concrete behavior target.
Summarizes, explains, and answers questions about Dave Farley's talk 'Vibe Coding — Is this really the best we can do?', including key arguments, frameworks, and recommendations.
Use when the user asks about Justin Cormack's AI Native DevCon talk on tests, observability, AI-generated behavior, evidence, instrumentation, and keeping AI systems honest when…
Create, review, or repair recurring Harness Engineering heartbeats that wake a thread, re-check live state, and route back into the right HE stage; use when PRs, CI, reviews,…
Ship skill changes to PRs when Codex skills need source edits, rooted sync, strict audit, reviewer evidence, commit, push, and PR status.
Summarizes, explains, and applies Lars Trieloff's AI Native DevCon talk on browser-native agents. Use for browser agents, running AI in the browser, browser-as-runtime…
Review, triage, and validate visual regression diffs. Use when the user wants snapshot-change analysis, layout regression evidence, Storybook diffs, Playwright screenshots, or…
Answers questions about, summarises key insights from, and applies the security guidance of Joseph Katsioloudes's talk 'Code Security Reinvented: Navigating the era of AI'.
Use when the user asks about Simon Obstbaum and Rob Willoughby's AI Native DevCon talk on measuring AI agents, output evals versus trajectory evals, instrumentation, compliance,…
Route ambiguous Harness Engineering requests to one lifecycle stage when users ask where to start, resume, plan, implement, review, debug, schedule a heartbeat, or resolve domain…
Explains James Moss's team-skills workflow and helps design skill governance: decomposition, ownership, versioning, eval scenarios, quality review, and lifecycle maintenance.
Automate until-green PR review, CI, merge, and cleanup follow-through. Use when open project PRs need GitHub/gh, CodeRabbit, CircleCI, Context7, Snyk, autofix, heartbeat, and…
Create or refresh evidence-bound Harness Engineering learning artifacts from verified solved problems.
Create, review, and validate an alignment checkpoint. Use when a request is ambiguous, high-stakes, multi-step, or requires explicit approval before tool use.
WHAT: Generate local Codex usage reports. WHEN: Use when users ask for usage analytics, weekly insights, session summaries, telemetry patterns, or prompting help.
Runs approved Harness Engineering plans in recurring phases: check live state, confirm continuation authority, execute only the active he-work slice, verify gates, update Linear…
Explains Oleg Selajev's Docker Sandboxes talk and helps design safe, conceptual agent-isolation policies: file-sharing boundaries, network policy, secret isolation, audit…
Use when the user asks about Patrick Debois's talk \"Coding Agents Don't Scale Themselves. Neither Do Your Teams.
Create and validate implementation-grade CLI specifications when command trees, JSON contracts, dry-run plans, errors, or agent-ready behavior need a binding spec.
Create evidence-backed HE reframe migration programs. Use when structural drift, routing ambiguity, or source-prompt gaps need phased rollback-safe execution.
Use when a Codex goal/task is stuck, hanging, not finishing, or needs status. Reads goal.md, state.yaml, receipts.jsonl; syncs reported status with board files; fixes invalid…
Use when the user asks about Steve Ruiz's AI Native DevCon talk on tldraw, Make Real, annotations as prompt input, canvas workflows, tldraw computer, and agents collaborating on…
Summarizes Rob Sloan's harness-engineering talk and creates safe design artifacts for agent context beyond code: product-intent packets, design constraints, acceptance criteria,…
Use when you need focused cleanup audits, safe removals, scoped quality-risk reductions, and evidence-backed cleanup plans before touching code.
Install, update, audit, diagnose, and explain @brainwav/coding-harness when repository governance, harness init, CI migration, or action-sync needs live command evidence.
Review, configure, and troubleshoot prek hooks when users need prek.toml edits, shim installs, hook validation, or pre-commit migration help.
Create, repair, and validate uv Python project setup. Use when initializing Python apps or libraries, managing uv dependencies, virtual environments, or CI-ready uv workflows.
Create, review, and maintain gold-standard Skills SDK eval scenarios before internal evals, dry Tessl staging, or live private Tessl scoring.
Review and prune stale branches safely. Use when branch cleanup needs evidence, protected-branch caution, PR awareness, and non-destructive recommendations.
Use when the user asks about Katie Roberts''s talk \"Stop Maintaining, Start Evolving: Applying AI-Native Practices to Brownfield Codebases\" — including questions about using AI…
Implement approved Harness Engineering work. Use when a plan, todo list, or tiny spec needs traceable delivery and validation.
Use when the user asks about Ryan Lopopolo's AI Native DevCon talk on harness engineering, steering coding agents with goals, constraints, context, tool scope, eval loops, and…
Run a bounded Harness Engineering lifecycle across multiple stages. Use when the user wants coordinated brainstorm, spec, plan, work, review, and fix flow rather than one isolated…
Install, repair, and validate Vale prose linting. Use when users need Vale config, style sync, docs lint gates, or broken Vale workflow diagnosis.
Plan execution work from specs, brainstorm outputs, bugs, or feature requests into an implementation-ready sequence.
Analyze, review, and plan architecture alternatives through a structured interview. Use when the user needs tradeoffs surfaced before implementation or a Linear decision note…
Generate and compare grounded product or engineering directions with tradeoffs. Use when users want possibilities, critique, or direction-setting before a spec.
Run, plan, and validate pnpm workspace operations. Use when a user needs pnpm monorepo installs, tests, builds, filters, changed-package selection, or publish routing.
Use when creating, auditing, upgrading, or validating Codex hook packs, hooks.json files, hook scripts, or repo-local/user-level .codex hook installs.