Skills

Skills

36 skills across 10 lifecycle groups. Each is a folder with a SKILL.md that follows the Role → Context → Workflow (Step 0 TodoWrite) contract. This table is generated from the real skill frontmatter at build time.

accessibility

SkillWhat it does
a11y-devAccessibility-first development assistant — applies WCAG 2.1/2.2 Level AA semantics, keyboard support, ARIA, and focus management while generating UI code, instead of retrofitting it afterward.
a11y-fixInvestigate and repair reported WCAG 2.1/2.2 AA defects — reproduce the original scenario, prove repository and source ownership, test root causes, apply a bounded native-first fix, and verify the exact build and required assistive-technology behavior. Explicit non-fix outcomes when evidence is missing.
a11y-reportConsolidate findings from a11y-scan, a11y-review, a11y-fix, and a11y-verify into a single shareable accessibility report — Markdown by default, or a self-contained filterable HTML file — with a WCAG severity/criterion summary. Can optionally file GitHub issues for open findings, gated on explicit approval.
a11y-reviewDeep, mode-based accessibility review of a live page beyond automated scanning — interactive widgets (keyboard/ARIA/focus), color contrast, color-only meaning, link purpose, display modes (dark/high-contrast/forced-colors), and responsive viewports. Uses Playwright + axe. WCAG 2.1/2.2 AA.
a11y-scanScan for WCAG 2.1/2.2 Level AA violations — statically over source (markup/components/styles) or at runtime against a live URL using axe-core. Produces a prioritized, deduplicated violation report with WCAG criteria and locations. WCAG-only; no vendor-specific standard.
a11y-test-genGenerate Playwright accessibility tests with axe-core for a component or page — automated WCAG scans plus keyboard-navigation, focus-management, and ARIA/state assertions — so accessibility regressions are caught in CI. WCAG 2.1/2.2 AA.
a11y-verifyVerify an accessibility repair on the exact candidate build against the original failing scenario, WCAG checks and required browser/assistive-technology matrix. Returns Fixed, Not fixed, Fixed but new issues, Blocked or Inconclusive with evidence; axe alone is not proof.

architecture

SkillWhat it does
architecture-docProduce or update a lightweight, current architecture document — components and responsibilities, key interfaces/contracts, data model, and how non-functional requirements map to design decisions — grounded in the actual code (via graph-repo). Realizes EASE-MAS A6.
graph-repoBuild a queryable code graph of a repository — modules, files, and symbols (functions/classes/exports) with import, call, and dependency edges — exported as JSON plus a Mermaid/HTML visualization. Powers impact analysis for review, refactor, debugging, and bounded repair.
threat-modelProduce a STRIDE-based threat model for a system or feature — assets, trust boundaries, data flows, threats per element, and prioritized mitigations mapped to controls — grounded in the actual architecture. Feeds security-review and the design of safe changes.

debugging

SkillWhat it does
analyze-bugTurn a raw failure (CI failure, crash, or report) into a structured, deduplicated bug — extract error, stack, logs, artifacts, correlation to the recent diff, and reproduction steps — and link it to the failing test/requirement in the traceability graph. Realizes EASE-MAS A15/A17.
fix-bugRepair a defect under bounded autonomy — take ranked hypotheses, localize and patch, then validate through unit, targeted E2E, regression, and security before an independent repair-validator gate, and open a draft PR (human merges). Enforces max attempts/diff and termination conditions. Realizes EASE-MAS A20–A22 (Algorithms 6/7).
fix-testRepair a failing or incorrect test when the test (not the product) is at fault — wrong assertion, drifted selector, bad fixture/data, or flakiness — without weakening coverage. First decides whether the failure is a real product regression (then it's a bug, not a test fix).
root-causeGiven a structured bug, produce ranked root-cause hypotheses (at least k competing causes) each with cited evidence and a confidence score, and flag ambiguity when the top hypothesis isn't clearly ahead. Read-only diagnosis that feeds fix-bug. Realizes EASE-MAS A19 (Algorithm 5).

design

SkillWhat it does
design-specTurn a brief or existing component into a design-system-aligned component spec — anatomy, variants/states, props/API, tokens, interaction, and accessibility requirements — ready for implementation and review. Framework-agnostic; grounded in the project's design tokens.
from-figmaExtract a design spec and design tokens from a Figma reference (via the Figma MCP or an exported file/link) and translate them into a design-spec plus token values that match the project's system. Optional integration — degrades gracefully when Figma access is unavailable.

development

SkillWhat it does
implement-changeImplement a bounded change (a bug fix or a small improvement/feature step) from understanding through reproduction, minimal edit, tests, and local validation — then self-review with the critic agents. Feature-branch, tests required, human merges.
refactorPerform a behavior-preserving refactor — reduce duplication, complexity, or unclear structure — safely, by first ensuring test coverage exists, then changing in small verified steps. No functional change; scope stays within ownership limits.
start-featureKick off a new feature the right way — turn intent into a spec, validate it, derive a lightweight architecture and an execution plan, get explicit human approval, then scaffold the branch, skeleton, and test stubs. Human-gated at the spec/plan (EASE-MAS A1–A7, permission L2).

devops

SkillWhat it does
fix-ciDiagnose and repair a failing CI pipeline — pull the failing run's logs (via the git provider CLI), classify the failure (build/test/lint/flake/infra), find the root cause, and apply a minimal fix or open a targeted follow-up. CLI-first (gh / az) with MCP fallback.
release-readinessRun a release-readiness gate before shipping — verify tests/coverage, quality gates (security, a11y, perf), open blockers, changelog/version, migrations, and rollback plan — and produce a go/no-go report with the evidence. Realizes EASE-MAS A23. Does not deploy.

repo

SkillWhat it does
sweep-codebaseRun a bounded codebase-hygiene sweep over a chosen scope — dead code, stale TODOs, lint/type issues, small inconsistencies, easy a11y/security nits — collect findings, deliberate, and route only approved, low-risk items to implement-change/refactor as small PRs. Bounded by design; never boils the ocean.

requirements

SkillWhat it does
from-requirementsTurn raw requirements (a brief, ticket, or conversation) into a structured PRD — goals, user stories with acceptance criteria, scope/non-goals, constraints, and non-functional requirements — and seed the traceability graph with Requirement and PRDItem nodes. Realizes EASE-MAS A1–A4.
grade-specGrade a PRD/spec for quality before build starts — completeness, testability, clarity, scope discipline, and NFR coverage — with a scored rubric and specific fixes. Blocks weak specs from entering the build loop. Realizes EASE-MAS A5.

review

SkillWhat it does
code-reviewRun a roster-driven, multi-specialist review of a code change (diff/PR). Selects the relevant critic agents by what the change touches, collects severity-tagged findings, de-duplicates, then hands them to `deliberate` for a verdict. Read-only — posts a review, never merges.
deliberateEvaluate a set of review findings and render a decision per finding and overall — Take Action, Stand Down, Defer, or Escalate — using evidence and severity, not majority vote. Used by code-review and sweep-codebase.
pr-learnExtract durable, reusable lessons from merged PRs (and their reviews) — recurring bug patterns, missed test cases, review blind spots, conventions — and record them in shared memory so future reviews and changes improve.

testing

SkillWhat it does
coverage-gapFind where tests are missing — requirements/acceptance criteria with no linked test, code paths/branches not exercised, and untested error/edge cases — and produce a prioritized gap report that feeds test-plan/author-test.
flaky-testDetect non-deterministic (flaky) tests from run history, classify the cause (timing/order/network/data/environment), quarantine to unblock CI, and propose or apply a durable fix — distinct from a real product failure. Realizes EASE-MAS A16 + failure classification (Algorithm 4).
test-planTurn a requirement/PRD into an executable, risk-based test plan — strategy, scenarios per acceptance criterion, and canonical test cases (id, requirement link, preconditions, steps, assertions) with traceability. Feeds author-test/a11y-test-gen. Realizes EASE-MAS A10–A12.
validate-scenarioValidate one test against its source scenario/requirement and return a verdict — Keep, UpdateTest, CreateTest, Deprecate, or Investigate — with an effort estimate and evidence, using a step-by-step comparison and a semantic-drift signal. Can author or fix the test on request.

testing/playwright

SkillWhat it does
author-testAuthor a Playwright test (unit-of-behavior or user journey) for a component, page, or requirement. Uses the Playwright MCP for interactive selector/snapshot discovery and the CLI to run the new test once. Links the test to its requirement in the traceability graph (EASE-MAS A13).
playwright-authAuthenticate against a configured test environment and save a reusable Playwright storage-state file (cookies/localStorage) so other Playwright skills can run authenticated tests without logging in each time. Supports none/storage-state/basic/cert/oidc auth from config.test_environments.
run-testsExecute Playwright tests deterministically via the CLI against a configured environment, collect artifacts (trace/video/screenshot/DOM/console/network), parse results, and record an Execution (and any Failures) in the traceability graph. This is the authoritative test runner (EASE-MAS A14).
update-testUpdate a Playwright test that has drifted from the current UI or requirement — fix selectors, assertions, fixtures, or flow so it once again asserts the intended behavior, without weakening it. Preserves the requirement link.
verify-testVerify a Playwright test is correct and non-flaky — it asserts the right behavior, passes for the right reason, fails when the behavior breaks, and is stable across repeated runs. Reports a verdict with evidence.