team-qa skill
Orchestrate the QA team through a full testing cycle — qa-lead strategy and test plan, qa-tester case writing, execution, sign-off.
Is the team-qa skill safe?
Clean: nothing in its files matched our rules. We read 1 file in the folder on 2026-09-28.
No findings.
Install the team-qa skill
A skill is a folder. Copy it into your agent's skills folder and the agent loads it when the task matches its description.
git clone --depth 1 https://github.com/Donchitos/Claude-Code-Game-Studios.git /tmp/Claude-Code-Game-Studios mkdir -p ~/.claude/skills cp -r /tmp/Claude-Code-Game-Studios/.claude/skills/team-qa ~/.claude/skills/team-qa
In the Claude apps, zip the folder and upload it from the Skills settings. The folder on GitHub
The instructions your agent would load
SKILL.md as published, without the frontmatter. Read it on GitHub
When this skill is invoked, orchestrate the QA team through a structured testing cycle.
Decision Points: At each phase transition, use AskUserQuestion to present the user with the subagent's proposals as selectable options. Write the agent's full analysis in conversation, then capture the decision with concise labels. In collaborative mode, the user must approve before moving to the next phase. In guided mode the pipeline advances automatically unless a phase is BLOCKED; in autonomous mode it runs end to end, recording each phase outcome via logdecision. Decisions in automationalwaysask categories (isalwaysaskcategory helper) always prompt regardless of mode. See .claude/docs/automation-modes.md.
Phase 0: Resolve Config
!bash "${CLAUDESKILLDIR}/../../hooks/yaml-helper.sh" resolveconfig --keys reviewmode,automation,team.size
Resolved above — use as-is; --review overrides review_mode. No block → defaults in .claude/docs/config-resolution.md.
review_mode sets gate depth:
- full — spawn all director and lead gates as described
- lean — skip director gates unless they are PHASE-GATE type (CD-PHASE-GATE, TD-PHASE-GATE, PR-PHASE-GATE, AD-PHASE-GATE)
- solo — skip all director gate spawning entirely; run the skill without any agent gates
automation drives the Decision Points note above. See the Decision Points note above and .claude/docs/automation-modes.md for how each mode changes pipeline behavior.
team.size: which agents are active (orthogonal to review_mode gate-depth and workflow docs).
Directors (CD/TD/PR) still spawn at phase gates regardless of size; a non-core agent needed at individual routes through the nearest active core agent with an informational note. "Phase gate" means any phase that ends in an AskUserQuestion decision point before the pipeline advances — not every phase. Apply the test literally: if the phase below has no decision point, it is not a gate, and an agent restricted to "phase gates only" is not spawned for it. This active-set scoping applies throughout the pipeline below: any phase that names an agent outside the active set routes through the nearest core agent rather than spawning it.
- individual (default): qa-tester only; qa-lead invoked at phase gates only.
- small: qa-lead + qa-tester pipeline (as documented).
- studio: qa-lead + per-story qa-tester spawn + sign-off.
Announce the active set before Phase 1 — never let the collapse be silent. Before spawning anything, state in one line which agents this run will actually spawn, and which the pipeline below names but will not spawn at the resolved team.size. For example:
Active set (team.size: ): .
Not spawned this run: — consulted
through . Raise team.size (or modes.rigor) to widen.
Fill it from the team.size list directly above and the agents this file's own pipeline names — not from an example. Both sets differ per orchestrator.
The pipeline below reads as a multi-agent fan-out and at the shipped default it is one or two agents — team-release names eight and runs one, team-narrative names six across five phases and runs writer alone. The collapse is correct: team.size is rigor-fronted and the narrow default is the token lever, measured at roughly 10x. What was wrong is that nothing said so, so a reader could not distinguish a correctly-collapsed run from a broken pipeline, and the per-agent "routes through the nearest core agent with an informational note" rule above fires at routing time and never states the shape of the run as a whole.
This is the same rule as the skipped-check reporting elsewhere in this file: a constraint that is enforced but never surfaced is indistinguishable, to the person reading the output, from one that was never enforced.
Team Composition
- qa-lead — QA strategy, test plan generation, story classification, sign-off report
- qa-tester — Test case writing, bug report writing, manual QA documentation
How to Delegate
Use the Agent tool to spawn each team member as a subagent:
- subagent_type: qa-lead — Strategy, planning, classification, sign-off
- subagent_type: qa-tester — Test case writing and bug report writing
Brief each agent — do not dump context. Read the shared inputs once and pass a distilled brief inline: the lines each agent actually needs, never a file path for a document you have already read (an agent handed a path re-reads the whole file). Pass a path only for a document you have not read and only that agent needs.
End every agent prompt with a return contract: "Write your full output to [path] — that named path is your write authorisation under the bounded exception below, so write it without a separate approval prompt. Return only (1) the path written, (2) a ≤5-bullet summary of decisions, (3) any BLOCKED/CONCERNS items, one line each. Do not restate the documents you read." Without it, an agent returns everything it read back into this session.
Why this does not violate the Collaboration Protocol. CLAUDE.md requires an agent to ask "May I write this to [filepath]?" before Write/Edit. A subagent spawned here writes without asking, and that is a deliberate, bounded exception rather than an oversight — the same call already made for consistency-check appending to active.md. The exception holds only when all three are true: (1) the path is one you named in the prompt, so the user approved the destination when they approved the phase; (2) it is a new artifact under production/, docs/ or tests/, never an edit to existing source or config; (3) the phase that produced it is itself gated by an AskUserQuestion before the pipeline advances. Outside those three, the agent must ask. Do not "fix" this by asking per subagent — a prompt per agent per phase makes an orchestrator unusable, which is why the exception exists.
Launch independent qa-tester tasks in parallel where possible (e.g., multiple stories in Phase 5 can be scaffolded simultaneously).
Pipeline
Phase 1: Load Context
Before doing anything else, gather the full scope:
- Detect the current sprint or feature scope from the argument:
- If argument is a sprint identifier (e.g., sprint-03): Glob production/sprints/ for files matching [sprint-identifier].md. Read the matched file. If multiple match, use the most recently modified.
- If argument is feature: [system-name]: glob story files tagged for that system
- If no argument: read production/session-state/active.md and production/sprint-status.yaml (if present) to infer the active sprint
- Read project.stage from project.yaml (fallback production/stage.txt) to confirm the current project phase.
- Count stories found and report to the user:
"QA cycle starting for [sprint/feature]. Found [N] stories. Current stage: [stage]. Ready to begin QA strategy?"
Phase 2: QA Strategy (qa-lead)
Spawn qa-lead via Agent to review all in-scope stories and produce a QA strategy.
Prompt the qa-lead to:
- Read each story file
- Classify each story by type: Logic / Integration / Visual/Feel / UI / Config/Data
- Identify which stories require automated test evidence vs. manual QA
- Flag any stories with missing acceptance criteria or missing test evidence that would block QA
- Estimate manual QA effort (number of test sessions needed)
- Before assessing smoke status, check for an existing smoke check report: Glob production/qa/smoke-.md and read the most recently modified file (if found). If a report exists, use its verdict and findings directly — do not re-interview the user. If no report exists, note: "No prior smoke check report found — run /smoke-check sprint before proceeding." and set smoke check status to UNKNOWN (treat as PASS WITH WARNINGS for the purpose of continuing). Produce a smoke check verdict: PASS / PASS WITH WARNINGS [list] / FAIL [list of failures] / UNKNOWN (no report found)**
- Produce a strategy summary table and smoke check result:
Smoke Check: [PASS / PASS WITH WARNINGS / FAIL / UNKNOWN] — [source: production/qa/smoke-[date].md or "no report found"] — [details if not PASS]
If the smoke check result is FAIL, the qa-lead must list the failures prominently. QA cannot proceed past the strategy phase with a failed smoke check.
Present the qa-lead's full strategy to the user, then use AskUserQuestion:
question: "QA Strategy Review"
options:
- "Looks good — proceed to test plan"
- "Adjust story types before proceeding"
- "Skip blocked stories and proceed with the rest"
- "Smoke check failed — fix issues and re-run /team-qa"
- "Cancel — resolve blockers first"If smoke check FAIL: do not proceed to Phase 3. Surface the failures from the smoke check report and stop. The user must fix them, re-run /smoke-check sprint, and then re-run /team-qa. If smoke check UNKNOWN: surface a warning — "No smoke check report found. Recommend running /smoke-check sprint before QA. Proceeding with caution." If smoke check PASS WITH WARNINGS: note the warnings for the sign-off report and continue. If blockers are present: list them explicitly. The user may choose to skip blocked stories or cancel the cycle.
Phase 3: Test Plan Generation
Using the strategy from Phase 2, produce a structured test plan document.
The test plan should cover:
- Scope: sprint/feature name, story count, dates
- Story Classification Table: from Phase 2 strategy
- Automated Test Requirements: which stories need test files, expected paths in tests/
- Manual QA Scope: which stories need manual walkthrough and what to validate
- Out of Scope: what is explicitly not being tested this cycle and why
- Entry Criteria: what must be true before QA can begin. Always include: (1) Smoke check PASS or PASS WITH WARNINGS report exists at production/qa/smoke-*.md, (2) build is stable (no crashes on launch), (3) all Must Have stories have Status: in-progress or done in production/sprint-status.yaml. Add any sprint-specific criteria beyond these.
- Exit Criteria: what constitutes a completed QA cycle (all stories PASS or FAIL with bugs filed)
Ask: "May I write the QA plan to production/qa/qa-plan-[sprint]-[date].md?"
Write only after receiving approval.
Phase 4: Test Case Writing (qa-tester)
Smoke check is performed as part of Phase 2 (QA Strategy). If the smoke check returned FAIL in Phase 2, the cycle was stopped there. This phase only runs when the Phase 2 smoke check was PASS, PASS WITH WARNINGS, or UNKNOWN.
For each story requiring manual QA (Visual/Feel, UI, Integration without automated tests):
Spawn qa-tester via Agent for each story (run in parallel where possible), providing:
explicitly in the prompt, one per story.
- The story file path
- The relevant section of the QA plan for that story
- The GDD acceptance criteria for the system being tested (if available)
- Instructions to write detailed test cases covering all acceptance criteria
- The output path: production/qa/test-cases/[story-slug]-cases.md. Name it
Why the path is stated here rather than left to the orchestrator. The
bounded write exception above holds only when "the path is one you named in
the prompt". This is the phase that spawns agents in parallel, so it is where
an unnamed destination does the most damage: each agent improvises its own, and
More skills from Donchitos/Claude-Code-Game-Studios
- AadoptBrownfield audit — do existing artifacts actually work? Numbered migration plan. Unlike /project-stage-detect, checks compliance not existence.
- Aarchitecture-decisionCreate an ADR documenting a technical decision: context, alternatives considered, consequences.
- Aarchitecture-reviewTraceability matrix mapping GDD requirements to ADRs. Finds gaps, cross-ADR conflicts, engine compatibility. PASS/CONCERNS/NOT ASSESSED/FAIL.
- Aart-bibleAuthor the Art Bible — visual identity gating asset production. Run before /map-systems.
- Aasset-auditAudit assets against naming conventions, file size budgets, format standards. Finds orphaned assets, missing references.
- Aasset-specPer-asset visual specs plus AI generation prompts from GDDs and character profiles. After the art bible.
- Abalance-checkFind balance outliers, broken progressions, degenerate strategies, economy imbalances in formulas and data. 'Check game balance'.
- AbrainstormGuided concept ideation using professional studio techniques, player psychology, creative exploration.
- Abug-reportStructured bug report from a description, or analyze code for potential bugs. Reproduction steps, severity.
- Abug-triageRe-evaluate open bugs — priority vs severity, assign to sprints, surface systemic trends. Run when the count grows.
- AchangelogAuto-generate a changelog from git commits and sprint data. Internal and player-facing versions.
- Acode-reviewArchitectural code review — coding standards, SOLID, testability, performance concerns.