skillZs
★ LIVE SKILL TAGS ★
>>> LIVE SKILLS INDEX <<<
* OPEN SOURCE *
NO LOGIN, NO TRACKING
※ REAL INSTALL DATA ※
← back to all skills
citypaul/.dotfiles182 installs

test-design-reviewer

Review test quality using Dave Farley's eight properties of good tests. Use when assessing a test file or suite for understandability, maintainability, repeatability, atomicity, necessity, granularity, speed, and evidence of test-first development.

How do I install this agent skill?

npx skills add https://github.com/citypaul/.dotfiles --skill test-design-reviewer
view source ↗

Is this agent skill safe to install?

  • Gen Agent Trust Hubpass

    The skill is a code review tool for evaluating test quality. It is safe for its intended use, though like any tool analyzing external code, it possesses a standard minor vulnerability to instructions that could be hidden within the files it reviews.

  • Socketpass

    No alerts

  • Snykpass

    Risk: LOW · No issues

What does this agent skill do?

Test Design Reviewer

Review tests as executable specifications and safety evidence. Read the tests before the implementation so their public story can stand on its own, then inspect the production boundary and repository constraints needed to judge the claims accurately.

Properties

PropertyInspectStrong evidence
UnderstandableNames, arrange/act/assert flow, domain vocabularyThe behavior and failure are clear without reconstructing internals
MaintainableCoupling, duplication, fixtures, public boundariesBehavior-preserving refactors do not require unrelated test rewrites
RepeatableTime, randomness, concurrency, network, shared resourcesRepeated and parallel runs have controlled inputs and cleanup
AtomicShared state, ordering, cleanup, failure isolationA test can run alone and its failure identifies one behavior
NecessaryDistinct risk or contract protectedRemoving the test would remove meaningful evidence
GranularScope of behavior and diagnostic qualityAssertions describe one coherent outcome; related assertions may stay together
FastMeasured feedback time at the appropriate layerThe suite is fast enough for its intended feedback loop
FirstEvidence of test-first developmentA captured RED run, development trace, or history demonstrates the test failed for the expected reason before production behavior changed

Rating

Rate each property Strong, Mixed, Weak, or Not assessed.

  • Use exact file locations and observed evidence.
  • Do not calculate an aggregate score; unequal risks and repository contexts make a weighted number falsely precise.
  • Mark First Not assessed when only the final tree is available. Static test shape cannot prove chronology.
  • Mark Fast Not assessed unless execution evidence or trustworthy timing is available.
  • Prefer the smallest change that strengthens observable behavior. Do not demand one assertion per test, one test per file, or unit tests where a higher-level contract is the honest evidence boundary.

Review Process

  1. Establish the claimed behavior, test layer, repository policy, and relevant risk.
  2. Read the tests without implementation and record what a failure would mean.
  3. Inspect the public production boundary, fixtures, and configured runner.
  4. Run focused tests or timing only when authorized and useful; report exactly what ran.
  5. Rate every property with evidence, including Not assessed where evidence is absent.
  6. Rank only actionable findings by severity and impact. Include the smallest credible fix.
  7. Separate test defects from production-design seams and local policy preferences.

Output

## Test design review: [scope]

| Property | Rating | Evidence |
|---|---|---|
| Understandable | Strong/Mixed/Weak/Not assessed | [file:line and reason] |
| Maintainable | ... | ... |
| Repeatable | ... | ... |
| Atomic | ... | ... |
| Necessary | ... | ... |
| Granular | ... | ... |
| Fast | ... | ... |
| First | ... | ... |

### Findings

1. **[severity] — [problem]** (`path:line`)
   Impact: [observable risk].
   Smallest fix: [action].

### Validation gaps

- [Anything not assessed and the evidence needed]

No findings is a valid result; do not invent work to populate the section.

Source And Attribution

The eight properties are drawn from Dave Farley's Properties of Good Tests. This version is a fresh, evidence-based implementation rather than a textual adaptation of an external skill. Read references/source-notes.md for the exact historical provenance and unresolved permission issue in older releases.

Add the canonical catalog link to the repository README so users can inspect current installs and available audits. The publishing guide covers the complete discovery path.

<a href="https://skillzs.dev/skills/citypaul/.dotfiles/test-design-reviewer">View test-design-reviewer on skillZs</a>