Agent Skills·Use case ·Test generation
Use case · 4,925 skills

Agent skills for test generation

Skills that write tests for you — unit, integration, edge cases — in your project's framework and conventions. Drop the SKILL.md into any compatible agent.

Browse all 4,925 Test generation skills →

Testing
quality-gate
Orchestrates the QUALITY pipeline stage for egregore work items, running code review, unbloat, and test updates
openclawqualityreviewtesting
Testing
test-updates
Updates, generates, and validates tests using git-workspace context and TDD/BDD methodology
openclawtddbddtests
Testing
python-testing
Python testing patterns with pytest, fixtures, TDD, mocking, async and integration tests
openclawpythonpytesttdd
Testing
pytest-config
Provides standardized pytest config, reusable fixtures, and CI integration patterns
openclawleylinepytestfixtures
Testing
testing-quality-standards
Defines testing quality metrics, coverage thresholds, and anti-patterns
openclawleylinetestingquality-metrics
Testing
subagent-testing
Test skills via TDD in fresh subagents
openclawtddsubagentstesting
Testing
quality-playbook
Explore any codebase from scratch and generate six quality artifacts: a quality constitution (QUALITY.md), spec-traced functional tests, a code review protocol with regression test…
claude-codecodexcursorgemini-cliqualitytestingcode-review
Testing
tc-formatter
Valida y estructura Casos de Prueba en formato FLIT estricto (QA_TC{##}_{MODULO}_{ALCANCE} - {ESCENARIO}), garantizando consecutivos correctos, trazabilidad con el escenario Gherki…
claude-codecodexcursorgemini-cliqatest-casesgherkin
Testing
write-lookahead-test
Generate rigorous look-ahead bias tests for any module that touches market data, features, or signals. Use this skill whenever the user adds or modifies code in data/feature_engine…
claude-codecodexcursorgemini-clitestinglookahead-biasmarket-data
Testing
react18-enzyme-to-rtl
Provides exact Enzyme → React Testing Library migration patterns for React 18 upgrades. Use this skill whenever Enzyme tests need to be rewritten - shallow, mount, wrapper.find(), …
claude-codecodexcursorgemini-clireacttesting-libraryenzyme
Engineering
onp-spec-driven
Spec-driven engineering with mechanical audit: Specify → Design → Tasks → Plan → Execute → Audit → Learn. Each acceptance criterion becomes an annotated test; assumptions and quest…
claude-codecodexcursorgemini-clitype:auditspec-drivenmechanical-audit
Testing
moviemonk-ai
Full build and safety test workflow for MovieMonk UX/Component changes. Validates that UI modifications maintain type safety, component consistency, responsive behavior, and test c…
claude-codecodexcursorgemini-climovieuxlinting
Testing
dual-agent-harness
Build or run a soulbis ⚔️ ⊥ soulbae 🧙 dual-agent harness — a proposer and a prover held apart by a Fiat-Shamir Gap, so no result survives that the proposer could have tuned to. Use…
claude-codecodexcursorgemini-clidual-agentfiat-shamirproposer
Testing
qa
This skill should be used when the user asks to "explore the app", "/qa", "auto-test the app", "find bugs in my app", "猴子测试", "自动探索", "随便点一下看看", "测一下全 app 看有没有崩", "smoke explore", …
claude-codecodexcursorgemini-cliandroidexploratory-testingcrash-detection
Testing
qa
QA-test a website or web app and return a 1-5 quality score (5 = flawless, 1 = broken) with evidence. Use when the user wants to test, QA, evaluate, score, or "check how good" a si…
claude-codecodexcursorgemini-cliqatestingbrowser
Engineering
minor-defect-fix
Handles Jira minor-defect tickets end-to-end: analyzes the issue, patches the code, expands tests to 80% coverage, syncs specs in the separate repo, logs progress in Jira, and prep…
claude-codecodexcursorgemini-clitool:jirajiracode-review
Testing
devlab-web-deep-acceptance
Deep functional acceptance for legacy web systems (L1-L4). Registry-driven point coverage, single-entry execution, false-success detection, and triple-alignment audit. Handles inve…
claude-codecodexcursorgemini-cliacceptance-testinge2elegacy-systems
Testing
qa
Behaviour QA of a PR/branch before merge in ayunis-core — spin up an isolated dev slot, seed, drive the changed flow end-to-end (API + headless browser), assert acceptance criteria…
claude-codecodexcursorgemini-cliqae2e-testingbrowser-automation
Testing
flake-axis-bisection
Isolates the trigger behind a known-flaky test by varying one axis at a time—order, isolation, workers, viewport, latency, repetition—while tracking pass/fail counts and checking w…
claude-codecodexcursorgemini-cliflaky-testsbisectionconfidence
Testing
parity-diff
Detects and classifies deterministic visual, property, and ARIA diffs between baseline and replacement apps using pixel, trait, and accessibility checks. Routes only “is this diff …
claude-codecodexcursorgemini-clivisual-testingaccessibilitydiff
Testing
test-playwright
Drive a live localhost app through browser automation like a real user, exercising every page, flow, and component touched in the session. Hunt for issues across UI, backend, and D…
claude-codecodexcursorgemini-clicloud:supabaseplaywrighte2e
Testing
dashboard-e2e-testing
Verify a dashboard or agent-facing surface actually works, not just that its components render. Use before treating any dashboard/backend change as done, when adding a new surface …
claude-codecodexcursorgemini-clie2edashboardverification
Testing
appbuilder-testing
Generates and executes tests for Adobe App Builder actions and UI components. Scaffolds Jest unit tests, integration tests against deployed actions, contract tests for Adobe API in…
claude-codecodexcursorgemini-clitype:generatortype:integrationadobe
Engineering
huntbug
Parallel bug hunt across four coordinated tracks: isolated git bisect worktree, minimal repro via curl→proxy→/gate→browser, data integrity/tenant/cache checks, and env/language/oau…
claude-codecodexcursorgemini-clidebugginggitreproduction
Testing
qa-pr
Runs an agent-driven QA pass on a PR preview URL using the kmikeym v0.01m SOP: cold sign-up, first prompt, in-app exploration, edit/theme change, publish, live URL test, and remix.…
claude-codecodexcursorgemini-cliai:agenttype:reviewqa
Testing
ddd-qa-chain
Executes a five-layer DDD quality-assurance chain—domain logic, IPC contracts, component interaction, E2E journeys, and visual regression—each with its own tool, command, and gate.…
claude-codecodexcursorgemini-clidddqatesting
Engineering
spec-driven-sprint
Drives a resumable, spec-driven sprint from a GitHub issue with anti-skip enforcement. Researches the fix surface, materializes a self-contained HTML spec with work cards and resum…
claude-codecodexcursorgemini-clisprinttddworktree
Testing
case
Import a TypR bug into cases/ and drive it from repro to fix. Use when the user hits a bug in a TypR project and wants it turned into a reproducible case (from a `typr case snapsho…
claude-codecodexcursorgemini-clibugreprotypr
Testing
block-testing
Automated testing guidance for custom AEM Edge Delivery Services blocks. Analyzes block JavaScript and CSS for common issues including missing null checks, unscoped CSS selectors, …
claude-codecodexcursorgemini-clilang:javascriptaemblocks
Testing
manual-test-planning
Produces plain-language manual test plans from supplied context: an executive summary, named test list, and step-by-step verification details with expected outcomes. Organizes cont…
claude-codecodexcursorgemini-climanual-testingtest-plansverification
Testing
bench-build
Builds benchmark tasks from pending backlog entries by delegating each to a subagent that fully authors the task (pin, overlay, prompt, assertions, weights, self-check) and produce…
claude-codecodexcursorgemini-clibenchmarktasksbacklog
AI / ML
forge-evals
Design evaluations for LLM features including golden datasets, rubric scoring, LLM-as-judge calibration, CI regression detection, online A/B tests, cost and latency budgets, and ad…
claude-codecodexcursorgemini-cliai:llmllmevaluation
Testing
qa-agent
Autonomous QA agent using playwright-cli. Reads the codebase, explores a running local app, attempts to reproduce a bug described in plain language, then creates a structured GitHu…
claude-codecodexcursorgemini-cliplaywrightqabug-reproduction
Testing
sap-sac-test-automation
Designs capability-gated browser discovery and deterministic Playwright test suites for SAP Analytics Cloud stories, dashboards, planning workflows, comments, permissions, and visu…
claude-codecodexcursorgemini-cliai:agentcloud:verceltype:integration
Testing
test-convert
Gera testes robustos e comportamentais para o Convert (React + Redux + Axios) seguindo BDD e as diretrizes do knowledge base. Analisa componentes, serviços, validações e cálculos e…
claude-codereactreduxaxios
Testing
qa
Compare extracted WXR content against the original source site page by page. Find missing text, headings, images, and links. Fix by patching the WXR or re-extracting individual pag…
claude-codecodexcursorgemini-cliqawxrcontent-validation
Testing
store-welle-usertest
Conducts a Windows Store wave test with the user: opens release apps sequentially, captures live feedback losslessly into task lists, prepares numbered asset reviews, generates sub…
claude-codecodexcursorgemini-clitype:reviewwindows-storeusability-testing
Testing
loop-evals
Designs the evaluation harness for agent loops, establishing trustworthy verification through a 7-layer suite with false-completion-rate and repair-productivity as core metrics. En…
claude-codecodexcursorgemini-cliai:agentloopsevaluation
Testing
cypress-tap
Controls a live Cypress open-mode session via `cypress tap` to execute specs, inspect reporter output, query DOM/accessibility/state, and rewind the app under test. Use for authori…
claude-codecodexcursorgemini-clitype:debugcypresse2e
Testing
agento11y-test-starter
Builds an offline test suite for an agent before deployment by analyzing its code, prompts, and tools to generate labeled cases (happy/edge/adversarial) with scoring recommendation…
claude-codecodexcursorgemini-clitest-suiteofflineevaluation
Testing
kelly-app-skill-creator-tests
Validates and executes integration tests for App-in-Skill projects, including server smoke tests, browser acceptance, Busabase integration, OAuth verification, persistence, and CI …
claude-codecodexcursorgemini-clitype:integrationtestingintegration
Testing
playwright-e2e-builder
Plans and builds comprehensive Playwright E2E test suites using Page Object Model, authentication state persistence, custom fixtures, visual regression, and CI integration. Conduct…
claude-codecodexcursorgemini-clitype:integrationplaywrighte2e
Testing
qa-loop
qa-loop runs a disciplined review cycle that stops on diminishing returns rather than zero defects. It anchors every pass to the implementation plan, fixes only plan-aligned issues…
claude-codecodexcursorgemini-cliqaregressionreview-cycle
Testing
forge-qa-warfare-v2
Delivers exhaustive actor-based QA coverage for FitnessMealPlanner across all roles, endpoints, states, inputs, and assertions. Executes a 9-phase pipeline integrating six speciali…
claude-codecodexcursorgemini-cliqafitnessmealplanneractor-based
Data
simulation-study
Generates reproducible Monte Carlo simulation studies in R: parameterized data-generating process, estimator grid, seeded replication loop, and summaries of bias, RMSE, coverage, s…
claude-codecodexcursorgemini-clitype:generatormonte-carlosimulation
Testing
chaos-tdd-fault-injection
Design deterministic fault-injection tests that declare guarded behavior first, then implement minimal guards at protocol seams for LLM and external I/O failures. Use to harden pip…
claude-codecodexcursorgemini-clitestingfault-injectionchaos-engineering
Testing
qa-gen
Generates a test plan from spec criteria, classifying each by verification method and writing 06_qa.md with scoped, traced scenarios. UI criteria map to browser runs; backend crite…
claude-codecodexcursorgemini-clitestingqatest-plan
Testing
testing-philosophy
Defines test quality by prioritizing observable behavior over implementation details, establishing a Testing Trophy with minimum end-to-end coverage for user-facing features. Appli…
claude-codecodexcursorgemini-clitype:reviewtestingquality
Testing
yamlet-tester
Projects a directory of EARS specs (.yamlet.yaml) into Gherkin `.feature` files with `yamlet tests`, force-regenerating the tree each run. Requires the specs directory; optional ta…
claude-codecodexcursorgemini-cliearsgherkinfeature-files
Testing
playwright-e2e-execution-run
Executes an existing Playwright end-to-end suite against a confirmed non-production target and returns a structured run report including pass/fail counts, flaky tests, durations, a…
claude-codecodexcursorgemini-clitype:reviewplaywrighte2e
Security
test-red-team
Red-teams any web, React Native, or hybrid app by driving browser automation, Android WebView, or native device interaction against a feature coverage matrix. Attacks UI, data flow…
claude-codecodexcursorgemini-clicloud:supabasered-teamsecurity
Automation
ab-scripting-feature-dev
Creates an AgenticBrowser Scripting DSL orchestration for end-to-end feature development: planning, TDD implementation, testing, and lint review. Outputs docs/scripting-features/fe…
claude-codecodexcursorgemini-cliai:agenttype:reviewdsl
Testing
shape-test-plan
Orchestrates stateful phased test rollouts for existing products by writing a durable plan at context/foundation/test-plan.md then sequencing through research, planning, and implem…
claude-codecodexcursorgemini-clitest-planphasedrollout
Testing
bruno-api-testing
Create, execute, and manage API test collections in Bruno format (OpenCollection YAML or legacy Bru). Supports generating collections from OpenAPI specs, writing requests with asse…
claude-codecodexcursorgemini-clibrunoapi-testingopenapi
Testing
schoolos-migration-guard
Checks Django 5.2 + PostgreSQL 16 migrations for unsafe operations on live data: exclusive locks, data loss, backward-incompatible changes, NOT NULL without defaults, irreversible …
claude-codecodexcursorgemini-clilang:pythondjangopostgresql
Testing
dual-testing
Language-agnostic testing strategy that pairs real-infrastructure integration tests (e.g. Testcontainers) for happy paths with mocked unit/slice tests for error handling and mappin…
claude-codecodexcursorgemini-clilang:typescripttype:integrationtesting
Testing
part-factory-test-data
Generates reusable test fixtures and builders via @kensio/part-factory, keeping setup concerns inside factories while passing only runtime dependencies at call time. Supports Stati…
claude-codecodexcursorgemini-clitypescriptfactoriesfixtures
AI / ML
prai-experiments
Supports reproducible ML experiments for PR&AI research: RQ breakdown, fair baselines with matched splits/compute, metric selection (accuracy/F1/mAP/IoU/AUC), multi-run statistics …
claude-codecodexcursorgemini-climachine-learningreproducibilityexperiments
Testing
unit-test-bean-validation
Provides patterns for unit testing Jakarta Bean Validation (JSR-380) constraints including @Valid, @NotNull, @Min, @Max, and @Email with Hibernate Validator. Generates custom valid…
claude-codecodexcursorgemini-clijavajunitjakarta-validation
Testing
user-uat
Executes a provided UAT or manual-smoke plan by running each step, capturing output, and auto-verifying mechanical checks (exit codes, output matches, refusal text, DB/HTTP/file/lo…
claude-codecodexcursorgemini-cliuatsmoke-testverification

Showing the top 60 of 4,925. See the full list →