Testing
quality-gate
Orchestrates the QUALITY pipeline stage for egregore work items, running code review, unbloat, and test updates
Testing
test-updates
Updates, generates, and validates tests using git-workspace context and TDD/BDD methodology
Testing
python-testing
Python testing patterns with pytest, fixtures, TDD, mocking, async and integration tests
Testing
pytest-config
Provides standardized pytest config, reusable fixtures, and CI integration patterns
Testing
testing-quality-standards
Defines testing quality metrics, coverage thresholds, and anti-patterns
Testing
subagent-testing
Test skills via TDD in fresh subagents
Testing
quality-playbook
Explore any codebase from scratch and generate six quality artifacts: a quality constitution (QUALITY.md), spec-traced functional tests, a code review protocol with regression test…
Testing
tc-formatter
Valida y estructura Casos de Prueba en formato FLIT estricto (QA_TC{##}_{MODULO}_{ALCANCE} - {ESCENARIO}), garantizando consecutivos correctos, trazabilidad con el escenario Gherki…
Testing
write-lookahead-test
Generate rigorous look-ahead bias tests for any module that touches market data, features, or signals. Use this skill whenever the user adds or modifies code in data/feature_engine…
Testing
react18-enzyme-to-rtl
Provides exact Enzyme → React Testing Library migration patterns for React 18 upgrades. Use this skill whenever Enzyme tests need to be rewritten - shallow, mount, wrapper.find(), …
Engineering
onp-spec-driven
Spec-driven engineering with mechanical audit: Specify → Design → Tasks → Plan → Execute → Audit → Learn. Each acceptance criterion becomes an annotated test; assumptions and quest…
Testing
moviemonk-ai
Full build and safety test workflow for MovieMonk UX/Component changes. Validates that UI modifications maintain type safety, component consistency, responsive behavior, and test c…
Testing
dual-agent-harness
Build or run a soulbis ⚔️ ⊥ soulbae 🧙 dual-agent harness — a proposer and a prover held apart by a Fiat-Shamir Gap, so no result survives that the proposer could have tuned to. Use…
Testing
qa
This skill should be used when the user asks to "explore the app", "/qa", "auto-test the app", "find bugs in my app", "猴子测试", "自动探索", "随便点一下看看", "测一下全 app 看有没有崩", "smoke explore", …
Testing
qa
QA-test a website or web app and return a 1-5 quality score (5 = flawless, 1 = broken) with evidence. Use when the user wants to test, QA, evaluate, score, or "check how good" a si…
Engineering
minor-defect-fix
Handles Jira minor-defect tickets end-to-end: analyzes the issue, patches the code, expands tests to 80% coverage, syncs specs in the separate repo, logs progress in Jira, and prep…
Testing
devlab-web-deep-acceptance
Deep functional acceptance for legacy web systems (L1-L4). Registry-driven point coverage, single-entry execution, false-success detection, and triple-alignment audit. Handles inve…
Testing
qa
Behaviour QA of a PR/branch before merge in ayunis-core — spin up an isolated dev slot, seed, drive the changed flow end-to-end (API + headless browser), assert acceptance criteria…
Testing
flake-axis-bisection
Isolates the trigger behind a known-flaky test by varying one axis at a time—order, isolation, workers, viewport, latency, repetition—while tracking pass/fail counts and checking w…
Testing
parity-diff
Detects and classifies deterministic visual, property, and ARIA diffs between baseline and replacement apps using pixel, trait, and accessibility checks. Routes only “is this diff …
Testing
test-playwright
Drive a live localhost app through browser automation like a real user, exercising every page, flow, and component touched in the session. Hunt for issues across UI, backend, and D…
Testing
dashboard-e2e-testing
Verify a dashboard or agent-facing surface actually works, not just that its components render. Use before treating any dashboard/backend change as done, when adding a new surface …
Testing
appbuilder-testing
Generates and executes tests for Adobe App Builder actions and UI components. Scaffolds Jest unit tests, integration tests against deployed actions, contract tests for Adobe API in…
Engineering
huntbug
Parallel bug hunt across four coordinated tracks: isolated git bisect worktree, minimal repro via curl→proxy→/gate→browser, data integrity/tenant/cache checks, and env/language/oau…
Testing
qa-pr
Runs an agent-driven QA pass on a PR preview URL using the kmikeym v0.01m SOP: cold sign-up, first prompt, in-app exploration, edit/theme change, publish, live URL test, and remix.…
Testing
ddd-qa-chain
Executes a five-layer DDD quality-assurance chain—domain logic, IPC contracts, component interaction, E2E journeys, and visual regression—each with its own tool, command, and gate.…
Engineering
spec-driven-sprint
Drives a resumable, spec-driven sprint from a GitHub issue with anti-skip enforcement. Researches the fix surface, materializes a self-contained HTML spec with work cards and resum…
Testing
case
Import a TypR bug into cases/ and drive it from repro to fix. Use when the user hits a bug in a TypR project and wants it turned into a reproducible case (from a `typr case snapsho…
Testing
block-testing
Automated testing guidance for custom AEM Edge Delivery Services blocks. Analyzes block JavaScript and CSS for common issues including missing null checks, unscoped CSS selectors, …
Testing
manual-test-planning
Produces plain-language manual test plans from supplied context: an executive summary, named test list, and step-by-step verification details with expected outcomes. Organizes cont…
Testing
bench-build
Builds benchmark tasks from pending backlog entries by delegating each to a subagent that fully authors the task (pin, overlay, prompt, assertions, weights, self-check) and produce…
AI / ML
forge-evals
Design evaluations for LLM features including golden datasets, rubric scoring, LLM-as-judge calibration, CI regression detection, online A/B tests, cost and latency budgets, and ad…
Testing
qa-agent
Autonomous QA agent using playwright-cli. Reads the codebase, explores a running local app, attempts to reproduce a bug described in plain language, then creates a structured GitHu…
Testing
sap-sac-test-automation
Designs capability-gated browser discovery and deterministic Playwright test suites for SAP Analytics Cloud stories, dashboards, planning workflows, comments, permissions, and visu…
Testing
test-convert
Gera testes robustos e comportamentais para o Convert (React + Redux + Axios) seguindo BDD e as diretrizes do knowledge base. Analisa componentes, serviços, validações e cálculos e…
Testing
qa
Compare extracted WXR content against the original source site page by page. Find missing text, headings, images, and links. Fix by patching the WXR or re-extracting individual pag…
Testing
store-welle-usertest
Conducts a Windows Store wave test with the user: opens release apps sequentially, captures live feedback losslessly into task lists, prepares numbered asset reviews, generates sub…
Testing
loop-evals
Designs the evaluation harness for agent loops, establishing trustworthy verification through a 7-layer suite with false-completion-rate and repair-productivity as core metrics. En…
Testing
cypress-tap
Controls a live Cypress open-mode session via `cypress tap` to execute specs, inspect reporter output, query DOM/accessibility/state, and rewind the app under test. Use for authori…
Testing
agento11y-test-starter
Builds an offline test suite for an agent before deployment by analyzing its code, prompts, and tools to generate labeled cases (happy/edge/adversarial) with scoring recommendation…
Testing
kelly-app-skill-creator-tests
Validates and executes integration tests for App-in-Skill projects, including server smoke tests, browser acceptance, Busabase integration, OAuth verification, persistence, and CI …
Testing
playwright-e2e-builder
Plans and builds comprehensive Playwright E2E test suites using Page Object Model, authentication state persistence, custom fixtures, visual regression, and CI integration. Conduct…
Testing
qa-loop
qa-loop runs a disciplined review cycle that stops on diminishing returns rather than zero defects. It anchors every pass to the implementation plan, fixes only plan-aligned issues…
Testing
forge-qa-warfare-v2
Delivers exhaustive actor-based QA coverage for FitnessMealPlanner across all roles, endpoints, states, inputs, and assertions. Executes a 9-phase pipeline integrating six speciali…
Data
simulation-study
Generates reproducible Monte Carlo simulation studies in R: parameterized data-generating process, estimator grid, seeded replication loop, and summaries of bias, RMSE, coverage, s…
Testing
chaos-tdd-fault-injection
Design deterministic fault-injection tests that declare guarded behavior first, then implement minimal guards at protocol seams for LLM and external I/O failures. Use to harden pip…
Testing
qa-gen
Generates a test plan from spec criteria, classifying each by verification method and writing 06_qa.md with scoped, traced scenarios. UI criteria map to browser runs; backend crite…
Testing
testing-philosophy
Defines test quality by prioritizing observable behavior over implementation details, establishing a Testing Trophy with minimum end-to-end coverage for user-facing features. Appli…
Testing
yamlet-tester
Projects a directory of EARS specs (.yamlet.yaml) into Gherkin `.feature` files with `yamlet tests`, force-regenerating the tree each run. Requires the specs directory; optional ta…
Testing
playwright-e2e-execution-run
Executes an existing Playwright end-to-end suite against a confirmed non-production target and returns a structured run report including pass/fail counts, flaky tests, durations, a…
Security
test-red-team
Red-teams any web, React Native, or hybrid app by driving browser automation, Android WebView, or native device interaction against a feature coverage matrix. Attacks UI, data flow…
Automation
ab-scripting-feature-dev
Creates an AgenticBrowser Scripting DSL orchestration for end-to-end feature development: planning, TDD implementation, testing, and lint review. Outputs docs/scripting-features/fe…
Testing
shape-test-plan
Orchestrates stateful phased test rollouts for existing products by writing a durable plan at context/foundation/test-plan.md then sequencing through research, planning, and implem…
Testing
bruno-api-testing
Create, execute, and manage API test collections in Bruno format (OpenCollection YAML or legacy Bru). Supports generating collections from OpenAPI specs, writing requests with asse…
Testing
schoolos-migration-guard
Checks Django 5.2 + PostgreSQL 16 migrations for unsafe operations on live data: exclusive locks, data loss, backward-incompatible changes, NOT NULL without defaults, irreversible …
Testing
dual-testing
Language-agnostic testing strategy that pairs real-infrastructure integration tests (e.g. Testcontainers) for happy paths with mocked unit/slice tests for error handling and mappin…
Testing
part-factory-test-data
Generates reusable test fixtures and builders via @kensio/part-factory, keeping setup concerns inside factories while passing only runtime dependencies at call time. Supports Stati…
AI / ML
prai-experiments
Supports reproducible ML experiments for PR&AI research: RQ breakdown, fair baselines with matched splits/compute, metric selection (accuracy/F1/mAP/IoU/AUC), multi-run statistics …
Testing
unit-test-bean-validation
Provides patterns for unit testing Jakarta Bean Validation (JSR-380) constraints including @Valid, @NotNull, @Min, @Max, and @Email with Hibernate Validator. Generates custom valid…
Testing
user-uat
Executes a provided UAT or manual-smoke plan by running each step, capturing output, and auto-verifying mechanical checks (exit codes, output matches, refusal text, DB/HTTP/file/lo…
Showing the top 60 of 4,925. See the full list →