Agent skills

Testing & QA skills

Read straight from the source repositories, not from submitted listings. Every skill shows what it does, what is inside, where it came from — and whether attention around its source is actually growing.

Toolclaude-code 29,140codex 4,755cursor 3,111copilot 976windsurf 55cline 34
CategoryWorkflow & Productivity 4,979AI & Agents 3,037Data & Analytics 2,345Code Review & Quality 1,376Backend & API 1,244Security 1,194Design & Presentation 1,154Documentation 965Content & Marketing 916Testing & QA 777DevOps & Cloud 576Databases 550Frontend 469Business & Finance 328Media & Video 257Other 9,833
1,723 found
1,5851,632 · page 34 / 36
snapshot-test-helperAssist with snapshot test helper operations. Auto-activating skill for Test Automation. Triggers on: snapshot test helper…jeremylongshorewritessnapshot-test-setupSet up SwiftUI visual regression testing with swift-snapshot-testing. Generates snapshot test boilerplate and CI configuration.…rshankraswritessnowflake-load-scaleImplement Snowflake load testing, warehouse scaling, and capacity planning. Use when testing query performance at scale…jeremylongshorewritessoak-test-plannerPlan soak test planner operations. Auto-activating skill for Performance Testing. Triggers on: soak test planner, soak test…jeremylongshorewritessoctheory-theory-constructionUse when building the actual theory for a Sociological Theory (ST) manuscript — turning a positioned problem into defined…brycewang-stanfordsoftware-vv-test-generatorMedical device software verification and validation test case generation skilla5c-aiwritessource-triangulationVerify a claim before repeating it — the independent-sources test (three citations of one press release is one source), the…mohitagw15856sparc-refineRun the SPARC Refinement and Completion phases — review code, improve test coverage, validate against specification, and generate…ruvnetwritesspeckit-clarifyStructured clarification workflow for underspecified requirements. Use before planning to resolve ambiguities through…foryourhealth111-pixelspike-consumer-adversarialOI-3 spike harness — heavy consumer, ADVERSARIAL arm. Worst-case early-exit test: the mid-workflow Skill call has no continuation…testdoublewritesspike-test-setupConfigure spike test setup operations. Auto-activating skill for Performance Testing. Triggers on: spike test setup, spike test…jeremylongshorewritesspq-topic-selectionUse when deciding whether a project fits Social Psychology Quarterly (SPQ) and whether to target an Article or a Note. SPQ is the…brycewang-stanfordspringboot-tddTest-driven development for Spring Boot using JUnit 5, Mockito, MockMvc, Testcontainers, and JaCoCo. Use when adding features…mturacspringboot-tddTest-driven development for Spring Boot using JUnit 5, Mockito, MockMvc, Testcontainers, and JaCoCo. Use when adding features…vibeevalspuerbarkeit-zwischenstaatlichkeit-ssnip-testWenn es um Spürbarkeit und Zwischenstaatlichkeit in Kartellrecht — Marktabgrenzungsprüfung geht: ordnet Akteninhalt, Belege…Klotzkettespy-setup-helperAssist with spy setup helper operations. Auto-activating skill for Test Automation. Triggers on: spy setup helper, spy setup…jeremylongshorewritesssnip-test-anwendungWenn es um SSNIP-Test — Anwendung in Kartellrecht — Marktabgrenzungsprüfung geht: prüft Frist, Form, Zuständigkeit, Rechtsweg und…Klotzkettestack-trace-go-probeInternal helper for meta-stack-trace-investigator. Use when a Go panic or stack trace needs Go-specific nil/error checks, go test…opensquillastack-trace-python-probeInternal helper for meta-stack-trace-investigator. Use when a Python traceback needs Python-specific root-cause checks, pytest…opensquillastack-trace-rust-probeInternal helper for meta-stack-trace-investigator. Use when a Rust panic or backtrace needs Rust-specific Result/Option checks…opensquillastackblitz-ci-integrationCI testing for WebContainer apps with Playwright browser tests. Use when working with WebContainers or StackBlitz SDK. 'jeremylongshorewritesstaff-schedulerBuild optimized staff schedules that match coverage to demand while respecting labor budgets, employee availability, and local…cosmicstack-labsstartup-idea-validatorPressure-test a startup idea the way a sharp investor or co-founder would — problem, market, wedge, moat, why-now, and the…mohitagw15856stata-skill-contributorGuide for contributing to the stata-skill project. Use when the user wants to run the eval pipeline, analyze test results…brycewang-stanfordstatistical-software-qaQuality assurance and testing protocols for statistical softwarebrycewang-stanfordstatistical-test-selectorSkill for selecting appropriate statistical tests for analysesa5c-aiwritesstatistical-testingApply statistical hypothesis testing, significance analysis, A/B test evaluation, and distribution comparisons for data science…a5c-aiwritesstatsmodels-statistical-modelingPython statistical modeling: regression (OLS, WLS, GLM), discrete (Logit, Poisson, NegBin), time series (ARIMA, SARIMAX, VAR)…BioTender-maxstory-origin-checkRecover the first public timestamp and canonical major coverage for a newsjacking signal, then decide whether newer coverage is…elvisunstrategy-red-teamRed-team a PRD, roadmap, or strategy by attacking its load-bearing assumptions before reality does. Steelmans then attacks each…phurynstress-test-configConfigure stress test config operations. Auto-activating skill for Performance Testing. Triggers on: stress test config, stress…jeremylongshorewritesstub-creatorCreate stub creator operations. Auto-activating skill for Test Automation. Triggers on: stub creator, stub creator Part of the…jeremylongshorewritesstyled_constrained_qaAnswer questions based on provided text with strict limits on sentence/paragraph count and per-sentence word length, adapting…ECNU-ICALKswe-benchRun SWE-bench instances with an OpenSquilla agent inside the official Docker images. Trigger when the user wants to…opensquillasynthesis-exploreSynthesis EXPLORE stage for the Diffmode growth-tactics pipeline (fuses the blind-combination draw + emergent-mechanism…acogoodt-demo-run-allDiscover all non-live Demo E2E tests and run them one file at a time in the main session, diagnosing and fixing failures inline…timzaakwritest-demo-runRun a single demo E2E test file, diagnose failures, dispatch fixes to agents, and re-run until pass.timzaakwritest-web-demo-run-allDiscover all non-live Demo E2E tests and run them one file at a time in the main session, diagnosing and fixing failures inline…timzaakwritest-web-demo-runRun a single demo E2E test file, diagnose failures, dispatch fixes to agents, and re-run until pass.timzaakwritestddThe test-first gate. Routed to by /feature (after the spec is approved), /fix, and /refactor before any implementation code is…arbiterForgetdd-bug-fixFix bugs using red-green-refactor — reproduce the bug as a failing test first, then fix it. Use when fixing bugs to ensure they…rshankraswritestdd-guideTest-first development route for TDD, writing failing tests first, RED -> GREEN -> REFACTOR, and behavior-changing…foryourhealth111-pixeltdd-masteryTest-driven development workflow with Red-Green-Refactor cycle across languagesrohitg00tdd-orchestratorMaster TDD orchestrator specializing in red-green-refactor discipline, multi-agent workflow coordination, and comprehensive…sickn33tddTest-driven development workflow with philosophy guide - plan → write tests → implement → validateparcadeitdd-refactor-guardPre-refactor safety checklist. Verifies test coverage exists before AI modifies existing code. Use before asking AI to refactor…rshankraswritestdd-repairTest-Driven Repair — given a failing test, spawn a bounded headless `claude -p` (Read/Edit/Bash only) that makes the test pass…ruvnetwritestddDrive a change through a red-green-refactor loop - failing test first, minimal code to pass, then clean up. Use when implementing…rohitg00
← Prev34 / 36Next →
How the catalog works
What is an agent skill?

A folder with a SKILL.md inside — instructions, and often scripts and assets, that an AI agent loads when the task matches. Claude Code, Codex, Cursor and Copilot all read the same format, so one skill usually works across them.

Where does this catalog come from?

We read 660 source repositories straight from their file trees rather than from submitted listings — what you see is what is actually published. 98 repositories were rejected because they advertise skills but contain none: link lists, not folders.

Why is there no install counter?

Because install counts live in the registry that serves `npx skills add`, and that is not ours — publishing a number we cannot verify would be worse than showing none. Instead we show where a skill comes from and whether attention around its source is actually growing, measured from our own weekly snapshots.

Do you deduplicate?

Yes, and it matters more than expected. Aggregator repositories republish the same skill in several places — one source carried 6,317 SKILL.md files for 2,001 actual skills. We collapse by folder name and keep the canonical copy, so the catalog counts things, not copies.

Keep going