Agent skills

Testing & QA skills

Read straight from the source repositories, not from submitted listings. Every skill shows what it does, what is inside, where it came from — and whether attention around its source is actually growing.

Toolclaude-code 29,140codex 4,755cursor 3,111copilot 976windsurf 55cline 34
CategoryWorkflow & Productivity 4,979AI & Agents 3,037Data & Analytics 2,345Code Review & Quality 1,376Backend & API 1,244Security 1,194Design & Presentation 1,154Documentation 965Content & Marketing 916Testing & QA 777DevOps & Cloud 576Databases 550Frontend 469Business & Finance 328Media & Video 257Other 9,833
1,723 found
1,4411,488 · page 31 / 36
percom-topic-selectionUse when deciding whether a pervasive-computing project belongs at IEEE PerCom or should be routed to ACM UbiComp/IMWUT, MobiCom…brycewang-stanfordperformance-test-designerPerformance test design skill for test planning, data collection, and acceptance criteria verificationa5c-aiwritesperformance-testingLoad, stress, spike testing with k6/Locust, bottleneck analysis, and performance test automationcosmicstack-labsperl-testingPerl testing patterns using Test2::V0, Test::More, prove runner, mocking, coverage with Devel::Cover, and TDD methodology.mturacperplexity-load-scaleLoad test Perplexity Sonar API integrations and plan capacity. Use when running performance tests, planning for traffic growth…jeremylongshorewritesperplexity-local-dev-loopConfigure Perplexity local development with mocking, testing, and hot reload. Use when setting up a development environment…jeremylongshorewritespersona-hello-worldCreate your first Persona identity verification inquiry and check its status. Use when learning Persona API basics, testing…jeremylongshorewritesphase-gated-commitsPhase-gated commit workflow for clean git history -- implement, review, test, commit per phasevibeevalphg-rebuttalUse when responding to a Progress in Human Geography (PiHG) decision letter (major/minor revision) — building a point-by-point…brycewang-stanfordphi-desensitizeUse this skill before any prompt, log message, error trace, test fixture, or external API call that may include patient health…charliehzmplan-a-feature-to-confluenceBuilds a feature specification from scratch with plan-a-feature and publishes it to a user-specified Confluence location, posting…testdoublewritesplan-compassStress-tests a plan through dependency-aware, easy-to-answer decision prompts. Use when the user wants a plan stress-test, plan…softcaneplan-interrogateStress-test a plan by walking its decision tree one question at a time. Use when the user wants to pressure-test a design before…rohitg00planning-oracle-to-postgres-migration-integration-testingCreates an integration testing plan for .NET data access artifacts during Oracle-to-PostgreSQL database migrations. Analyzes a…githubplatform-detectionReference data for detecting the test platform (VSTest vs Microsoft.Testing.Platform) and test framework (MSTest, xUnit, NUnit…dotnetplaywright-automation-fill-in-formAutomate filling in a form using Playwright MCPgithubplaywright-explore-websiteWebsite exploration for testing using Playwright MCPgithubplaywright-generate-testGenerate a Playwright test based on a scenario using Playwright MCPgithubГенерация регулярного выражения для миграции Playwright (page.click -> page.locator)Создает регулярное выражение для VS Code для массовой замены конструкции page.click('selector') на…ECNU-ICALKpmf-strategyWhen the user wants to validate product-market fit, measure PMF, or plan before scaling. Also use when the user mentions "PMF,"…kostja94pmla-topic-selectionUse when deciding whether a literary or language-studies project fits PMLA (Publications of the Modern Language Association) and…brycewang-stanfordpnas-fitUse before any writing begins to stress-test whether a result clears PNAS's bar — high quality and broad significance to a…brycewang-stanfordpnas-statisticsUse to enforce PNAS's statistics and reproducibility reporting — n and replication, test choice and assumptions, effect sizes…brycewang-stanfordpnas-workflowUse when deciding which pnas-* sub-skill to invoke next, or when sequencing a manuscript from significance test through reviewer…brycewang-stanfordpnasnexus-fitUse before any writing — as the first gate — to stress-test whether a result fits PNAS Nexus — high-quality, broadly significant…brycewang-stanfordpnasnexus-statisticsUse to enforce PNAS Nexus's statistics and reproducibility reporting — n and replication, test choice and assumptions, effect…brycewang-stanfordpnasnexus-workflowUse when deciding which pnasnexus-* sub-skill to invoke next, or when sequencing a manuscript from scope/significance test…brycewang-stanfordpods-topic-selectionUse when deciding whether a data-management project belongs at ACM PODS (the database-theory symposium) or should be routed to…brycewang-stanfordpolicy-renewal-reviewRun a pre-renewal review of an insurance programme: scan coverage gaps against current operations, test limit adequacy against…mohitagw15856popdevr-topic-selectionUse when deciding whether a project fits Population and Development Review (PDR, Wiley / Population Council) and which article…brycewang-stanfordpopl-experimentsUse when designing the empirical component of a POPL paper — deciding whether evidence should be a mechanization, a prototype…brycewang-stanfordpoq-survey-design-and-measurementUse when defending the survey design and measurement of a Public Opinion Quarterly (POQ) manuscript through the Total Survey…brycewang-stanfordpoq-topic-selectionUse when deciding whether a project fits Public Opinion Quarterly (POQ) and which submission type to target. POQ is the leading…brycewang-stanfordposthog-local-dev-loopConfigure PostHog local development with mocking, debug mode, and testing. Use when setting up a development environment, mocking…jeremylongshorewritesppsych-revisionUse when responding to a Perspectives on Psychological Science (PoPS) editor and reviewer decision letter — coverage gaps…brycewang-stanfordpr-reviewReviews a GitHub pull request and posts inline comments plus one consolidated summary, adapting to any codebase by discovering…tech-leads-clubpraktikabilitaet-vollzug-testWenn es um NKR-Praktikabilitaet im Vollzug in Normenkontrollrat (NKR) — Prüfung von Gesetzentwuerfen geht: prüft Frist, Form…Klotzketteprd-writer輕量版 PRD 撰寫工具 — 適合個人專案、單一功能、無合規/金流/風控需求、單團隊開發, 快速產出可直接交付工程的「施工藍圖」等級文件。 當使用者說「幫我寫 PRD」、「產品需求文件」、「寫 spec」、「功能規格」、「write a PRD」、…skinnerlee1225press-and-prWhen the user wants to get press coverage, media mentions, or editorial features for their app — including writing press…Eronredpricing-testTest pricing strategies with synthetic data. Use when: simulating willingness to pay, price sensitivity, or optimal price points.indranilbanerjeeprioritize-assumptionsPrioritize assumptions using an Impact × Risk matrix and suggest experiments for each. Use when triaging a list of assumptions…phurynproduct-lensUse this skill to validate the "why" before building, run product diagnostics, and pressure-test product direction before the…mturacproduct-page-optimizationGenerates A/B test plans and optimization checklists for your App Store product page — icon, screenshots, and app previews. Use…rshankraswritesprompt-optimizerDiagnose and rewrite an underperforming LLM prompt so it produces reliable, well-structured output. Use when asked to improve a…mohitagw15856prompt-proximity-architectureTurn an approved measurement charter, ICPs, and buyer jobs into a budget-aware prompt coverage blueprint across proximity bands…elvisunprompt-regression-suiteDesign a regression test suite that catches an LLM feature getting worse when the prompt, model, or context changes. Use when…mohitagw15856prompt-set-qaGate a prompt universe for schema and provenance completeness, target or campaign contamination, evidence entailment…elvisunprompt-testA/B test content variations. Use when: comparing quality scores across prompt approaches, headline styles, or content versions.indranilbanerjee
← Prev31 / 36Next →
How the catalog works
What is an agent skill?

A folder with a SKILL.md inside — instructions, and often scripts and assets, that an AI agent loads when the task matches. Claude Code, Codex, Cursor and Copilot all read the same format, so one skill usually works across them.

Where does this catalog come from?

We read 660 source repositories straight from their file trees rather than from submitted listings — what you see is what is actually published. 98 repositories were rejected because they advertise skills but contain none: link lists, not folders.

Why is there no install counter?

Because install counts live in the registry that serves `npx skills add`, and that is not ours — publishing a number we cannot verify would be worse than showing none. Instead we show where a skill comes from and whether attention around its source is actually growing, measured from our own weekly snapshots.

Do you deduplicate?

Yes, and it matters more than expected. Aggregator repositories republish the same skill in several places — one source carried 6,317 SKILL.md files for 2,001 actual skills. We collapse by folder name and keep the canonical copy, so the catalog counts things, not copies.

Keep going