Agent skills

Testing & QA skills

Read straight from the source repositories, not from submitted listings. Every skill shows what it does, what is inside, where it came from — and whether attention around its source is actually growing.

Toolclaude-code 29,140codex 4,755cursor 3,111copilot 976windsurf 55cline 34
CategoryWorkflow & Productivity 4,979AI & Agents 3,037Data & Analytics 2,345Code Review & Quality 1,376Backend & API 1,244Security 1,194Design & Presentation 1,154Documentation 965Content & Marketing 916Testing & QA 777DevOps & Cloud 576Databases 550Frontend 469Business & Finance 328Media & Video 257Other 9,833
1,723 found
9611,008 · page 21 / 36
ab-test-setupStructured guide for setting up A/B tests with mandatory gates for hypothesis, metrics, and execution readiness.sickn33ab-test-store-listingWhen the user wants to A/B test App Store product page elements to improve conversion rate. Also use when the user mentions "A/B…Eronredabtretung-adversarial-testWenn es um Abtretung in AGB-Recht-Prüfer geht: ordnet Sachverhalt, Norm, Beweislast, Gegenargumente und nächsten Schritt; liefert…Klotzketteaccessibility-test-planCreate accessibility testing plans covering assistive technologies and WCAG criteria.Owl-Listeneracr-identificationUse when the empirical identification strategy is the bottleneck for a 《会计研究》 (Accounting Research) manuscript — exogenous…brycewang-stanfordad-campaign-analyzerUse this skill when the user shares ad campaign performance data and asks what to cut, scale, or test. Trigger for prompts like…githubad-copyWrite platform-native paid ad copy with multiple angles to test. Use when asked to write ad copy…mohitagw15856adaptive_multi_level_simplifierSimplifies complex academic, historical, political, economic, technical, consulting, hydrological, and physics concepts across…ECNU-ICALKadobe-load-scaleImplement load testing, auto-scaling, and capacity planning for Adobe API integrations with k6 scripts targeting Firefly, PDF…jeremylongshorewritesadobe-local-dev-loopConfigure Adobe local development with App Builder CLI, Runtime actions, hot reload, and mock testing for Firefly/PDF/Photoshop…jeremylongshorewritesads-testDesign and evaluate paid-ad experiments with hypotheses, randomization units, sample-size and duration assumptions, guardrails…AgriciDanielads-validateValidate Claude Ads contracts, scoring inputs, run bundles, capabilities, source freshness, safety, installation, uninstall, or…AgriciDanieladversarial-test-agbWenn es um Adversarial Test AGB in AGB-Recht-Prüfer geht: ordnet Akteninhalt, Belege, Lücken und Nachforderungen; liefert eine…Klotzketteagent-browserAutomates browser interactions for web testing, form filling, screenshots, and data extraction. Use when the user needs to…actionbookagent-browserAutomates browser interactions for web testing, form filling, screenshots, and data extraction. Use when the user needs to…code-yeongyuagent-browserBrowse the web for any task — research topics, read articles, interact with web apps, fill forms, take screenshots, extract data…nanocoaiagent-browserBrowser automation CLI for AI agents. Use when the user needs to inspect, test, or automate browser behavior: navigating pages…nexu-ioagent-qa-testingAgent davranis testi ve protokol uyumluluk dogrulamasi. Agent'larin tanimli rollerine uygun davranip davranmadigini…vibeevalagent-test-long-runnerAgent skill for test-long-runner - invoke with $agent-test-long-runnerruvnetagent-workspace-linuxUse when a task needs an isolated hidden Linux desktop or workspace-owned browser: GUI app QA, web/browser/shopping automation…agent-shagentic-evalPatterns and techniques for evaluating and improving AI agent outputs. Use this skill when: - Implementing self-critique and…githubagf-running-apple-sitUse when apple-dev has finished feature code + Unit tests (Swift Testing) and is about to enter code-review. Provides the Apple…pcliangxagf-running-sit-testsUse when an execution-layer dev (frontend-dev / backend-dev / ai-agent-dev / ml-engineer / miniapp-dev) has finished feature code…pcliangxagf-writing-qa-reportUse when qa-engineer (or miniapp-qa-engineer) is about to publish an E2E or UAT report. Provides the report skeleton…pcliangxagsy-topic-selectionUse when deciding whether a project fits Agricultural Systems (AgSy) and which article type to target. AgSy is a systems-science…brycewang-stanfordahr-topic-selectionUse when deciding whether a history project fits The American Historical Review (AHR) and how to frame its significance. The AHR…brycewang-stanfordai-slop-cleanerPost-implementation cleanup that removes AI-generated bloat while preserving functionality. Runs pass-by-pass with test…vibeevalai-visibility-panel-designSelect QA-approved canonical intent cells into a versioned AI-visibility tracking panel with partitions, variants, lanes…elvisunaistats-related-workUse when positioning an AISTATS submission against AI, machine-learning, statistics, and uncertainty literature, including arXiv…brycewang-stanfordajs-topic-selectionUse when deciding whether a sociology project fits the American Journal of Sociology (AJS) and how to frame it. AJS is the…brycewang-stanfordalgolia-local-dev-loopConfigure Algolia local development with separate dev index, mocking, and testing. Use when setting up a development environment…jeremylongshorewritesamann-literature-synthesisUse when systematically gathering, coding, and synthesizing a management/organization literature for an Academy of Management…brycewang-stanfordamann-revisionUse when responding to an Academy of Management Annals (Annals) proposal decision or full-review referee letter —…brycewang-stanfordamann-tables-figuresUse when building exhibits for an Academy of Management Annals (Annals) review — the signature framework figure, who-studied-what…brycewang-stanfordamanthro-topic-selectionUse when deciding whether an anthropology project fits American Anthropologist (AA) and which section to target. AA is the…brycewang-stanfordamr-theory-developmentUse when building the actual theory for an Academy of Management Review (AMR) manuscript — turning a positioned puzzle into…brycewang-stanfordanth-ci-integrationConfigure CI/CD pipelines for Anthropic Claude API integrations. Use when setting up automated testing, prompt regression tests…jeremylongshorewritesanth-load-scaleImplement load testing, auto-scaling, and capacity planning for Claude API. Use when running performance benchmarks, planning for…jeremylongshorewritesanth-local-dev-loopConfigure a local development workflow for Anthropic Claude API projects. Use when setting up dev environment, configuring hot…jeremylongshorewritestesting-anti-patternsReviews test code to identify and fix common testing anti-patterns including flaky tests, over-mocking, brittle assertions, test…rohitg00aos-tables-figuresUse when building the exhibits of an Accounting, Organizations and Society (AOS) manuscript — data-inventory and evidence tables…brycewang-stanfordAP Prep Book Selection based on Author Expertise and CrammabilityEvaluates and recommends Advanced Placement (AP) exam preparation books by strictly prioritizing author credentials (AP readers…ECNU-ICALKAP Prep Book Selection by Author Expertise and CrammabilitySelects AP exam preparation books by evaluating author credentials (specifically AP teaching, reading, or consulting experience)…ECNU-ICALKapi-test-generatorGenerate api test generator operations. Auto-activating skill for Test Automation. Triggers on: api test generator, api test…jeremylongshorewritesapi-testingUse when testing HTTP endpoints, probing URLs for status and headers, or running declarative API test suites from plain text filesjeremylongshorewritesapollo-hello-worldCreate a minimal working Apollo.io example. Use when starting a new Apollo integration, testing your setup, or learning basic…jeremylongshorewritesapp-analyticsWhen the user wants to set up, interpret, or improve their app analytics and tracking. Also use when the user mentions…Eronredapp-icon-optimizationWhen the user wants to design, test, or improve their app icon to increase tap-through rate and conversions in App Store search…Eronred
← Prev21 / 36Next →
How the catalog works
What is an agent skill?

A folder with a SKILL.md inside — instructions, and often scripts and assets, that an AI agent loads when the task matches. Claude Code, Codex, Cursor and Copilot all read the same format, so one skill usually works across them.

Where does this catalog come from?

We read 660 source repositories straight from their file trees rather than from submitted listings — what you see is what is actually published. 98 repositories were rejected because they advertise skills but contain none: link lists, not folders.

Why is there no install counter?

Because install counts live in the registry that serves `npx skills add`, and that is not ours — publishing a number we cannot verify would be worse than showing none. Instead we show where a skill comes from and whether attention around its source is actually growing, measured from our own weekly snapshots.

Do you deduplicate?

Yes, and it matters more than expected. Aggregator repositories republish the same skill in several places — one source carried 6,317 SKILL.md files for 2,001 actual skills. We collapse by folder name and keep the canonical copy, so the catalog counts things, not copies.

Keep going