Agent skills

Testing & QA skills

Read straight from the source repositories, not from submitted listings. Every skill shows what it does, what is inside, where it came from — and whether attention around its source is actually growing.

Toolclaude-code 29,140codex 4,755cursor 3,111copilot 976windsurf 55cline 34
CategoryWorkflow & Productivity 4,979AI & Agents 3,037Data & Analytics 2,345Code Review & Quality 1,376Backend & API 1,244Security 1,194Design & Presentation 1,154Documentation 965Content & Marketing 916Testing & QA 777DevOps & Cloud 576Databases 550Frontend 469Business & Finance 328Media & Video 257Other 9,833
1,723 found
1,1051,152 · page 24 / 36
click-test-planDesign click/first-click tests to evaluate navigation and information findability.Owl-Listenerclickup-hello-worldMake your first ClickUp API v2 calls: list workspaces, spaces, and create a task. Use when starting a new ClickUp integration…jeremylongshorewritesclinical-dialogue-agents-guidePapers on AI agents for clinical dialogue and medical QAbrycewang-stanfordcnvkit-copy-numberDetect somatic CNVs from WES/WGS/targeted BAMs (CNVkit v0.9.x). Bin coverage in target/antitarget regions, normalize vs…BioTender-maxcode-healthScans the codebase for dead code, tech debt, outdated dependencies, and code quality issues. Delegates to the Centinela (QA)…davepooncode-showcase-systematic-debuggingFour-phase debugging methodology with root cause analysis. Use when investigating bugs, fixing test failures, or troubleshooting…sickn33code-showcase-testing-patternsJest testing patterns, factory functions, mocking strategies, and TDD workflow. Use when writing unit tests, creating test…sickn33code-taskPREFERRED way to change code in a REAL repository: fix a GitHub issue, fix a bug, add/implement a function or feature, or make…opensquillacoderabbit-observabilityMonitor CodeRabbit review effectiveness with metrics, dashboards, and alerts. Use when tracking review coverage, measuring…jeremylongshorewritescogpsych-theory-and-hypothesesUse when stating the theory, formalizing the model, and deriving predictions for a Cognitive Psychology (Elsevier) manuscript.…brycewang-stanfordcohere-local-dev-loopConfigure Cohere local development with mocking, testing, and hot reload. Use when setting up a development environment…jeremylongshorewritescolm-experimentsUse when designing or auditing the empirical core of a COLM paper — contamination analysis for evaluation data, fair baselines…brycewang-stanfordcolm-topic-selectionUse when deciding whether language-model research belongs at COLM or should route to ACL/EMNLP, ICLR, NeurIPS, ICML, or a…brycewang-stanfordcommit-gateThe only path to a commit. Routed to when the user invokes /commit or otherwise instructs codeArbiter to persist staged changes.…arbiterForgecompetitor-pagesCompetitor page gap analysis — take a query you want to win and the pages currently outranking you, and produce a concrete brief…nowork-studiocompliance-reviewUse this skill at Step 10 of the v2 SOP, executed by the Compliance-Agent (whose backing model must be heterogeneous from the…charliehzmconbio-topic-selectionUse when deciding whether a project fits Conservation Biology and which article type to target. Conservation Biology is the…brycewang-stanfordcontract-test-creatorCreate contract test creator operations. Auto-activating skill for Test Automation. Triggers on: contract test creator, contract…jeremylongshorewritesconversion-optimizationWhen the user wants to improve conversion rates, run A/B tests, optimize funnels, or reduce friction. Also use when the user…kostja94conversion-value-mapperUse when the user asks to "set up conversion values so tROAS optimizes profit not orders", "map margin onto my purchase value"…aaron-he-zhuConvert JSON Q&A to alternating line text fileGenerates Python code to convert a JSON dataset containing questions and answers into a text file with questions and answers on…ECNU-ICALKcoreweave-hello-worldDeploy a GPU workload on CoreWeave with kubectl. Use when running your first GPU job, testing inference, or verifying CoreWeave…jeremylongshorewritescoverage-report-analyzerAnalyze coverage report analyzer operations. Auto-activating skill for Test Automation. Triggers on: coverage report analyzer…jeremylongshorewritescoverage-trackerRun a Google Alerts-style keyword coverage tracker. Uses news-search for recent keyword queries, lets the LLM dedupe and classify…elvisuncoverage-tracker-setupSet up a lightweight Google Alerts-style coverage tracker for any number of keywords. Creates a tracker config with each keyword…elvisuncpp-testingUse only when writing/updating/fixing C++ tests, configuring GoogleTest/CTest, diagnosing failing or flaky tests, or adding…mturaccrabboxDetect and use Crabbox for repository tests and validation on remote runners. Use when crabbox.yaml or .crabbox.yaml exists, the…openclawcrap-scoreCalculates targeted CRAP (Change Risk Anti-Patterns) scores for a named .NET method, class, or single source file. Use when the…dotnetCreate Complex Summarization Prompts for Model EvaluationGenerate complex, challenging prompts for Math, Physics, Law, Chemistry, or Biology to test 'Summarization - No Guidance'…ECNU-ICALKcreate-skillScaffolds new agent skills for the dotnet/skills repository. Use when creating a new skill, generating SKILL.md files, writing a…dotnetcreate-skill-testScaffolds eval.yaml evaluation specs for agent skills in the dotnet/skills repository. Use when creating skill tests, writing…dotnetcreate-test-runCreate a Kobiton test run from a test case or suite, then offer to monitor it. When the user gives only partial details (or just…jeremylongshorecreating-oracle-to-postgres-migration-integration-testsCreates integration test cases for .NET data access artifacts during Oracle-to-PostgreSQL database migrations. Generates…githubcreative-testing-frameworkDesign structured ad creative tests with A/B test plans, multivariate creative strategies, sample size calculations, and…indranilbanerjeecrim-topic-selectionUse when deciding whether a project fits Criminology (ASC / Wiley) and whether to target a full Article or a Research Note.…brycewang-stanfordcsharp-testingC# and .NET testing patterns with xUnit, FluentAssertions, mocking, integration tests, and test organization best practices.mturacctx-dispatchCoordinate nontrivial work in the ctx repository when a task has independent investigation, implementation, or test lanes…stevesoluncurranthro-topic-selectionUse when deciding whether an anthropology project fits Current Anthropology (CA) and which article type to target. CA is the…brycewang-stanfordcurriculum-gap-analysisIdentify gaps, overlaps, and misalignments in curriculum coverage through systematic comparison with standards and learning…a5c-aiwritescypress-skillGenerates production-grade Cypress E2E and component tests in JavaScript or TypeScript. Supports local execution and TestMu AI…sickn33data-finderFind and assess datasets for a research question. Dispatches Explorer agents to search across data source categories, then…brycewang-stanfordwritesdata-quality-checksDesign the data quality checks for a table or pipeline across the standard dimensions. Use when asked to add data quality tests…mohitagw15856database-test-helperAssist with database test helper operations. Auto-activating skill for Test Automation. Triggers on: database test helper…jeremylongshorewritesdatabase-testingTest database operations — schema migration validation, query correctness, transaction integrity, and data consistency checks.a5c-aiwritesdb-seedGenerate database seed scripts with realistic sample data. Reads Drizzle schemas or SQL migrations, respects foreign key…jezwebwritesdbt-test-creatorCreate dbt test creator operations. Auto-activating skill for Data Pipelines. Triggers on: dbt test creator, dbt test creator…jeremylongshorewritesddd-aggregateScaffold an aggregate root with entity, value objects, repository interface, domain events, and test stubs. Use when adding a new…ruvnetwritesdebateStructured AI debate templates and synthesis. Use when orchestrating multi-round debates between AI tools, 'debate topic', 'argue…agent-sh
← Prev24 / 36Next →
How the catalog works
What is an agent skill?

A folder with a SKILL.md inside — instructions, and often scripts and assets, that an AI agent loads when the task matches. Claude Code, Codex, Cursor and Copilot all read the same format, so one skill usually works across them.

Where does this catalog come from?

We read 660 source repositories straight from their file trees rather than from submitted listings — what you see is what is actually published. 98 repositories were rejected because they advertise skills but contain none: link lists, not folders.

Why is there no install counter?

Because install counts live in the registry that serves `npx skills add`, and that is not ours — publishing a number we cannot verify would be worse than showing none. Instead we show where a skill comes from and whether attention around its source is actually growing, measured from our own weekly snapshots.

Do you deduplicate?

Yes, and it matters more than expected. Aggregator repositories republish the same skill in several places — one source carried 6,317 SKILL.md files for 2,001 actual skills. We collapse by folder name and keep the canonical copy, so the catalog counts things, not copies.

Keep going