Agent skills

Testing & QA skills

Read straight from the source repositories, not from submitted listings. Every skill shows what it does, what is inside, where it came from — and whether attention around its source is actually growing.

Toolclaude-code 29,140codex 4,755cursor 3,111copilot 976windsurf 55cline 34
CategoryWorkflow & Productivity 4,979AI & Agents 3,037Data & Analytics 2,345Code Review & Quality 1,376Backend & API 1,244Security 1,194Design & Presentation 1,154Documentation 965Content & Marketing 916Testing & QA 777DevOps & Cloud 576Databases 550Frontend 469Business & Finance 328Media & Video 257Other 9,833
1,723 found
241288 · page 6 / 36
azure-aigatewayConfigure Azure API Management as an AI Gateway for AI models, MCP tools, and agents. WHEN: semantic caching, token limit…microsoftdata-warehouse-experimentationRunning experiments out of the data warehouse instead of via dedicated experiment platforms. SQL-based assignment, exposure…rampstackcoexecuting-distributed-system-testsUse when running a previously designed distributed-systems test plan against a real or simulated cluster — driving fault…shenliinitiating-coverageFull equity research initiation: company research, financial model, valuation, charts, 30-50 page reportginlix-aimatlab-generate-gnss-waveformGenerate GNSS baseband waveforms (GPS, Galileo, NavIC) with physically realistic or user-specified channel impairments using the…matlabmatlab-generate-wlan-waveformGenerate standard-compliant IEEE 802.11 waveforms using MATLAB WLAN Toolbox. Use when creating WLAN waveforms, PPDU packets, or…matlabmatlab-testingGenerate and run MATLAB unit tests using matlab.unittest and matlab.uitest. Parameterized tests, fixtures, mocking, coverage…matlabrag-perfPerformance benchmarking for a deployed NVIDIA RAG Blueprint server: profiling pass + aiperf load test driven by a single YAML…NVIDIAstartup-positioningMarket positioning strategy using the April Dunford framework, enriched with JTBD discovery, Moore positioning statement, and…ferdinandobonsstatsmodelsStatistical models library for Python. Use when you need specific model classes (OLS, GLM, mixed models, ARIMA) with detailed…K-Dense-AIwritesads-creative-developmentHow to produce ad creative that converts at performance scale. Hook patterns, format selection, video pacing, variation systems…rampstackcoagentsop-aiderSOP for terminal-based, git-native AI pair programming with Aider (git work-tree + tree-sitter repo-map + edit-format +…agentsopeanalyzeConfirmatory hypothesis testing matched to pre-registration, with full assumption testing, effect sizes, confidence intervals…brycewang-stanfordcoverage-analysisProject-wide code coverage and CRAP (Change Risk Anti-Patterns) score analysis for .NET projects. Calculates CRAP scores per…dotnetdoca-capsUse this skill when the user wants to invoke the read-only doca_caps CLI to ask what DOCA sees on this host — listing DOCA…NVIDIAdoca-flow-dpa-perfUse this skill when the user is invoking doca_flow_dpa_perf on DPA-capable hardware (ConnectX-7 minimum supported, ConnectX-8…NVIDIAdoca-sha-offload-engineUse this skill when wiring the DOCA SHA Offload Engine (an OpenSSL ENGINE) into an existing OpenSSL pipeline to offload one-shot…NVIDIAexperimentation-platform-orchestratorA platform decision framework for experimentation. When to use Statsig vs PostHog vs GrowthBook vs Optimizely vs Amplitude vs…rampstackcointegration-orchestratorGenerate a phased delivery orchestration plan for creative-direction-driven work: which skills run when, what locks at which…rampstackconextflowBuild, run, and debug Nextflow data pipelines and nf-core workflows end to end. Use whenever the user mentions Nextflow, nf-core…K-Dense-AIovernight-developmentAutomates software development overnight using git hooks to enforce test-driven Use when appropriate context detected. Trigger…jeremylongshorewritesskill-creatorCreate new skills, modify and improve existing skills, and measure skill performance. Use when users want to create a skill from…himself65vector-forgeMutation-driven test vector generation. Finds implementations of a cryptographic algorithm or protocol, runs mutation testing to…trailofbitsandroid-app-factoryPlan, build, test, and release a production-grade native Android app from a product idea through Google Play. Use for requests to…JasonColapietroclean-codeWrite readable, maintainable code through disciplined naming, small functions, and clean error handling. Use when the user…wondelaibio-causal-genomics-colocalization-analysisTest whether two or more traits share a causal variant at a locus using Bayesian colocalization (coloc.abf, coloc.susie…BioTender-maxcudaq-guideCUDA-Q onboarding guide for installation, test programs, GPU simulation, QPU hardware, and quantum applications.NVIDIAeducation-data-source-pseoPSEO — Census data linking graduates to employment via LEHD wage records. Earnings percentiles at 1/5/10 years post-graduation by…brycewang-stanfordfigma-load-scaleLoad test Figma API integrations and plan for scale. Use when benchmarking API throughput, testing rate limit behavior, or…jeremylongshorewritesgolang-namingGo (Golang) naming conventions — covers packages, constructors, structs, interfaces, constants, enums, errors, booleans…samberwritesgolang-spf13-viperGolang configuration library using spf13/viper — layered precedence (flag > env > file > KV > default), BindPFlag/BindPFlags…samberwritesgtarsHigh-performance toolkit for genomic interval analysis in Rust with Python bindings. Use when working with genomic regions, BED…BioTender-maxgtarsHigh-performance toolkit for genomic interval analysis in Rust with Python bindings. Use when working with genomic regions, BED…foryourhealth111-pixelgtarsHigh-performance toolkit for genomic interval analysis in Rust with Python bindings. Use when working with genomic regions, BED…LeonChaoXhypothesis-generationStructured hypothesis formulation from observations. Use when you have experimental observations or data and need to formulate…foryourhealth111-pixelwriteshypothesis-generationStructured hypothesis formulation from observations. Use when you have experimental observations or data and need to formulate…LeonChaoXwritesinspired-productBuild empowered product teams using discovery and delivery dual-track. Use when the user mentions "product discovery", "empowered…wondelailegal-test-builder-patrick-munroBuilds a high-fidelity interactive legal assessment as a single self-contained HTML artifact. Output includes a live countdown…lawve-aimcp-case-overridesPer-case mocked MCP override smoke skill used by skill-up e2e tests.alibabameasure-experiment-designDesigns an A/B test or experiment with variants, success metrics, sample size, and duration for an existing hypothesis. Use when…product-on-purposemom-testTalk to customers without leading them using Mom Test rules: discuss their life not your idea, ask about specifics in the past…wondelainestjs-expertCreates and configures NestJS modules, controllers, services, DTOs, guards, and interceptors for enterprise-grade TypeScript…Jeffallanopenrouter-hello-worldSend your first OpenRouter API request and understand the response. Use when learning OpenRouter, testing setup, or verifying…jeremylongshorewritesorchestrating-test-executionTest coordinate parallel test execution across multiple environments and frameworks. Use when performing specialized testing.…jeremylongshorewritesplaywright-interactivePersistent browser and Electron interaction through `js_repl` for fast iterative UI debugging.Haohao-endplaywright-javaScaffold, write, debug, and enhance enterprise-grade Playwright E2E tests in Java using Page Object Model, JUnit 5, Allure…sickn33polars-bioHigh-performance genomic interval operations and bioinformatics file I/O on Polars DataFrames. Overlap, nearest, merge, coverage…BioTender-maxpolars-bioHigh-performance genomic interval operations and bioinformatics file I/O on Polars DataFrames. Overlap, nearest, merge, coverage…K-Dense-AIwrites
← Prev6 / 36Next →
How the catalog works
What is an agent skill?

A folder with a SKILL.md inside — instructions, and often scripts and assets, that an AI agent loads when the task matches. Claude Code, Codex, Cursor and Copilot all read the same format, so one skill usually works across them.

Where does this catalog come from?

We read 660 source repositories straight from their file trees rather than from submitted listings — what you see is what is actually published. 98 repositories were rejected because they advertise skills but contain none: link lists, not folders.

Why is there no install counter?

Because install counts live in the registry that serves `npx skills add`, and that is not ours — publishing a number we cannot verify would be worse than showing none. Instead we show where a skill comes from and whether attention around its source is actually growing, measured from our own weekly snapshots.

Do you deduplicate?

Yes, and it matters more than expected. Aggregator repositories republish the same skill in several places — one source carried 6,317 SKILL.md files for 2,001 actual skills. We collapse by folder name and keep the canonical copy, so the catalog counts things, not copies.

Keep going