Agent skills

Testing & QA skills

Read straight from the source repositories, not from submitted listings. Every skill shows what it does, what is inside, where it came from — and whether attention around its source is actually growing.

Toolclaude-code 29,140codex 4,755cursor 3,111copilot 976windsurf 55cline 34
CategoryWorkflow & Productivity 4,979AI & Agents 3,037Data & Analytics 2,345Code Review & Quality 1,376Backend & API 1,244Security 1,194Design & Presentation 1,154Documentation 965Content & Marketing 916Testing & QA 777DevOps & Cloud 576Databases 550Frontend 469Business & Finance 328Media & Video 257Other 9,833
1,723 found
337384 · page 8 / 36
working-with-legacy-codeSafely change and test untested codebases using Feathers'' "Working Effectively with Legacy Code". Use when the user mentions…wondelai24-ai-avatar-production-globalAI Avatar production pipeline for global markets — 3-tier tools (Free/Pro/Enterprise), 4 workflows (single avatar, translate…minhnv0807agentsop-test-fix-loopDecision protocol for wiring a verify-then-fix loop around a code-editing LLM agent. The agent edits → runs lint/test → reads the…agentsopeai-team-orchestrationBootstrap and run a lightweight multi-agent development team. Use when starting or adopting a project, planning work…githubbio-causal-genomics-colocalization-analysisTest whether two traits share a causal variant at a genomic locus using Bayesian colocalization with coloc. Computes posterior…FreedomIntelligencebio-de-edger-basicsPerform differential expression analysis using edgeR in R/Bioconductor. Use for analyzing RNA-seq count data with the…FreedomIntelligencecompetitor-analysisUse when the user asks to "analyze competitors" or "竞品分析"; benchmarks competitor keywords, content, backlinks, AI citations, and…aaron-he-zhucontext-modeUse context-mode tools (ctx_execute, ctx_execute_file) instead of Bash/cat when processing large outputs. Triggers: "analyze…mksgludefine-hypothesisDefines a testable hypothesis with clear success metrics and a validation approach. Use when forming assumptions to test or…product-on-purposebio-de-edger-basicsPerform differential expression analysis using edgeR in R/Bioconductor. Use for analyzing RNA-seq count data with the…BioTender-maxi4h-catheter-navigation-e2eEnd-to-end smoke for catheter navigation covering setup, digital twin, DRR, and unit tests. Use when asked to run the full…NVIDIAi4h-catheter-navigation-render-drrRender a single DRR fluoroscopy frame from a CT cache or synthetic phantom. Use when asked to render DRR, generate a fluoro…NVIDIAi4h-catheter-navigation-smokeRun CPU-only fluorosim smoke tests (imports, preprocessing, CLI parsers). Use when asked to smoke-test catheter navigation in CI…NVIDIAi4h-workflow-e2eRun the full end-to-end agentic pipeline (record → mimic → annotate → replay → convert → visualize → finetune → validate). Use…NVIDIAideateTake a fuzzy idea (or an existing thing you want to improve) through a disciplined funnel — explore → pressure-test → converge —…nelsonwerdmatlab-generate-5g-waveformGenerate 3GPP-compliant 5G NR downlink and uplink baseband waveforms. Use to create NR signals, test model (TM) waveforms, fixed…matlabmcore-testingTest system for Megatron-LM. Covers test layout, recipe YAML structure, adding and running unit and functional tests, golden…NVIDIAmeasure-experiment-resultsDocuments the results of a completed experiment or A/B test with statistical analysis, learnings, and recommendations. Use after…product-on-purposemeta-tags-optimizerOptimize title tags, meta descriptions, Open Graph, and Twitter cards for maximum click-through rate. Generates multiple A/B test…ViryaZhengminiprogram-developmentWeChat Mini Program development skill for building, debugging, previewing, testing, publishing, and optimizing mini program…TencentCloudBasenemo-rl-docsDocumentation conventions for NeMo-RL. Covers docs/index.md updates and docstring format. Do NOT use for: bug fixes, test fixes…NVIDIAnew-loopSpin up a new loop (domain) in a file-based knowledge base — bootstrap the substrate if it's missing, gather the loop's charter…AI-Builder-Clubpr-review**AUTHOR SKILL (internal to microsoft/aspire-skills).** Reviews pull requests *into this repo* for problems only — bugs…microsoftpublic-relationsWhen the user wants help with public relations, earned media, press coverage, journalist outreach, or media strategy (not pull…sickn33pysamGenomic file toolkit. Read/write SAM/BAM/CRAM alignments, VCF/BCF variants, FASTA/FASTQ sequences, extract regions, calculate…FreedomIntelligencepysamGenomic file toolkit. Read/write SAM/BAM/CRAM alignments, VCF/BCF variants, FASTA/FASTQ sequences, extract regions, calculate…LeonChaoXsparring-partnerThis skill should be used whenever the user wants to develop, refine, or stress-test an idea, plan, design, research direction…QinghongLinsuede-ab-testingSuede-owned experimentation discipline for hypotheses, sample sizing, test duration, significance, and repeatable experiment…JasonColapietrotemporal-python-testingComprehensive testing approaches for Temporal workflows using pytest, progressive disclosure resources for specific testing…sickn33temporal-python-testingTest Temporal workflows with pytest, time-skipping, and mocking strategies. Covers unit testing, integration testing, replay…wshobsontest-guardReview generated or changed test code against universal testing rules before it ships or is presented for approval.sickn33testingWrite or repair Elixir tests with ExUnit, sandbox isolation, async reliability, Mox, ExMachina, and LiveViewTest. Use for test…oliver-kriskatilegym-adding-cutile-kernelAdd a new cuTile GPU kernel operator to TileGym. Covers dispatch registration in ops.py, cuTile backend implementation…NVIDIAwindsurf-test-generationGenerate comprehensive test suites using Cascade. Activate when users mention "generate tests", "test coverage", "write unit…jeremylongshorewritesab-test-plannerDesign statistically rigorous A/B tests for product features, UI changes, onboarding flows, and pricing experiments. Use when…mohitagw15856ab-testingWhen the user wants to plan, design, or implement an A/B test or experiment, or build a growth experimentation program. Also use…sickn33agentsop-bounded-loopUniversal discipline for any LM-driven loop — agent retries, plan-act-observe, multi-agent handoffs, optimiser passes, test-fix…agentsopeagentsop-regression-gateBuild a held-out eval set, run it on every prompt/model change, and block regressions in CI. An LM change is a code change — gate…agentsopebio-ecological-genomics-biodiversity-metricsCalculates species richness, diversity, and turnover using the Hill number framework with iNEXT coverage-based…BioTender-maxbuild-loopDrive a build to actually-works, near-finish-line craft by looping build → see → exercise → check → critique → rebuild over the…nelsonwerdcontent-gap-analysisUse when the user asks to "find content gaps", "竞品写了什么", or "还应该写什么"; builds a competitor-relative coverage map of missing…aaron-he-zhucontent-quality-auditorUse when auditing content quality, E-E-A-T, or publish readiness; runs a typed 80-item CORE-EEAT profile with evidence coverage…aaron-he-zhugoogle-ads-copyGenerate and A/B test Google Ads copy. Use when asked to write ad copy, headlines, descriptions, create ad variants, test ad…nowork-studiocreate-sdi-runImport data into the Simulation Data Inspector (SDI) from MAT, CSV, or Excel files, from workspace variables, or from a Simulink…matlabdurable-objectsCreate and review Cloudflare Durable Objects. Use when building stateful coordination (chat rooms, multiplayer games, booking…cloudflaree2e-cucumber-playwrightUse when writing, changing, or reviewing Cucumber and Playwright tests under `e2e/`, including feature files, step definitions…langgeniusevent-studyUse this skill whenever the user wants to conduct an event study, create event study plots, test for parallel trends, implement…brycewang-stanfordfitness-functionsWrite architecture fitness functions — deterministic tests that enforce a project's hard rules (module boundaries, offline…rshankraswrites
← Prev8 / 36Next →
How the catalog works
What is an agent skill?

A folder with a SKILL.md inside — instructions, and often scripts and assets, that an AI agent loads when the task matches. Claude Code, Codex, Cursor and Copilot all read the same format, so one skill usually works across them.

Where does this catalog come from?

We read 660 source repositories straight from their file trees rather than from submitted listings — what you see is what is actually published. 98 repositories were rejected because they advertise skills but contain none: link lists, not folders.

Why is there no install counter?

Because install counts live in the registry that serves `npx skills add`, and that is not ours — publishing a number we cannot verify would be worse than showing none. Instead we show where a skill comes from and whether attention around its source is actually growing, measured from our own weekly snapshots.

Do you deduplicate?

Yes, and it matters more than expected. Aggregator repositories republish the same skill in several places — one source carried 6,317 SKILL.md files for 2,001 actual skills. We collapse by folder name and keep the canonical copy, so the catalog counts things, not copies.

Keep going