Agent skills

Testing & QA skills

Read straight from the source repositories, not from submitted listings. Every skill shows what it does, what is inside, where it came from — and whether attention around its source is actually growing.

Toolclaude-code 29,140codex 4,755cursor 3,111copilot 976windsurf 55cline 34
CategoryWorkflow & Productivity 4,979AI & Agents 3,037Data & Analytics 2,345Code Review & Quality 1,376Backend & API 1,244Security 1,194Design & Presentation 1,154Documentation 965Content & Marketing 916Testing & QA 777DevOps & Cloud 576Databases 550Frontend 469Business & Finance 328Media & Video 257Other 9,833
1,723 found
1,0091,056 · page 22 / 36
app-preview-videoWhen the user wants to plan, script, produce, or optimize App Store Preview videos or Google Play promo videos — the autoplay…Eronredappium-skillGenerates production-grade Appium mobile automation scripts for Android and iOS in Java, Python, or JavaScript. Supports real…sickn33applyAssisted job-application filler. Given a job posting URL, open it in a browser via Playwright, extract the JD and company, score…extrasmall0apsr-topic-selectionUse when deciding whether a political-science project fits the American Political Science Review (APSR) and which of its five…brycewang-stanfordarecon-literature-synthesisUse when systematically gathering, reading, and synthesizing a body of economic research for an Annual Review of Economics (ARE)…brycewang-stanfordarecon-revisionUse when responding to Editorial Committee and referee feedback on an Annual Review of Economics (ARE) review — coverage gaps…brycewang-stanfordarpsych-comprehensiveness-and-balanceUse when checking that an Annual Review of Psychology (ARPsych) review covers the literature even-handedly — across labs…brycewang-stanfordarpsych-literature-synthesisUse when systematically gathering, reading, and structuring the literature for an Annual Review of Psychology (ARPsych) review so…brycewang-stanfordarpsych-organizing-frameworkUse when imposing an analytical structure or taxonomy on a psychology literature for an Annual Review of Psychology (ARPsych)…brycewang-stanfordarpsych-proposal-and-commissioningUse when getting a topic in front of the Annual Review of Psychology (ARPsych) Editorial Committee, suggesting a topic/author, or…brycewang-stanfordarsoc-comprehensiveness-and-balanceUse when calibrating completeness vs. selectivity and ensuring fairness across theoretical schools, methods, authors, and debates…brycewang-stanfordarsoc-literature-synthesisUse when systematically gathering, reading, and synthesizing a large body of sociological research for an Annual Review of…brycewang-stanfordarsoc-revisionUse when responding to Editorial Committee and referee feedback on an Annual Review of Sociology (ARSoc) review — coverage gaps…brycewang-stanfordarsoc-transparency-and-reproducibilityUse when documenting the coverage account of an Annual Review of Sociology (ARSoc) review and the transparency obligations of any…brycewang-stanfordartbull-topic-selectionUse when deciding whether an art-history project fits The Art Bulletin and how to frame its contribution. The Art Bulletin is the…brycewang-stanfordasc-testflight-orchestrationOrchestrate TestFlight distribution, groups, testers, and What to Test notes using asc. Use when rolling out betas.rorkaiase-writing-styleUse when shaping the prose and structure of an ASE (IEEE/ACM Automated Software Engineering) research paper, covering the…brycewang-stanfordasplos-topic-selectionUse when deciding whether a project belongs at ASPLOS or at a single-community venue — applying the cross-layer deletion test…brycewang-stanfordasr-topic-selectionUse when deciding whether a sociology project fits the American Sociological Review (ASR) and whether to write a full Article or…brycewang-stanfordassemblyai-local-dev-loopConfigure AssemblyAI local development with hot reload and testing. Use when setting up a development environment, configuring…jeremylongshorewritesassessment-design-guidePsychometrics and educational assessment design for researchersbrycewang-stanfordassumption-bountyExtract every hidden assumption from a plan or document and put a price on each one — what it costs if wrong, what it costs to…mohitagw15856async-insteadConvert a meeting into async work that actually decides — the doc-plus-deadline format that replaces the room, the comment-window…mohitagw15856attio-ci-integrationConfigure CI/CD pipelines for Attio integrations with GitHub Actions, mock-based unit tests, and live API integration tests. "CI…jeremylongshorewritesattribution-reconcilerUse when platform-reported conversions disagree with GA4/ecommerce, when you suspect Meta and Google are double-counting the same…aaron-he-zhuAudio Dataset Loading and STFT Feature ExtractionLoad audio files from a directory, parse labels from filenames, generate random VAD segments, extract STFT features (mean along…ECNU-ICALKauto-arenaAutomatically evaluate and compare multiple AI models or agents without pre-existing test data. Generates test queries from a…agentscope-aiauto-prUse when the user invokes `/auto-pr <repo-url>` or asks to "open N PRs against <repo>", "auto-contribute to <repo>", or "raise…vouchdevauto-researchResearch uncertain questions with an explicit, user-approved web search or ChatGPT consultation, then present options and wait…sickn33Automated Unit Root Testing in RProvides a single R command or function to perform batch unit root testing (ADF, PP, DF-GLS) on multiple variables across…ECNU-ICALKawt-e2e-testingAI-powered E2E web testing — eyes and hands for AI coding tools. Declarative YAML scenarios, Playwright execution, visual…sickn33azure-microsoft-playwright-testing-tsRun Playwright tests at scale using Azure Playwright Workspaces (formerly Microsoft Playwright Testing). Use when scaling browser…microsoftazure-microsoft-playwright-testing-tsRun Playwright tests at scale with cloud-hosted browsers and integrated Azure portal reporting.sickn33azure-resource-manager-playwright-dotnetAzure Resource Manager SDK for Microsoft Playwright Testing in .NET. Use for MANAGEMENT PLANE operations: creating/managing…microsoftazure-resource-manager-playwright-dotnetAzure Resource Manager SDK for Microsoft Playwright Testing in .NET.sickn33backend-developmentBackend API design, database architecture, microservices patterns, and test-driven development. Use for designing APIs, database…MoizIbnYousafbackpropBug → spec protocol. When a bug is found or a test fails, trace the cause, decide whether a new §V invariant would catch…JuliusBrusseebamboohr-local-dev-loopConfigure BambooHR local development with hot reload, mocking, and testing. Use when setting up a development environment…jeremylongshorewritesBatch Unit Root Testing in RGenerate a reusable R script to perform ADF, PP, and DF-GLS unit root tests across multiple variables, transformations…ECNU-ICALKbdd-container-updateUpdate the BDD test container image when its dependencies or runtime change. Determines version bump (major vs minor vs patch)…NVIDIA-AI-Blueprintsbedtools-genomic-intervalsGenomic interval ops on BED/BAM/GFF/VCF. Find overlaps, merge intervals, compute coverage, extract FASTA, find nearest features.…BioTender-maxBilingual Statistics Q&AProvides answers to statistics questions in both English and Chinese.ECNU-ICALKbinlog-generationGenerate MSBuild binary logs (binlogs) for build diagnostics and analysis. USE FOR: adding /bl:{} to any dotnet build, test…dotnetbiocompatibility-test-selectorBiocompatibility test selection and protocol recommendation skill based on device categorizationa5c-aiwritesbjps-topic-selectionUse when deciding whether a political-science project fits the British Journal of Political Science (BJPS) and which of its three…brycewang-stanfordborzoiPredict genome-wide functional tracks (RNA-seq, CAGE, DNase, ChIP) from DNA sequence with Borzoi. Use this skill when: (1)…BioTender-maxborzoiPredict genome-wide functional tracks (RNA-seq, CAGE, DNase, ChIP) from DNA sequence with Borzoi. Use this skill when: (1)…xuzhougengbrainstormAI DevKit · Use when the user asks to brainstorm, ideate, generate ideas, expand options, challenge ideas, pressure-test ideas…codeaholicguy
← Prev22 / 36Next →
How the catalog works
What is an agent skill?

A folder with a SKILL.md inside — instructions, and often scripts and assets, that an AI agent loads when the task matches. Claude Code, Codex, Cursor and Copilot all read the same format, so one skill usually works across them.

Where does this catalog come from?

We read 660 source repositories straight from their file trees rather than from submitted listings — what you see is what is actually published. 98 repositories were rejected because they advertise skills but contain none: link lists, not folders.

Why is there no install counter?

Because install counts live in the registry that serves `npx skills add`, and that is not ours — publishing a number we cannot verify would be worse than showing none. Instead we show where a skill comes from and whether attention around its source is actually growing, measured from our own weekly snapshots.

Do you deduplicate?

Yes, and it matters more than expected. Aggregator repositories republish the same skill in several places — one source carried 6,317 SKILL.md files for 2,001 actual skills. We collapse by folder name and keep the canonical copy, so the catalog counts things, not copies.

Keep going