Agent skills

Testing & QA skills

Read straight from the source repositories, not from submitted listings. Every skill shows what it does, what is inside, where it came from — and whether attention around its source is actually growing.

Toolclaude-code 29,140codex 4,755cursor 3,111copilot 976windsurf 55cline 34
CategoryWorkflow & Productivity 4,979AI & Agents 3,037Data & Analytics 2,345Code Review & Quality 1,376Backend & API 1,244Security 1,194Design & Presentation 1,154Documentation 965Content & Marketing 916Testing & QA 777DevOps & Cloud 576Databases 550Frontend 469Business & Finance 328Media & Video 257Other 9,833
1,723 found
289336 · page 7 / 36
prompt-engineerWrites, refactors, and evaluates prompts for LLMs — generating optimized prompt templates, structured output schemas, evaluation…Jeffallanrun-smoke-testsInspect an unfamiliar repository, interpret a broad Markdown user journey at runtime, operate the real product through its…tamdogoodsentry-load-scaleScale Sentry for high-traffic applications handling millions of events per day. Use when optimizing SDK performance at high…jeremylongshorewritessplitting-datasetsProcess split datasets into training, validation, and testing sets for ML model development. Use when requesting "split dataset"…jeremylongshorewritestesting-load-balancersValidate load balancer behavior, failover, and traffic distribution. Use when performing specialized testing. Trigger with…jeremylongshorewritescausal-inference-mixtapeThis skill should be used when the user asks to "implement a DiD regression", "write a causal inference pipeline", "set up an…brycewang-stanfordgame-theory-paper-writerGenerate, continue, revise, polish, and stress-test game theory research papers. Use when the user provides a game theory topic…brycewang-stanfordchange-traceability-reviewUse for change traceability review across specs, commits, PRs, branches, local diffs, and git history, including Spec…QoderAIdeliver-acceptance-criteriaGenerates structured Given/When/Then acceptance criteria for a user story or feature slice, covering the happy path, key failure…product-on-purposedeliver-edge-casesDocuments edge cases, error states, boundary conditions, race conditions, and recovery paths for a feature - the systematic…product-on-purposedigital-health-clinical-asr-setupStage 1 of Clinical ASR Flywheel. Use when bootstrapping a cycle: NVCF+MW disclosure, NVIDIA_API_KEY check, deps install, TTS+ASR…NVIDIAeducation-data-explorerDiscovers education data from Urban Institute Portal: endpoints, variables, year coverage, join keys (CCD, IPEDS, CRDC…brycewang-stanfordeducation-data-source-nacuboNACUBO endowment data (~650 institutions, 2012-2022). Portal: 7 columns only (total endowment, per-FTE, YoY change). Use for…brycewang-stanfordexam2knowledgeReverse-engineers exam questions into high-frequency test points and reusable solving patterns. Use for past papers, practice…mingchen666genotoxicGraph-informed mutation testing triage. Parses codebases with Trailmark, runs mutation testing and necessist, then uses survived…trailofbitsgolang-refactoringGolang refactoring — the safe, at-scale process for restructuring existing Go code: a coverage-adaptive safety net, tool-driven…samberwritesgolang-testingProduction-ready Golang tests — table-driven tests, testify suites and mocks, parallel tests, fuzzing, fixtures, goroutine leak…samberwritesgrowing-outside-in-systemsDrive feature development using Outside-In TDD with Hexagonal Architecture. Design emerges through inline code, in-memory fakes…a5c-aihow-to-write-componentUse when implementing or refactoring React/TypeScript components and the task requires decisions about component ownership…langgeniushsb-testExecute QA test plans on Holoscan Sensor Bridge hardware. Reads a user-provided test document, filters tests by the user's setup…NVIDIAwritescommunity-import-smoke-testA portable community plugin for validating Open Design plugin import flows.nexu-iojohnny-suede-designDesign and write polished product surfaces people understand fast: landing pages, dashboards, campaigns, restyles, UI copy, and…JasonColapietrokeyword-researchUse when the user asks to "find keywords", "挖词", or "搜什么词"; prioritizes search volume, keyword difficulty, intent, and topic…aaron-he-zhuklingai-hello-worldCreate your first Kling AI video generation with a minimal working example. Use when learning Kling AI or testing your setup.…jeremylongshorewriteslangchain-ci-integrationWire LangChain 1.0 / LangGraph 1.0 tests into a GitHub Actions pipeline\ \ \u2014\nunit tests with FakeListChatModel, VCR-gated…jeremylongshorewriteslangchain-local-dev-loopBuild a fast, deterministic local test loop for LangChain 1.0 / LangGraph\ \ 1.0\n\u2014 FakeListChatModel fixtures, pytest…jeremylongshorewriteslangchain-prompt-engineeringManage LangChain 1.0 prompts like code \u2014 LangSmith prompt hub versioning,\n\ XML-tag conventions for Claude, few-shot…jeremylongshorewritesmatlab-integrate-antennasIntegrate antennas into RF systems using MATLAB Antenna Toolbox and RF Toolbox. Covers impedance matching network design…matlabnovel-to-gameTurn a novel into a fully playable game on the selected target platform. Orchestrates the whole adaptation pipeline —…worldwondererphx-workExecute Elixir/Phoenix plan tasks with progress tracking. Use after /skill:phx-plan to implement features with mix compile and…oliver-kriskaplaywright-expertUse when writing E2E tests with Playwright, setting up test infrastructure, or debugging flaky browser tests. Invoke to write…Jeffallanpython-proUse when building Python 3.11+ applications requiring type safety, async programming, or robust error handling. Generates…Jeffallanrails-expertRails 7+ specialist that optimizes Active Record queries with includes/eager_load, implements Turbo Frames and Turbo Streams for…Jeffallanrender-glassy-matte-grwmAssemble a multi-scene GRWM beauty-demo ad from a config — a locked-identity creator applies ~5 products step by step while a…gooseworks-airender-multiworldAssemble a silent, music-led 3-world product-tour ad — trim and hard-cut-concat the per-world WIDE-arrival + top-down-macro…gooseworks-airender-podcast-skitAssemble a two-host fake-podcast skit ad from a config — per-line lipsync clips hard-concatenated in script order, scaled/padded…gooseworks-airender-stopmotion-hand-swatch-cycleAssemble a stop-motion hand-swatch-cycle product-demo ad from a config — a sequence of still PLATES (one hand swiping a…gooseworks-airender-vo-anchored-motion-listicleAssemble an expert/educator motion-graphic LISTICLE video ad from a config — a spoken authoritative voiceover carries a numbered…gooseworks-aisemgrep-rule-variant-creatorCreates language variants of existing Semgrep rules. Use when porting a Semgrep rule to specified target languages. Takes an…trailofbitswritesstatsmodelsStatistical models library for Python. Use when you need specific model classes (OLS, GLM, mixed models, ARIMA) with detailed…LeonChaoXsustainable-opposing-counsel-reviewProduces an adversarial attack on a legal argument that survives reply. Runs the opposing-counsel discipline twice: an…lawve-aitddTest-driven development with red-green-refactor loop. Use when user wants to build features or fix bugs using TDD, mentions…stevesoluntest-guardReview generated or changed test code against universal testing rules before it ships. Best used reactively after an agent…amElnagdytestingVitest testing guide. Use when writing or updating tests, fixing failing tests, improving coverage, debugging test issues, or…lobehubtraceUse when encountering bugs, test failures, runtime errors, broken builds, or "this doesn't work" reports. Systematic root-cause…jeremylongshorevercel-load-scaleLoad test and scale Vercel deployments with concurrency tuning and capacity planning. Use when running performance tests…jeremylongshorewritesvue-expert-jsCreates Vue 3 components, builds vanilla JS composables, configures Vite projects, and sets up routing and state management using…JeffallanworkExecute Elixir/Phoenix plan tasks with progress tracking. Use after /phx:plan to implement features with mix compile and mix test…oliver-kriska
← Prev7 / 36Next →
How the catalog works
What is an agent skill?

A folder with a SKILL.md inside — instructions, and often scripts and assets, that an AI agent loads when the task matches. Claude Code, Codex, Cursor and Copilot all read the same format, so one skill usually works across them.

Where does this catalog come from?

We read 660 source repositories straight from their file trees rather than from submitted listings — what you see is what is actually published. 98 repositories were rejected because they advertise skills but contain none: link lists, not folders.

Why is there no install counter?

Because install counts live in the registry that serves `npx skills add`, and that is not ours — publishing a number we cannot verify would be worse than showing none. Instead we show where a skill comes from and whether attention around its source is actually growing, measured from our own weekly snapshots.

Do you deduplicate?

Yes, and it matters more than expected. Aggregator repositories republish the same skill in several places — one source carried 6,317 SKILL.md files for 2,001 actual skills. We collapse by folder name and keep the canonical copy, so the catalog counts things, not copies.

Keep going