Agent skills

Testing & QA skills

Read straight from the source repositories, not from submitted listings. Every skill shows what it does, what is inside, where it came from — and whether attention around its source is actually growing.

Toolclaude-code 29,140codex 4,755cursor 3,111copilot 976windsurf 55cline 34
CategoryWorkflow & Productivity 4,979AI & Agents 3,037Data & Analytics 2,345Code Review & Quality 1,376Backend & API 1,244Security 1,194Design & Presentation 1,154Documentation 965Content & Marketing 916Testing & QA 777DevOps & Cloud 576Databases 550Frontend 469Business & Finance 328Media & Video 257Other 9,833
1,723 found
769816 · page 17 / 36
patiently-aiPatiently AI simplifies medical documents for patients. Takes doctor's letters, test results, prescriptions, discharge summaries…FreedomIntelligenceperformance-directionProvider-independent performance direction for generated video, avatar/spokesperson clips, animation, ads, film scenes, social…calesthiopersonal-board-of-directorsFive standing advisors — the Operator, the Skeptic, the CFO, the Coach, the Customer — debate your decision on paper and vote.…mohitagw15856pester-should-migrationExperimental (preview) Pester skill for migrating classic Should -Be (v5) assertion syntax to the new Should-* (v6) assertions…githubphone-agentRun a real-time AI phone agent using Twilio, Deepgram, and ElevenLabs. Handles incoming calls, transcribes audio, generates…majiayu000phx-helpChoose Phoenix review, plan, debug, or test command. Use when user asks which /skill:phx-* command or plugin skill handles a…oliver-kriskaphx-verifyVerify Elixir/Phoenix changes — compile, format, and test in one loop. Use after implementation, before PRs, or after fixing bugs.oliver-kriskaplan-writingTransform research findings into actionable implementation plans with stakes-based rigor, test-first strategy, and granular task…a5c-aiwritesplaywrightPlaywright E2E testing, page objects, fixtures, visual regression, accessibility testing, and CI integration patterns.a5c-aiwritesPlaywright E2E TestingDeep integration with Playwright for browser automation and end-to-end testinga5c-aiwritesplaywright-electron-configConfigure Playwright for comprehensive Electron application testing including E2E tests, visual regression, accessibility audits…a5c-aiwritespolicy-qaAnswer policy questions with strict citations. Refuses to answer without sources.majiayu000analyze-projectUse when starting work on an unfamiliar project or needing to understand a codebase - performs comprehensive analysis discovering…majiayu000post-ocr-cleanupClean post-OCR text: correction, QA, multilingual handling, provenance.brycewang-stanfordpre-motion-andrew-birdAdversarial premortem for England & Wales civil litigation - builds the strongest version of a case, then attacks it from four…lawve-aipremium-web-build-gatePremium website pre-build gate for cinematic, editorial, Squarespace-polished, Raycast/Linear/Vercel-precise, or motion-heavy web…TheGoat395press-coverage-page-generatorWhen the user wants to create a press coverage page, "As Seen In" section, or media mentions aggregation. Also use when the user…kostja94press-mediaGet press coverage and media attention for your indie app. Covers press kit preparation, finding journalists/bloggers/YouTubers…rshankraswritesprice-optimization-toolEvaluate ecommerce price candidates using unit economics, historical observations, elasticity analysis, scenario modeling, and…nexscope-aiprocedural-canvas-animationProvider-independent production guidance for deterministic Canvas 2D and p5.js animation. Use for particles, fields, trails…calesthioproduct-appeal-analyzerEvaluate product desirability, market positioning, and emotional resonance—the complement to friction analysis. Assess whether…majiayu000writesproduction-design-directionProvider-independent production design direction for AI-generated media. Use when translating narrative, brand, advertising…calesthioprompt-engineeringOptimize prompts for LLMs and AI systems with structured techniques, evaluation patterns, and synthetic test data generation. Use…majiayu000prompt-governanceUse when managing prompts in production at scale: versioning prompts, running A/B tests on prompts, building prompt registries…majiayu000prompt-labIterate on LLM prompts with structured evaluation and self-correction. Test prompts against ground truth, compare models, track…majiayu000writesprompt-optimizerA/B test CLAUDE.md instruction changes against eval benchmarks. Capture baselines, test variants, compare results.majiayu000prompt-optimizerA/B test CLAUDE.md instruction changes against eval benchmarks. Capture baselines, test variants, compare results.majiayu000prompt-regression-testerCompares old vs new prompts across test cases with diff summaries, stability metrics, breakage analysis, and fix suggestions. Use…majiayu000prompt-tunerImprove embedded LLM system prompt based on evaluation test failuresmajiayu000promptfooPromptfoo evaluation framework for testing and comparing LLM outputs. Use when writing eval configs, creating test cases…majiayu000writesproof-apiBuild API test suites — endpoint testing, contract testing, load testing for REST/GraphQL/gRPC APIs. Use when asked to "test this…jeremylongshorewritesproof-e2eBuild E2E test specs for critical user journeys — Playwright or Cypress, page objects, setup/teardown, CI config. Use when asked…jeremylongshorewritesproof-reconTesting reconnaissance — inventory all tests, frameworks, coverage, CI integration, and assess testing maturity for project…jeremylongshorewritesproof-strategyProduce a test strategy for a project or feature — risk map, test type decisions, coverage targets, CI config. Use when asked to…jeremylongshorewritespydantic-evalsTest and evaluate AI agents and LLM outputs using code-first evaluation framework with strong typing. Use when the user wants to…majiayu000pysamGenomic file toolkit. Read/write SAM/BAM/CRAM alignments, VCF/BCF variants, FASTA/FASTQ sequences, extract regions, calculate…majiayu000pytest-ml-testerML-specific testing skill using pytest with fixtures for data, models, and predictions.a5c-aiwritespytest-skillGenerates production-grade pytest tests in Python with fixtures, parametrize, markers, mocking, and conftest patterns. Use when…sickn33pytest TestingExpert pytest framework for Python unit, integration, and functional testinga5c-aiwritespython-testing-patternsImplement comprehensive testing strategies with pytest, fixtures, mocking, and test-driven development. Use when writing Python…sickn33qt-test-fixture-generatorGenerate Qt Test fixtures with mock QObject signals and slots, data-driven tests, and GUI testing setupa5c-aiwritesquality-checklistValidate implementation quality through custom checklists, scoring against constitution standards, specification coverage, and…a5c-aiwritesquantized-exportExport a promoted fine-tuned model in the right deployment format — merged safetensors, LoRA-only, GGUF with imatrix, or FP8. Use…wshobsonquery-decompositionQuery decomposition for multi-concept retrieval. Use when handling complex queries spanning multiple topics, implementing…majiayu000qwen3-ttsProduce text-to-speech with Alibaba/Qwen Qwen3-TTS through DashScope/Model Studio or open-weight Qwen3-TTS checkpoints. Use when…calesthiorag-patternsRetrieval-Augmented Generation patterns and best practices. Implement chunking, embedding, retrieval, reranking, and generation…majiayu000RAN Reinforcement Learning EngineerReinforcement learning engineering for RAN systems with policy gradients, experience replay, and AgentDB integration. Implements…majiayu000rdd-analysisEconometrics skill for Regression Discontinuity Design (RDD). Activates when the user asks about: "regression discontinuity"…brycewang-stanford
← Prev17 / 36Next →
How the catalog works
What is an agent skill?

A folder with a SKILL.md inside — instructions, and often scripts and assets, that an AI agent loads when the task matches. Claude Code, Codex, Cursor and Copilot all read the same format, so one skill usually works across them.

Where does this catalog come from?

We read 660 source repositories straight from their file trees rather than from submitted listings — what you see is what is actually published. 98 repositories were rejected because they advertise skills but contain none: link lists, not folders.

Why is there no install counter?

Because install counts live in the registry that serves `npx skills add`, and that is not ours — publishing a number we cannot verify would be worse than showing none. Instead we show where a skill comes from and whether attention around its source is actually growing, measured from our own weekly snapshots.

Do you deduplicate?

Yes, and it matters more than expected. Aggregator repositories republish the same skill in several places — one source carried 6,317 SKILL.md files for 2,001 actual skills. We collapse by folder name and keep the canonical copy, so the catalog counts things, not copies.

Keep going