Agent skills

Testing & QA skills

Read straight from the source repositories, not from submitted listings. Every skill shows what it does, what is inside, where it came from — and whether attention around its source is actually growing.

Toolclaude-code 29,140codex 4,755cursor 3,111copilot 976windsurf 55cline 34
CategoryWorkflow & Productivity 4,979AI & Agents 3,037Data & Analytics 2,345Code Review & Quality 1,376Backend & API 1,244Security 1,194Design & Presentation 1,154Documentation 965Content & Marketing 916Testing & QA 777DevOps & Cloud 576Databases 550Frontend 469Business & Finance 328Media & Video 257Other 9,833
1,723 found
865912 · page 19 / 36
statistical-analysis-advisorRecommends appropriate statistical methods (T-test vs ANOVA, etc.) based on dataset characteristics, performs assumption…majiayu000statistical-analysisGuided statistical analysis with test selection and reporting. Use when you need help choosing appropriate tests for your data…majiayu000statistical-analysisGuided statistical analysis with test selection and reporting. Use when you need help choosing appropriate tests for your data…majiayu000statistical-analystRun hypothesis tests, analyze A/B experiment results, calculate sample sizes, and interpret statistical significance with effect…majiayu000still-image-retouching-finishingProvider-independent still-image post-production skill for RAW/rendered intake, nondestructive development, exposure, white…calesthiostress-testingStress-test plans, proposals, and strategies. Use for pre-mortems, assumption audits, risk registers, evaluating business ideas…majiayu000strict-tddStrict RED->GREEN->REFACTOR test-driven development with enforcement. Never write production code before a failing test. Atomic…a5c-aiwritesstring-database-ppiQuery STRING REST API for PPIs (59M proteins, 20B interactions, 5000+ species). Retrieve networks, run GO/KEGG enrichment, find…BioTender-maxStryker Mutation TestingStryker mutation testing for assessing test suite quality and effectivenessa5c-aiwritesstudy-notes-synthesizerTurn lecture notes, slides, and readings into one exam-ready study guide — synthesis, not summary. Use when asked to make a study…mohitagw15856subagent-driven-developmentThe implementation engine. Routed to by /sprint (full plan, autonomous) and by executing-plans (scoped batch, checkpoint-gated).…arbiterForgesubject-line-labUse when the user asks to "generate subject line variants", "pre-score my subject lines", or "will this subject get truncated /…aaron-he-zhusuede-ai-evalDesign AI evals that catch regressions before users do: rubrics, test cases, failure modes, acceptance gates, and AI-SPEC…JasonColapietrosuede-codex-fleetClaude-directed parallel OpenAI Codex CLI worker fleet for bulk generation. Use when a job is high-volume, well-specified, and…JasonColapietrosuede-designMake Suede interfaces feel intentional: tokens, color, components, type, visual hierarchy, motion, dark mode, and visual QA for…JasonColapietrosuede-launch-packagingPackage finished work so people can use it: README, docs, install commands, proof links, QA, release copy, and handoff notes.JasonColapietrosuede-mcp-qaCatch MCP drift before release: skill catalogs, tool and resource schemas, prompts, install paths, JSON-RPC behavior, and docs…JasonColapietrosuper-ralph-wiggumSuper Ralph Wiggum - autonomous iteration loops with templates, PRD support, progress tracking, and browser testing. This skill…majiayu000surge-experimentGrowth experiment design — structure a growth hypothesis, define metric, baseline, expected lift, and kill condition for a single…jeremylongshorewritesswap-analyzerAnalyze swap compatibility and safety. Use when evaluating proposed schedule swaps to ensure they maintain ACGME compliance and…majiayu000synthesia-avatar-videoProduce presenter-led AI avatar videos with Synthesia Studio and Synthesia API. Use for Synthesia-specific avatar video planning…calesthiosynthetic-controlEconometrics skill for Synthetic Control Method (SCM). Activates when the user asks about: "synthetic control", "SCM", "synthetic…brycewang-stanfordsystematic-debuggingUse when encountering any bug, test failure, or unexpected behavior, before proposing fixes. Requires root cause investigation…a5c-aisystematic-debuggingUse this skill whenever a test fails, a code review surfaces a defect, or a bug is reported, to perform structured root-cause…charliehzmtavus-replica-videoUse for producing Tavus AI-human videos and real-time avatar conversations with Tavus Faces/Replicas, PALs/Personas, async Video…calesthiotddAI DevKit · Test-driven development — write a failing test before writing production code. Use when implementing new…codeaholicguytdd-enforcementRed-Green-Refactor TDD methodology with mandatory failing tests, minimal implementation, quality refactoring, and 80% coverage…a5c-aiwritestdd-featureRed-green-refactor scaffold for building new features with TDD. Write failing tests first, then implement to pass. Use when…rshankraswritestencent-hunyuanvideoGenerate and operate Tencent Hunyuan video through the managed TokenHub HY-Video-1.5 API or official local HunyuanVideo…calesthiohive.terminal-tools-foundationsRequired reading whenever any shell_* tool is available. Teaches the foreground/background dichotomy (terminal_exec auto-promotes…aden-hivetest-analysisAnalyze creative test results using heatmaps and data visualization to identify statistical winners and recommend next steps…majiayu000test-assumptionIdentify hidden assumptions about users in code that could exclude or harm. A focused analysis asking "What am I assuming about…majiayu000test-case-generatorGenerate comprehensive test cases including edge cases, stress tests, and counter-examples for algorithm correctness…a5c-aiwritestest-coverage-analyzerAnalyze test coverage and identify gaps before migration to ensure adequate safety netsa5c-aiwritestest-data-generationSynthetic test data generation and management using Faker.js and similar tools. Generate realistic test data, create data…a5c-aiwritestest-driven-developmentTest-first development practice where test specifications are written before production code, integrated into plan tasks as…a5c-aiwritestest-driven-developmentUse when implementing any feature or bugfix, before writing implementation codeECNU-ICALKtest-driven-developmentUse when implementing any feature or bugfix, before writing implementation codeFreedomIntelligencetest-driven-developmentUse when the user explicitly requests strict or test-first TDD, or when the current conversation already contains an explicit…GanyuanRantest-driven-developmentUse when implementing any feature or bugfix, before writing implementation codeInfrasity-Labstest-driven-developmentUse when implementing any feature or bugfix, before writing implementation codesickn33test-enforcementAutomated test validation, coverage checking, and quality metrics with aggressive defaultsa5c-aiwritesthe-due-diligence-callSimulate the due-diligence call where an acquirer's or investor's analyst takes your metrics apart — the questions behind the…mohitagw15856the-price-pushbackSimulate the client who grinds on your price — the budget theater, the competitor quote, the scope squeeze — against your actual…mohitagw15856thinking-scientific-methodHypothesis → Prediction → Test → Revise with explicit falsification. Use for debugging, feature experimentation, performance…majiayu000thinking-thought-experimentTest ideas through hypothetical scenarios when empirical testing is impractical. Use for architecture evaluation, edge case…majiayu000threejs-scene-compositionProduction guidance for planning, building, animating, capturing, and reviewing complete Three.js scenes for rendered media. Use…calesthiotime-seriesEconometrics skill for time series analysis. Activates when the user asks about: "time series", "stationarity", "unit root test"…brycewang-stanford
← Prev19 / 36Next →
How the catalog works
What is an agent skill?

A folder with a SKILL.md inside — instructions, and often scripts and assets, that an AI agent loads when the task matches. Claude Code, Codex, Cursor and Copilot all read the same format, so one skill usually works across them.

Where does this catalog come from?

We read 660 source repositories straight from their file trees rather than from submitted listings — what you see is what is actually published. 98 repositories were rejected because they advertise skills but contain none: link lists, not folders.

Why is there no install counter?

Because install counts live in the registry that serves `npx skills add`, and that is not ours — publishing a number we cannot verify would be worse than showing none. Instead we show where a skill comes from and whether attention around its source is actually growing, measured from our own weekly snapshots.

Do you deduplicate?

Yes, and it matters more than expected. Aggregator repositories republish the same skill in several places — one source carried 6,317 SKILL.md files for 2,001 actual skills. We collapse by folder name and keep the canonical copy, so the catalog counts things, not copies.

Keep going