mcpbeat

Testing Claude Skills

3 535 testing skills from 425 authors. They run test suites, check accessibility and catch what broke. Half of them fit into 1 800 tokens or less — that is what one costs your context window when the agent loads it. 404 ship runnable scripts rather than instructions alone. 2 of them cannot work without an MCP server, most often rube. We also found 403 copies of these same skills sitting in other people's repositories — counted once here, not 403 times.

3 535 unique 425 authors 2 151 updated this month 319 from vendors

1 800
tokens, median
what a typical one costs in context
404
ship scripts
code that runs, not instructions alone
2
need a server
most often rube
403
copies elsewhere
counted once here, not once per repository

1 009–1 056 of 3 535

page 22 of 74
Common Tdd
by HoangNguyen0403

Implements a strict Red-Green-Refactor loop to ensure zero production code is written without a prior failing test. Use when: creating new features, fixing bugs, or expanding test coverage.

3k tokens
Quality Engineering Appium MCP
by HoangNguyen0403

Drives iOS/Android mobile devices via Appium MCP. Use for verifying mobile bugs, E2E tests, and navigating real device clouds (LambdaTest/BrowserStack).

3k tokens
Common Web Visual Testing
by HoangNguyen0403

Standardizes visual audits, responsive design, and behavioral testing for web apps. Use to verify a web UI fix or cross-browser behavior; defer backend API refactors, Playwright installation/tooling setup, and Appium/mobile automation.

5k tokens
JavaScript Tooling
by HoangNguyen0403

Development tools, linting, and testing for JavaScript projects.

958 tokens
Quality Engineering Quality Assurance
by HoangNguyen0403

Write or review manual Zephyr test cases with 1-condition-per-TC granularity, Module_Action on Screen when Condition naming, platform prefix rules, and High/Normal/Low priority classification. Use for test-case authoring and review; defer Jira traceability, linking, and pushing cases to Zephyr.

1k tokens
Quality Engineering Zephyr Coverage Analysis
by HoangNguyen0403

Audit test coverage health, gaps, and QE debt for Jira stories or epics. Produces coverage_analysis_report.md with AC-to-TC heatmap, risk scores, and prioritized action plan. Use when assessing coverage percentage, pre-release readiness, sprint readiness, or identifying missing test cases. Do NOT use for TC creation — use zephyr-test-generation instead.

1k tokens
Typescript Tooling
by HoangNguyen0403

Development tools, linting, and build config for TypeScript. Use when configuring ESLint, Prettier, Jest, Vitest, tsconfig, or any TS build tooling.

1k tokens
Battle Test
by HoangNguyen0403

Deep audit of a skills directory against the Skill Creator standard. Produces a scored report and phased remediation plan.

613 tokens
Implement Feature
by HoangNguyen0403

Implement an approved feature plan with fresh-context slices, TDD, evidence, and PR-ready output.

982 tokens
Implementation Readiness
by HoangNguyen0403

Verify BRD-lite, PRD, SRS/FRS, UX, and test prerequisites before implementation starts.

875 tokens
Incident Hotfix
by HoangNguyen0403

Mitigate a production incident or urgent regression first, then route to root-cause remediation and a postmortem.

754 tokens
Review Ticket
by HoangNguyen0403

Review a ticket or PR through focused specialist lenses: scope, architecture, security, tests, AC coverage, and PR metadata.

1k tokens
Android Concurrency
by HoangNguyen0403

Write correct coroutine scopes, lifecycle collection, and dispatcher injection in Android production code. Use for suspend functions, coroutine scopes, and dispatcher mechanics; defer ViewModel StateFlow/LiveData architecture, Fragment lifecycle recipes, persistence/notifications, and unit-test recipes to their specific skills.

1k tokens
Android Testing
by HoangNguyen0403

Write Android unit, Compose UI, and Hilt-integrated tests. Use when designing test behavior with MockK or coroutine test utilities; defer database/WorkManager-specific recipes to the owning feature skill.

1k tokens
Angular Testing
by HoangNguyen0403

Write Angular component tests using TestBed, ComponentHarness, and HttpTestingController with proper signal input handling. Use when writing component tests, mocking HTTP calls, or testing signal inputs.

2k tokens
Ac Criteria Validator
by majiayu000

Validate acceptance criteria and feature completion. Use when checking if features pass, validating test results, verifying acceptance criteria, or determining feature completion status.

860 tokens
Agent Context Generator
by majiayu000

Generate project-level AGENTS.md guides that capture conventions, workflows, and required follow-up tasks. Use when a repository needs clear agent onboarding covering structure, tooling, testing, task flow, README expectations, and conventional commit summaries.

2k tokens
Ln 23 Test Suite Auditor
by levnikolaevich

Audits whether an existing test suite proves important product behavior with strong, isolated tests. Use when test confidence is uncertain; not to design or implement new tests.

3k tokens
Ln 42 Acceptance Test Builder
by levnikolaevich

Creates and runs reproducible acceptance tests for stated requirements using project-native tooling. Use when executable acceptance evidence is needed; not for audits or product fixes.

2k tokens
Ln 41 Test Strategy Planner
by levnikolaevich

Designs a risk-based test strategy and prioritized scenarios without changing code. Use when requirements need a test plan; not for auditing or implementing tests.

2k tokens
Threejs Visual Validation
by scottstts

Validate advanced Three.js graphics as authored systems rather than subjective screenshots. Use for fixed-view visual contracts, field and pass diagnostics, no-post baselines, seed sweeps, camera-scale tests, temporal stability checks, GPU budgets, and regression evidence for procedural scenes.

3k tokens
Retrieval Practice Generator
by GarethManning

Generate retrieval practice questions at varied difficulty levels for a topic or concept. Use when creating quiz starters, revision activities, or low-stakes testing materials.

4k tokens
Accessibility Testing
by vibeeval

axe-core integration, WCAG 2.2 AA checklist, keyboard navigation testing, screen reader testing, and ARIA pattern validation.

2k tokens
Agent QA Testing
by vibeeval

Agent davranis testi ve protokol uyumluluk dogrulamasi. Agent'larin tanimli rollerine uygun davranip davranmadigini assertion-based test'lerle olcer. Personality drift, role violation ve output kalite regresyonu tespit eder.

2k tokens
Agent Tamagotchi
by vibeeval

Terminal pet that lives in your statusline. 12 species, 5 stats (DEBUGGING, PATIENCE, CHAOS, WISDOM, SPEED). Reacts to your workflow - happy when tests pass, sad when builds fail, excited during swarm mode. Deterministic species from user ID.

911 tokens
API Patterns
by vibeeval

API design, versioning, testing, schema validation, and contract testing patterns for REST and GraphQL APIs.

4k tokens
Chaos Engineering
by vibeeval

Failure injection patterns, blast radius control, steady state hypothesis, and gameday planning for resilience testing.

2k tokens
Commit Trailers
by vibeeval

Commit mesajlarina yapilandirilmis karar trailer'lari ekle. Constraint, Rejected, Confidence, Scope-risk, Not-tested trailer'lari ile karar baglamini koru.

1k tokens
Component Library Patterns
by vibeeval

Design system token management, component API design, Storybook, and visual regression testing patterns

764 tokens
Contract Testing Patterns
by vibeeval

Pact consumer-driven contracts, provider verification, schema evolution

932 tokens
Debug Hooks
by vibeeval

Systematic hook debugging workflow. Use when hooks aren't firing, producing wrong output, or behaving unexpectedly.

893 tokens
Django Tdd
by vibeeval

Django testing strategies with pytest-django, TDD methodology, factory_boy, mocking, coverage, and testing Django REST Framework APIs.

5k tokens
Experiment Loop
by vibeeval

Autonomous experiment loop: hypothesize > modify > test > evaluate > keep/discard > repeat. Run N experiments automatically with measurable metrics. Works for performance optimization, A/B testing, prompt engineering, and any measurable improvement task.

2k tokens
Factcheck Guard
by vibeeval

Use this skill when making any factual claim about the codebase — existence, absence, or behavior. Converts the claim-verification rule into a systematic action protocol that prevents false assertions from grep-only results.

2k tokens
Feature Flag Patterns
by vibeeval

Gradual rollout strategies, kill switches, A/B testing integration, flag lifecycle management, and technical debt prevention.

2k tokens
Golang Testing
by vibeeval

Go testing patterns including table-driven tests, subtests, benchmarks, fuzzing, and test coverage. Follows TDD methodology with idiomatic Go practices.

4k tokens
Hook Developer
by vibeeval

Complete Claude Code hooks reference - input/output schemas, registration, testing patterns

4k tokens
Load Testing Patterns
by vibeeval

k6 script templates, load profiles, response time thresholds, SLO validation, and performance testing strategies.

2k tokens
Mutation Testing
by vibeeval

Mutation testing ile test suite kalitesini olc. Stryker, mutmut, go-mutesting destegi.

2k tokens
Performance Testing
by vibeeval

Load testing with k6/Artillery, response time thresholds, memory leak detection, N+1 query detection, and CI integration.

1k tokens
Property Based Testing
by vibeeval

Property-based testing (PBT) patterns with fast-check (JS/TS), Hypothesis (Python), and gopter (Go). Generate random inputs, define invariants, shrink failures to minimal cases. Adapted from Trail of Bits. Use when testing pure functions, parsers, serializers, state machines, or any code where example-based tests miss edge cases.

2k tokens
Prove
by vibeeval

Formal theorem proving with research, testing, and verification phases

2k tokens
Python Testing
by vibeeval

Python testing strategies using pytest, TDD methodology, fixtures, mocking, parametrization, and coverage requirements.

5k tokens
21 Ads Audit Global
by minhnv0807

Comprehensive ads audit for Meta/Google/TikTok with Health Score 0-100. 84 checkpoints across 6 dimensions (Account/Campaign/AdSet/Ad/Tracking/Optimization). Has 4 region variants for benchmarks. INCLUDES Dropshipping Audit Checklist (creative testing velocity, CRO signals, attribution accuracy). Trigger: 'ads audit', 'ad account audit', 'Meta audit', 'Google Ads audit', 'TikTok audit', 'dropshipping ads audit'.

16k tokens
19 Ab Test Setup Global
by minhnv0807

Design valid A/B tests for global marketing — hypothesis formulation, sample size calculation, statistical significance, multi-arm testing, primary vs secondary metrics. Tools: Optimizely, VWO, Google Optimize (sunset 2023, alternatives), built-in platform tests (Meta, Google). Trigger: 'A/B test', 'split test', 'multivariate test', 'experiment design', 'statistical significance', 'sample size calculator'.

4k tokens
19 Ab Test Setup
by minhnv0807

Khi nguoi dung can thiet lap A/B test cho ads, landing page, email, hoac product. Cung dung khi nguoi dung nhac 'A/B test', 'split test', 'test creative', 'test copy', 'test landing', 'test gia', 'kiem dinh thong ke', 'sample size'. Skill nay giup chon bien can test, tinh sample size, setup tracking, va phan tich ket qua thong ke (significance) — phu hop cho team khong phai data scientist.

2k tokens
Deliver Acceptance Criteria
by product-on-purpose

Generates structured Given/When/Then acceptance criteria for a user story or feature slice, covering the happy path, key failure scenarios, and non-functional expectations in testable form. Use when turning requirements into verifiable scenarios for engineering handoff and QA sign-off. For a dedicated catalog of boundary conditions, error states, and recovery paths across a feature, use deliver-edge-cases; to write the stories themselves, use deliver-user-stories.

4k tokens
Deliver Edge Cases
by product-on-purpose

Documents edge cases, error states, boundary conditions, race conditions, and recovery paths for a feature - the systematic catalog of what can go wrong and the failure modes to design for. Use during specification to map the failure surface and ensure comprehensive coverage, or during QA planning to identify boundary and limit scenarios to test. Distinct from deliver-acceptance-criteria, which writes story-level Given/When/Then checks; this skill produces the whole-feature edge-case catalog.

6k tokens