mcpbeat Sign in

Playwright Regression Testing Agent Skill

Govern Playwright TypeScript regression suites across many tests. Use when asked to plan, select, tier, execute, or optimize suites with risk/change analysis, tags, CI/CD, sharding, flaky-test management, or suite-health metrics; not for authoring one UI spec. Keywords: regression strategy, smoke tests, test selection, CI pipeline, flaky tests, test sharding, impact analysis, git diff.

9k tokens
context cost
the whole folder, loaded on every use
8
files
instructions only
0
copies elsewhere
how many repositories repackaged it
209
stars on the repo
on the repository, not the skill itself

Install

one command, takes just this skill from the repository
npx skills add https://github.com/fugazi/test-automation-skills-agents --skill playwright-regression-testing

The instruction itself

8 sections, as written by the author

Playwright Regression Testing (TypeScript)

Strategy and best practices for automated regression testing of web applications using Playwright with TypeScript.

> Activation: This skill is triggered when working with regression test strategy, test suite selection, test prioritization, CI/CD pipeline testing, flaky test management, test sharding, or optimizing test execution for web applications using Playwright.

When to Use This Skill

  • Plan regression suites with risk-based and change-based test selection
  • Organize tests into tiers (smoke, sanity, selective, full regression)
  • Optimize execution with parallelization, sharding, and time-budget strategies
  • Integrate with CI/CD using GitHub Actions pipelines
  • Manage flaky tests with quarantine, retry policies, and root cause tracking
  • Monitor suite health with execution time, flake rate, and detection metrics
  • Select tests after changes using git diff analysis and impact mapping

Do NOT Use For

  • Authoring a single UI spec or page-object model (use playwright-e2e-testing).
  • Driving a live browser interactively for debugging (use playwright-cli).
  • Selenium/Java regression suites (use webapp-selenium-testing).
  • API contract testing in isolation (use api-testing).

Prerequisites

| Requirement | Details |

| -------------- | ---------------------------------------- |

| Node.js | v18+ recommended |

| Playwright | @playwright/test package |

| TypeScript | typescript configured in project |

| Browsers | Installed via npx playwright install |

| Git | Required for change-based test selection |

| GitHub Actions | Recommended CI/CD platform |


Quick Reference

Tiers: Smoke (<2min, every commit) → Sanity (<10min, every PR) → Selective (<30min, on merge) → Full (<60min, nightly/pre-release).

Key tags: @smoke, @sanity, @regression, @critical, @slow, @quarantine, @a11y.

CLI: npx playwright test --grep @smoke | --grep @regression | --grep-invert @quarantine | --shard=1/4 | --last-failed

For full tier model, regression types table, and tag taxonomy, see references/regression-catalogs.md.


Red Flags

  • Treating flaky tests as "fixed" by adding retries or waitForTimeout — quarantine and root-cause instead.
  • Running the full suite on every commit — use tiered selection (smoke on commit, full nightly).
  • Quarantining tests silently with no tracking ticket — quarantine must be temporary and owned.
  • No change-based selection — running everything regardless of what changed wastes CI budget.
  • Ignoring suite-health metrics (rising duration, climbing flake rate) until they block releases.

References

| Document | Content |

| ----------------------------------------------------------------- | ------------------------------------------------------------------------------------------------------------------------------------------------ |

| Regression Strategy | Tier model (smoke→full), regression types, triggers, directory layout, test tagging and tag taxonomy |

| Regression Selection | Test selection (change-based, risk-based, historical, time-budget) and test naming conventions |

| Regression Best Practices | Locator priority, web-first assertions, test independence, test.step() reporting, complete worked example test |

| CI/CD Integration | GitHub Actions tiered pipeline, sharding, merge reports, Playwright config, performance optimization, CLI reference |

| Flaky Management | Retry policies, quarantine strategies, detection checklist, suite health metrics, troubleshooting |


Verification

  • [ ] Smoke test subset identified — Tagged @smoke tests run in under 2 minutes
  • [ ] No test duplication — Each scenario tested exactly once at the appropriate level
  • [ ] Test isolation verified — Running tests in random order produces same results as sequential
  • [ ] Flaky test baseline established — All tests pass 5/5 consecutive runs

Other skills for the same job

different authors, same section of the catalogue
Webapp Testing
by anthropics
vendor ×12

Toolkit for interacting with and testing local web applications using Playwright. Supports verifying frontend functionality, debugging UI behavior, capturing browser screenshots, and viewing browser logs.

6k tokens scripts
Azure Microsoft Playwright Testing Ts
by lingxling
×1

Run Playwright tests at scale with cloud-hosted browsers and integrated Azure portal reporting.

2k tokens
Chrome Devtools
by christophacham
×1

Browser debugging, performance profiling, and automation via Chrome DevTools MCP. Use when user says "debug this page", "take a screenshot", "check network requests", "profile performance", "inspect console errors", or "analyze page load". Do NOT use for full E2E test suites (use playwright-skill) or non-browser debugging.

1k tokens
QA
by browser-use

QA-test a website or web app and return a 1-5 quality score (5 = flawless, 1 = broken) with evidence. Use when the user wants to test, QA, evaluate, score, or "check how good" a site, page, flow, or app — including a local dev server (e.g. "qa test localhost:5173", "does the checkout work?", "rate this landing page"). Drives a real Browser Use cloud browser, tunneling localhost automatically.

9k tokens
Playwright Component Testing
by microsoft
vendor

Set up component testing with Playwright using a story gallery — scaffold stories and a gallery dev page driven by the built-in mount fixture, no dedicated component-testing runtime. Use when asked to test React or Vue components in isolation with Playwright, or to migrate off @playwright/experimental-ct-react / -vue.

8k tokens scripts
Browser Testing With Devtools
by addyosmani

Tests in real browsers via Chrome DevTools MCP. Use when building or debugging anything that runs in a browser. Use when you need to inspect the DOM, capture console errors, analyze network requests, profile performance, or verify visual output with real runtime data. Requires the chrome-devtools MCP server to be configured.

4k tokens
Browser Harness
by browser-use

Always use browser-harness for any web interaction: automation, scraping, testing, or site/app work.

493k tokens scripts
Help Center UI Test
by Automattic

Run a browser-based UI review of the WordPress.com Help Center across multiple surfaces, looking for visual and behavioral issues. Use when asked to test the Help Center UI.

2k tokens

How to use it

Copy the folder

Take fugazi/playwright-regression-testing from the repository into ~/.claude/skills for personal use, or into .claude/skills inside a project.

Check the name does not clash

The agent identifies a skill by the name field in its header. Two skills with the same name cannot sit side by side — one of them will be ignored.

Install what it needs

The instructions reference npx. Without those the skill loads but fails at the first command.