mcpbeat Sign in

Designing Experiments Agent Skill

Design experiments and quasi-experiments before analysis. Use when choosing study design, treatment/control structure, outcomes, assumptions, validation plans after scientific experiment failure, or which of DiD, ITS, synthetic control, or regression discontinuity fits the research question. For fitting models or estimating effects on existing data, use performing-causal-analysis instead.

599 tokens
context cost
the whole folder, loaded on every use
1
files
instructions only
0
copies elsewhere
how many repositories repackaged it
2583
stars on the repo
on the repository, not the skill itself

Install

one command, takes just this skill from the repository
npx skills add https://github.com/foryourhealth111-pixel/Vibe-Skills --skill designing-experiments

The instruction itself

4 sections, as written by the author

Designing Experiments

Helps choose and specify a research design before data analysis starts. This skill owns study-design decisions: what is treated, what is compared, what outcome is measured, which assumptions are required, which validation or recovery experiment should follow a failed scientific experiment, and which design is defensible.

It does not fit causal models, estimate treatment effects, interpret fitted model output from existing data, or debug software/build failures.

Decision Framework

  • Control Group?
  • Yes: Go to Step 2.
  • No: Consider Interrupted Time Series (ITS).
  • Unit Structure?
  • Single Treated Unit:
  • With multiple controls: Synthetic Control (SC).
  • No controls: ITS.
  • Multiple Treated Units:
  • With control group: Difference-in-Differences (DiD).
  • Time Structure?
  • Panel Data (Multiple units over time): Required for DiD and SC.
  • Time Series (Single unit over time): Required for ITS.

Method Quick Reference

  • Difference-in-Differences (DiD): Compares trend changes between treated and control groups. Assumes Parallel Trends.
  • Interrupted Time Series (ITS): Analyzes trend/level change for a single unit after intervention. Assumes Trend Continuity.
  • Synthetic Control (SC): Constructs a synthetic counterfactual from weighted control units. Assumes Convex Hull (treated unit within range of controls).

Failed Experiment Recovery

When a scientific experiment or optimization plan produces weak or contradictory results, use the same design surface to:

  • Separate implementation or measurement errors from design-assumption failures.
  • Identify which assumption should be tested next.
  • Define a minimal validation experiment before abandoning the approach.
  • State the decision rule for continuing, revising, or stopping the line of work.

Other skills for the same job

different authors, same section of the catalogue
Hypogenic
by christophacham
×3

Automated LLM-driven hypothesis generation and testing on tabular datasets. Use when you want to systematically explore hypotheses about patterns in empirical data (e.g., deception detection, content analysis). Combines literature insights with data-driven hypothesis testing. For manual hypothesis formulation use hypothesis-generation; for creative ideation use scientific-brainstorming.

7k tokens
Statistical Analysis
by ComeOnOliver
×3

Statistical analysis toolkit. Hypothesis tests (t-test, ANOVA, chi-square), regression, correlation, Bayesian stats, power analysis, assumption checks, APA reporting, for academic research.

33k tokens scripts
Hypogenic
by ComeOnOliver
×2

Automated hypothesis generation and testing using large language models. Use this skill when generating scientific hypotheses from datasets, combining literature insights with empirical data, testing hypotheses against observational data, or conducting systematic hypothesis exploration for research discovery in domains like deception detection, AI content detection, mental health analysis, or other empirical research tasks.

12k tokens
Swarm Advanced
by ComeOnOliver
×2

Advanced swarm orchestration patterns for research, development, testing, and complex distributed workflows

9k tokens
Ux Researcher Designer
by ComeOnOliver
×2

UX research and design toolkit for Senior UX Designer/Researcher including data-driven persona generation, journey mapping, usability testing frameworks, and research synthesis. Use for user research, persona creation, journey mapping, and design validation.

8k tokens scripts
Statistical Analysis
by K-Dense-AI
×1

Guided statistical analysis for research data - test selection, assumption checking, effect sizes, power analysis, Bayesian alternatives, and APA-formatted reporting. Use whenever a user wants to compare groups, test a hypothesis, analyze experimental or survey data, check statistical assumptions, compute required sample sizes, or write up results - even if they never name a specific test. Covers t-tests, ANOVA, chi-square, correlation, regression, non-parametric and Bayesian methods. For low-level model APIs, see the statsmodels and pymc skills.

29k tokens scripts
Pre Mortem
by phuryn

Run a pre-mortem risk analysis on a PRD or launch plan. Categorizes risks as Tigers (real problems), Paper Tigers (overblown concerns), and Elephants (unspoken worries), then classifies as launch-blocking, fast-follow, or track. Use when preparing for launch, stress-testing a product plan, or identifying what could go wrong.

1k tokens
Wage Hour QA
by anthropics
vendor

> Jurisdiction-aware wage/hour and employment Q&A — classification, overtime, meal/rest breaks, leave, final pay — answered for the specific state/country with the controlling rule researched and cited rather than stated from memory. Use when the user asks any employment law question, or says "what's the rule in [state]", "is this exempt", "do we have to pay overtime for", or "can we classify this as".

3k tokens

How to use it

Copy the folder

Take foryourhealth111-pixel/designing-experiments from the repository into ~/.claude/skills for personal use, or into .claude/skills inside a project.

Check the name does not clash

The agent identifies a skill by the name field in its header. Two skills with the same name cannot sit side by side — one of them will be ignored.