mcpbeat

Academic Experiments

joshua-zyy/academic-experiments

Audit, run, or verify experimental evidence for CS/AI/ML papers. Produces Evidence Inventory with evidence_type annotations (newly_run/preexisting_artifact/user_claim) and Protocol Risk assessments. Use when: checking if experiment results are reproducible, auditing existing experiment artifacts, running minimal reproducible commands, evaluating checkpoints without full retraining, documenting protocol risks like data leakage or missing baselines. Triggers on: 复核实验, run experiments, 实验结果, experiment evidence, verify results, 实验验证, evidence inventory, protocol risk, 跑实验, check results, reproduce experiments, 实验审计.

6k tokens
context cost
the whole folder, loaded on every use
11
files
ships runnable scripts
0
copies elsewhere
how many repositories repackaged it
104
stars on the repo
on the repository, not the skill itself

Install

one command, takes just this skill from the repository
npx skills add https://github.com/joshua-zyy/academic-paper-writer --skill academic-experiments

What comes with it

21 071 bytes besides the instruction
agents/experiment_agent.md
manifest.yaml
references/evidence-inventory.md
references/protocol-risks.md
references/run-strategy.md
scripts/evidence_scanner.py
static/core/anti-patterns.md
static/core/output-contract.md
static/core/red-lines.md
static/core/stance.md

The instruction itself

5 sections, as written by the author

Academic Experiments

将此 skill 视为"实验取证代理",目标是建立最短且可信的证据链,而不是尽量多跑实验。

Router Protocol

  • Read manifest.yaml. It declares always_load files, axes, and references.on_demand.
  • Read every file listed under always_load. These are the skill's binding rules — not reference material.
  • Apply the loaded material as constraints:
  • stance.md defines non-negotiable rules, evidence type semantics, failure degradation, and scope.
  • red-lines.md defines absolute prohibitions. Do not negotiate these.
  • output-contract.md defines deliverables per mode and claim-readiness classification.
  • anti-patterns.md defines known failure modes and their correct alternatives.
  • Detect the mode using the manifest's mode axis: experiment-evidence-pass, evidence-inventory-only, or minimal-reproducible-run. Align evidence type semantics to ../shared/core/evidence-policy.md.
  • Echo the selected mode to the user before executing.
  • Reach for references/ only when the manifest's references.on_demand condition is satisfied.

Modes

| Mode | Use when |

|---|---|

| experiment-evidence-pass | Full audit: inventory + run + record + risk analysis |

| evidence-inventory-only | Inventory existing artifacts only, no execution |

| minimal-reproducible-run | Execute minimal reproducible command (e.g. eval existing checkpoint) |

Agent Dispatch

agents/experiment_agent.md is dispatched by academic-paper-writer orchestrator at Step 4. The agent may run experiments but must not modify project source code or data files, nor write paper prose independently.

Independent Use

| Input | Mode | Priority | Behavior |

|---|---|---|---|

| repo_path + no run mode | experiment-evidence-pass | 2 (path trigger) | Full audit: inventory → env → minimal run → risk |

| repo_path + "inspect only" | evidence-inventory-only | 1 (explicit) | Inventory only, no commands |

| repo_path + specific command | minimal-reproducible-run | 1 (explicit) | Verify env → execute → record |

| No repo_path | — | 3 (no input) | Ask path, or auto-detect entry files |

| Scenario | Recommended |

|---|---|

| Just auditing/reproducing evidence | This skill (standalone) |

| Writing results into paper prose | academic-paper-writer orchestrator |

| Draft results need verification | This skill → academic-reviser |

How to use it

Copy the folder

Take joshua-zyy/academic-experiments from the repository into ~/.claude/skills for personal use, or into .claude/skills inside a project.

Check the name does not clash

The agent identifies a skill by the name field in its header. Two skills with the same name cannot sit side by side — one of them will be ignored.