evo-hq/evo-report
Read-only evo run reporting. Use when the user invokes /evo:report, asks what happened overnight, asks what improved recently, asks for the best/frontier candidates, asks for a quick score chart without opening the dashboard, or wants the scatter plot in chat output. Never run benchmarks, gates, Slurm commands, evo run, or ad-hoc verification scripts for report requests.
This is a copy. The original lives at evo-hq/report.
npx skills add https://github.com/evo-hq/evo --skill report
Report the current evo workspace from recorded state only. A report request is
read-only, even if the user phrases it casually as "what happened?", "what got
better?", "what should I pay attention to?", or "I just woke up".
Do not spend compute while reporting:
evo run, evo gate check, benchmark commands, or project evalscripts.
python bench.py, python slurm_eval.py, sbatch, srun,squeue, sacct, or scancel to verify a result.
Use stored evo state instead: evo report, evo status, evo tree,
evo frontier, evo show <id>, evo diff <id>, and immutable artifacts under
.evo/run_*/experiments/<exp>/attempts/<NNN>/.
For chart requests, render the dashboard's scatter plot as a colored terminal
block, one chart per run, sized to the current terminal.
Mirrors the web dashboard's score scatter (left rail of evo dashboard):
pruned with prune_kind=exhausted can still be best; prune_kind=invalid and its descendants cannot.Every run in the workspace is rendered, stacked top-to-bottom, with a header line showing run_id · target · metric.
Run:
evo report
That is it. Print the output verbatim in your reply so the user sees the chart. Do not summarize the chart in prose — the visual is the point.
Flags:
--color always|never|auto — force or suppress ANSI color. Default auto (color when stdout is a TTY). Pass --color always if you are piping through a host that strips TTY but renders ANSI in chat.--watch [SECONDS] — live-refresh mode (like nvidia-smi -l). Re-reads the workspace every N seconds (default 2) and redraws in place. Ctrl-C to exit. Use this when you want to babysit a running optimization without manually re-invoking the report.evo status or evo show <id> is faster.evo tree is the right command.evo dashboard instead.When the user asks what happened recently or what improved, summarize from
recorded evo state:
evo status, evo frontier, and evo tree.evo show <id> for the best node and any recent committed/evaluatednodes you mention.
evo diff <id> only to explain what changed in a recorded experiment.outcome.json,benchmark.log, or declared artifacts for that experiment. Treat missing
artifacts as "not recorded", not as permission to rerun.
Report:
If the user wants fresh validation or reruns, ask them to explicitly start a new
optimization or evaluation command. Do not infer that from a report request.
Take evo-hq/evo-report from the repository into ~/.claude/skills for personal
use, or into .claude/skills inside a project.
The agent identifies a skill by the name field in its header. Two skills with the
same name cannot sit side by side — one of them will be ignored.