willoscar/unit-executor
Execute exactly one eligible Unit in an existing research Workspace; use for stepwise or manual semantic execution when status, Attempt, Artifact, Manifest, checkpoint, and acceptance evidence must remain synchronized.
npx skills add https://github.com/WILLOSCAR/research-units-pipeline-skills --skill unit-executor
The leading principle is atomicity: one invocation owns one Unit Attempt and
either commits one accepted Completion or records one diagnosable block. It
never starts a second Unit.
UNITS.csv and the selected Unit row.inputs field.DECISIONS.md when the Unit is checkpoint-gated.outputs field.UNITS.csv, Run Evidence, and optional STATUS.md projection.output/QUALITY_GATE.md when strict quality checks block Completion.Inspect the Workspace through the Pipeline adapter. Select the requested Unit,
or the first TODO Unit whose dependencies are DONE. Stop when a HUMAN
checkpoint, unresolved Decision, open Attempt, or integrity failure prevents
selection.
Completion criterion: exactly one eligible Unit is selected, or one blocking
condition is recorded with a concrete next action.
Start semantic work through the adapter, never by editing a status cell:
uv run python scripts/pipeline.py mark \
--workspace workspaces/<name> \
--unit-id <U###> \
--status DOING \
--note "starting semantic execution"
Completion criterion: the Unit is DOING and one matching open Attempt owns
the execution.
Read the selected Unit's Skill and only the context pointers required by this
branch. Produce the declared outputs without changing unrelated Workspace
artifacts.
Completion criterion: every required output exists or the failure is specific
enough to commit as BLOCKED.
Evaluate the Unit acceptance rule and strict quality contract when requested.
Commit through the adapter:
uv run python scripts/pipeline.py mark \
--workspace workspaces/<name> \
--unit-id <U###> \
--status DONE \
--note "acceptance checked"
Use BLOCKED with a concrete reason when acceptance fails. Do not directly
edit UNITS.csv; the adapter aligns Attempt, Artifact, Manifest, Decision, and
status projections.
Completion criterion: Completion is DONE with acceptance and provenance
evidence, or BLOCKED with a diagnosable Failure.
Refresh the Workspace projection and report the completed or blocked Unit. Do
not claim end-to-end completion and do not start the next eligible Unit.
Completion criterion: exactly one Unit changed execution state during this
invocation and the next operator can resume from Workspace files.
UNITS.csv owns dependencies, inputs, outputs,acceptance, checkpoint, and Skill identity.
research-pipeline-runner for automatic continuation across Units.uv run python .codex/skills/unit-executor/scripts/run.py \
--workspace workspaces/<name>
--workspace <path>: existing Workspace.--unit-id <U###>: execute a specific eligible Unit.--inputs, --outputs, --checkpoint: Pipeline-runner compatibilityarguments.
--strict: block scaffold-like outputs and write the quality-gate report.Run exactly one strict Unit:
uv run python .codex/skills/unit-executor/scripts/run.py \
--workspace workspaces/<name> \
--strict
Equivalent adapter command:
uv run python scripts/pipeline.py run-one \
--workspace workspaces/<name> \
--strict
The helper returns 0 for DONE or IDLE, and 2 for BLOCKED or ERROR.
open Attempts before changing status.
DONE Unit has missing outputs, reopen it through the adapter with anexplanatory note; never repair the CSV projection alone.
Take willoscar/unit-executor from the repository into ~/.claude/skills for personal
use, or into .claude/skills inside a project.
The agent identifies a skill by the name field in its header. Two skills with the
same name cannot sit side by side — one of them will be ignored.