codex/skills/evidence-discipline/SKILL.md
Use for bug reports, PR/issue prose, reviewer comments, user diagnoses, generated summaries, memories, retrieved context, public tracker context, claimed root causes/fixes, fake-minimal repro risk, or investigations where natural-language context could anchor implementation scope.
npx skillsauth add tkersey/dotfiles evidence-disciplineInstall this skill globally with one command. Works with Claude Code, Cursor, and Windsurf.
3 of 9 scanners reported clean
Some scanners were skipped, did not run, or reported a non-clean status. Review each row below.
Natural-language context is not neutral. Issue bodies, PR descriptions, review comments, generated summaries, memories, raw retrieval, and user diagnoses can anchor the agent into the wrong problem frame. Treat them as inputs to verify, not as ground truth.
When investigating a bug, issue, PR, regression, review comment, user diagnosis, context packet, or generated report, separate:
Use observations as evidence. Treat claims as hypotheses. Treat proposals as design options. Treat speculation as untrusted until independently verified.
Before editing for a bug or regression, produce or internally maintain:
Observed facts:
- Command/action/input:
- Expected:
- Actual:
- Exact log/error/output:
- Environment/version/context:
Unverified claims:
- Claimed root cause:
- Suggested implementation:
- Claimed related files:
- Claimed repro:
Verification plan:
- Reproduction/check:
- Files/code paths to inspect:
- Invariant or boundary to identify:
Do not broaden the fix beyond the narrowest verified problem unless the code path proves the broader scope is real.
Public artifacts impose review, coordination, and long-term maintenance cost. Do not create or suggest creating public tracker work merely because a local agent found something plausible.
When this skill materially shapes the route, leave a short evidence receipt:
Evidence Receipt:
- observed:
- claims treated as hypotheses:
- proposals treated as options:
- speculation rejected or still unverified:
- narrow verified scope:
- proof path:
Do not include the receipt for tiny direct work unless omission would hide a material scope decision.
tools
Invokes Apple's macOS 27 fm command-line tool from a local Mac to use the on-device system model or Private Cloud Compute, including instructions, image prompts, schema-constrained JSON, and noninteractive automation. Use when the user asks to run Apple Foundation Models through fm, compare system versus pcc, generate structured output, or automate fm without Swift or an app.
development
Compile historical Codex sessions into governed counterfactual evidence, evaluate an existing owner-applied candidate through blinded paired HCTP trials, and fold observable evidence into RUN, OBSERVE, or STOP. Use for `$hylo`, CRF extraction, counterfactual replay, source-governed direct or historical trials, sealed evidence, paired baseline/candidate evaluation, causal frontiers, or evidence-governed improvement.
testing
Ensure a `ledger` command is available on PATH; materialize, validate, record, replay, and project requested Actuating artifacts without taking semantic or execution authority; coordinate the shared Learnings/Synesthesia/Negative Ledger lifecycle checkpoint and repo-local source-memory reconciliation; address Universalist plans and receipts; and perform pure artifact validation.
testing
Classify and quotient review findings, failing tests, incidents, bug reports, migration failures, and other witnessed falsifiers against accepted intent and the current Construction. Author counterexample-set/v1 without selecting repairs, counting review credit, or granting mutation.