skills/office-hours/SKILL.md
YC-partner-style interrogation of a raw idea. Six forcing questions before you shape anything: demand reality, status quo, desperate specificity, narrowest wedge, observation & surprise, future-fit. Operationalizes the AGENTS.md "Diverge Before You Converge" doctrine at the ideation stage — problem-diamond, pre-/shape. Use when: user arrives with a rough idea, backlog item is fuzzy, you can't already write a one-sentence goal with a testable outcome, or the phrase "what should we build" would be useful. Trigger: /office-hours, /oh, /interrogate.
npx skillsauth add phrazzld/agent-skills office-hoursInstall this skill globally with one command. Works with Claude Code, Cursor, and Windsurf.
3 of 9 scanners reported clean
Some scanners were skipped, did not run, or reported a non-clean status. Review each row below.
Raw idea → sharpened problem statement. Six forcing questions that catch
premature scoping, proxy goals, and solution-shaped problems before they
reach /shape.
/shape on an unshaped conceptSkip for small, well-understood work. Office hours is for moments where the idea is not yet an idea.
Ask each, in order. Record the answers.
Who specifically wants this, and how do you know? Not "users" — which users. What did they actually say? If you can't name three specific people or instances, the demand is hypothetical.
What do they do today? If the answer is "they live with it," check whether they actually feel pain. If the answer is "they use tool X," your bar is "10× better than X," not "exists."
Describe the worst outcome of not building this. Concrete: what breaks, who suffers, what do they do next? If the answer is "they keep going, fine," the need isn't desperate — scope accordingly.
What's the smallest, ugliest version that would make ONE named user happy? If you can't name it, you're designing for a crowd you haven't met.
What would we learn by shipping this that we don't already know? If the outcome is predictable, the value is bounded — ship the prediction cheaply if at all.
Three years out, does this survive? If the use case evaporates when (a) LLMs improve, (b) underlying tools change, (c) the team reorgs, it's a tactical patch — name it as such.
## Office Hours: <idea>
### Demand Reality
[who, when, what they said — specific]
### Status Quo
[what they do today; bar to clear]
### Desperate Specificity
[concrete failure mode of not building]
### Narrowest Wedge
[smallest ship that makes one named user happy]
### Observation & Surprise
[what we'd learn; confidence level]
### Future-Fit
[three-year survival: yes / no / tactical-patch]
### Sharpened Problem Statement
<one sentence ready for /shape, OR "needs more demand reality before shape">
/shape already ran,
use /ceo-review — office hours is the wrong tool for an already-shaped
plan./groom surfaces raw themes from a codebase. Pipe the most promising
theme into /office-hours before shaping./shape accepts a sharpened problem statement as input. If office
hours surfaces absent demand reality, do not shape — the work isn't ready./ceo-review is the post-shape counterpart. Office hours sharpens
the problem; CEO review challenges the plan.testing
Capture one compounding repo-technical learning while a solved problem is still fresh. Use when: after a bug fix, diagnosis, delivery, review, or incident reveals a reusable pattern worth adding to `docs/solutions/`. Trigger: /compound, /capture-learning, /learning.
testing
Route Misty Step factory application capabilities. Use when choosing, auditing, integrating, or operating Canary, Powder, Landmark, Aesthetic, or Bitterblossom: production observability, incidents, health checks, error logging, backlog/work-card state, release intelligence, UI/UX system adoption, or supervised/unsupervised agent dispatch. Trigger: /factory-apps, /factory-stack.
testing
Prove a skill beats no-skill with a falsifiable A/B eval, or retire it. Design, generate, run, and maintain a skill-specific eval: name the one claim the skill must earn, run it skill-on vs raw same-model, grade blind with objective checks first, return a keep/adapt/cut verdict. Use when: "eval this skill", "does this skill help", "prove the skill beats no skill", "write an eval for", "benchmark a skill", "is this skill worth it", "skill A/B", "skill regression test", "generate skill evals". Trigger: /skill-eval, /eval-skill, /prove-skill.
tools
> Template. Copy to `<target-repo>/.agents/skills/<repo>-<domain>/SKILL.md` > and fill every bracketed placeholder from the live target repo. Delete this > line and every other `> ` guidance line before committing. See > `../../references/repo-local-skill-generation.md` for the full process. --- name: <repo>-<domain> description: | [One paragraph: what this skill verifies/runs/operates for <repo>, stated in terms of the repo's real shape (service/CLI/library/etc.), not generic process. En