The persona layer your agent already speaks.
One SKILL.md. Filtered per intensity by get_lean_instructions(mode) and injected in the system slot by per-provider adapters with Anthropic prompt-cache markers. Cuts LLM output size, cost, and latency · measured, not implied.
skills/lean/SKILL.md · read once, shipped everywhere
The plugin, MCP server, benchmark arms, and FastAPI adapters all read the same file. Bump it once · everything downstream picks it up. The system slot is marked cache_control: ephemeral, so the persona charges once per Anthropic prompt-cache TTL, not per turn.
Three payloads, one file.
Minimum payload · small models, tight context, cost-sensitive calls.
Default · production balance of guidance and payload.
Maximum guidance · long agentic sessions with over-build risk.
Ships across every major agent host.
Eight slash commands, one persona.
Optional statusline: point statusLine.command at hooks/lean-statusline.sh to show the active level.