Skill decay
Agent Guards for long-lived AI agents: context-budget (audit per-turn token weight, fail CI on bloat) + claim-audit (flag unverified factual claims). Standalone tools + CI-ready, Claude Code / Codex / OpenCode compatible.
npx -y skills add trac3r00/agent-guards --skill skill-decayAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
2 things to look at
- 29 days oldThe repository was created 29 days ago. New is not bad, but a brand new repository carrying a familiar-sounding name is the shape a typosquat arrives in, and there has been no time for anyone else to find a problem with it.
- 0 stars0 stars. Stars are a popularity signal and not a quality one, but at this level it is likely that nobody has read this closely except its author, and you would be relying on your own review.
What its author says it does
Copied from the file, not written here
Find declared-but-unused skills, tools, or plugins by cross-referencing an inventory (SKILL.md files or a name list) against real usage logs, and fail CI when the dead weight grows past a budget.
The file declares its own license as MIT. That is the author’s claim about this one file, and it is not the same thing as the license GitHub reports for the repository, which is listed with the other numbers below.
SKILL.md
5.9 KB, as published. Nobody here has run it
Skill Decay
Turn "we have 200 skills" into "we use 40 of them" — with receipts.
Overview
Every skill, tool, or plugin an agent declares is loaded, described, and paid for in the prompt on every single turn. The inventory only ever grows: adding one looks free, removing one requires proof it's dead, and nobody gathers that proof. So the context window fills with capabilities that haven't been invoked in months, quietly taxing latency and token cost on requests that will never touch them.
skill_decay.py is that proof. It cross-references a declared inventory
against a usage log and reports the decay:
- never — declared, loaded, invoked zero times in the logs.
- stale — used once, but not in the last N days (default 30).
- live — actually earning its prompt slot.
It exits non-zero when the number of decay candidates blows a budget, so "the
inventory grew faster than it's used" fails CI instead of silently bloating
every request. Nothing from the inventory is imported or executed — SKILL.md
is read as text, logs are read as text. Pure stdlib.
This is the other half of a context budget: context-budget tells you how
heavy each file is, skill-decay tells you whether anyone uses it. A file
that is both heavy and never-used is the first thing to cut.
When to use
- An agent/plugin host declares dozens+ of skills/tools and per-turn prompt cost is climbing, but you can't tell which ones are dead.
- You want CI to fail when the count of never/stale capabilities exceeds a budget, forcing a prune-or-justify decision.
- You're deciding what to archive and want a usage table, not a vibe.
Not for: proving a skill is good (a rarely-used skill can still be critical —
read never as a review prompt, confirm it isn't a break-glass tool before
deleting), or measuring token weight (that's context-budget).
The method
- Point it at the inventory and the usage.
Inventory names come from each# inventory = a tree of SKILL.md files; usage = your agent logs python scripts/skill_decay.py --skills-dir ~/.hermes/skills \ --logs ~/.hermes/logs --stale-days 30 --max-decay 20SKILL.mdfrontmattername:(or the containing directory name). If you don't haveSKILL.mdfiles, pass a flat list instead:--names plan,codex,dogfood,spike. - Read the three tiers.
- never: declared, never invoked. The prime prune candidates.
- stale: used once, but not within
--stale-days. Decaying. - live: actively used — leave them alone. Each row carries the invocation count and days-since-last-use so you can rank the cut list by cost-of-keeping.
- Wire it into CI. The tool exits non-zero over the decay budget:
Or be strict with# fail the build when >20 skills are dead or decaying - run: python scripts/skill_decay.py --skills-dir skills --logs logs --max-decay 20--fail-on-neverto block any never-used skill from merging. Now inventory bloat is a red build, not a slow creep. - Prune the top, re-measure. Archive a confirmed-dead skill, re-run. The
decay_candidatescount is the scoreboard.
How usage is matched
- Each inventory name is matched as a whole token (word/path-segment
boundaries), so
plannever counts a hit insideplanetorplanner. Hyphens and underscores are treated as word-internal, soskill-decaymatches exactly and notskill. - Any
YYYY-MM-DDon a usage line is read as that invocation's timestamp; the latest one becomes the item's last-seen date, which drives staleness. Logs with no dates still produce never/live verdicts — you just lose the stale tier (everything used islive). - Usage sources: one or more
--logsfiles/dirs (repeatable), or stdin. A directory is walked recursively.
Anti-patterns
- Deleting a
neverskill on sight. Zero invocations is a strong signal, not a verdict — a break-glass / disaster-recovery skill should be rarely used. Confirm it isn't intentionally dormant before archiving. - Auditing against too short a log window. If your usage log only covers a
week, a monthly skill looks dead. Point
--logsat a window at least as long as your longest legitimate usage cadence. - Matching substrings. Naive
grep skillnameover-counts (planinsideplanet) and under-cuts real dead weight. This tool matches whole tokens on purpose — don't "fix" it back to substring matching. - Treating the tool as its own exception. A decay auditor that itself never runs is dead weight. Put it in CI so it runs on every change.
Example
$ python scripts/skill_decay.py --skills-dir skills --logs agent.log --stale-days 30 --max-decay 5
Skill-decay report
==============================
inventory: 42 as-of: 2026-07-07 stale>30d
live: 30 stale: 5 never: 7 decay-candidates: 12
name calls last-use decay
------------------------------------------------
touchdesigner-mcp 0 - never
godmode 0 - never
...
p5js 1 2026-04-02 stale (96d)
plan 171 2026-07-07 live (0d)
FAIL: decay_candidates 12 > max_decay 5
$ echo $?
1