agentsclimarketplace

Gpt55 critical auditor

Skill richfrem/agent-plugins-skills/plugins/agent-agentic-os/skills/gpt55-critical-auditor

repo for reusable plugins and skills

Install
npx -y skills add richfrem/agent-plugins-skills --skill gpt55-critical-auditor

Assembled from the repository path, not quoted from the project. Check it against their README if it does not work.

One thing to look at

  • 4 stars4 stars. Stars are a popularity signal and not a quality one, but at this level it is likely that nobody has read this closely except its author, and you would be relying on your own review.

What its author says it does

Copied from the file, not written here

Conducts a full-system adversarial audit of agent plugins, skills, and orchestration against enforced runtime contracts using deep reasoning (GPT5.5-equivalent thinking).

SKILL.md

2.1 KB, as published. Nobody here has run it

Purpose

This skill performs a deep, adversarial system audit designed to break:

  • runtime enforcement guarantees
  • mutation safety assumptions
  • eval coverage completeness
  • plugin isolation boundaries

This is NOT a compliance review.

This is a failure-seeking audit.


Audit Mandate

You must assume:

  • The system is incorrect until proven otherwise
  • Enforcement can be bypassed unless proven impossible
  • Evals are incomplete
  • Plugins will attempt to violate boundaries
  • Mutation logic will corrupt state unless isolated

Required Focus Areas

1. Execution Enforcement

  • Can HALT be bypassed?
  • Are there code paths where violations do not interrupt execution?
  • Are exceptions truly global?

2. Mutation Integrity

  • Can partial mutations persist?
  • Are sandbox boundaries leak-proof?
  • Can state escape temp/sandbox/?

3. Eval Authority

  • What behaviors are NOT covered by evals?
  • Can mutations exploit blind spots?
  • Can regression be reclassified as acceptable?

4. Baseline Integrity

  • Can baselines be indirectly manipulated?
  • Are diffs actually verified or just declared?

5. Plugin Isolation

  • Do any skills:
    • chain multiple actions?
    • make decisions?
    • orchestrate workflows?

6. Orchestration Leakage

  • Are there hidden multi-step flows inside skills?
  • Are sub-agents behaving like orchestrators?

7. Debt System Exploits

  • Can debt accumulate without blocking progress?
  • Can agents route around debt enforcement?

Output Requirements

For every issue:

  • ID
  • Severity (P0/P1/P2)
  • Exploit scenario (MANDATORY)
  • Exact failure path
  • Why enforcement fails
  • System impact
  • Recommended fix
  • Can agent exploit automatically? (YES/NO)

Critical Rule

If you cannot prove a guarantee is enforced:

→ It must be treated as NOT enforced

Keep looking

Skills are one crate of 328,083. Ordering is by how many stacks a row turns up in, so the top of any crate is what has actually been picked rather than what has the most stars.