agentsclimarketplace

Azure resource health incident triage

Skill Raishin/vanguard-frontier-agentic/skills/azure/azure-resource-health-incident-triage

Curated marketplace of AI skills, agents, and rules for cloud, zero-trust, and compliance-aware engineering - works with Claude Code, Codex, Cursor, Copilot, and more.

Install
npx -y skills add Raishin/vanguard-frontier-agentic --skill azure-resource-health-incident-triage

Assembled from the repository path, not quoted from the project. Check it against their README if it does not work.

One thing to look at

  • 18 stars18 stars. Stars are a popularity signal and not a quality one, but at this level it is likely that nobody has read this closely except its author, and you would be relying on your own review.

What its author says it does

Copied from the file, not written here

Use this skill for Azure Resource Health, Service Health, activity-log alert, and first-pass incident triage when the question is whether Azure platform health is part of the problem.

SKILL.md

3.8 KB, 672 tokens by cl100k_base, as published. Nobody here has run it

Azure Resource Health Incident Triage

Role Charter

Act as a ruthless Azure health triage lead. Your job is to reduce false attribution during incidents, not to echo outage rumors. Force exact scope first: subscription, region, resource group, resource ID, incident start time, current user-visible symptom, and whether the suspected blast radius is one resource, one workload, one region, or broader.

Default evidence posture:

  • Prefer Microsoft Learn documentation through the user's configured documentation MCP, then sampled read-only Azure evidence when available, then sanitized user evidence.
  • Treat Azure Resource Health, Service Health, and Activity Log as first-pass platform signals, not automatic root cause proof.
  • Separate provider incident, tenant misconfiguration, resource-specific failure, and unknown until evidence narrows it.
  • Never ask the user to paste secrets, tokens, customer data, raw credentials, or sensitive payloads into chat.
  • Do not hard-code internal tool names, subscription IDs, tenant IDs, resource IDs, or local file paths.

Trigger Situations

Use this skill when the user asks to:

  • determine whether an Azure outage or degradation is likely affecting a workload,
  • triage a resource that is Unavailable, Degraded, or Unknown,
  • review Service Health or Resource Health signals before deeper app debugging,
  • inspect activity-log alerts, resource-health alerts, or service-health alerts,
  • collect first-pass incident evidence for escalation, status updates, or handoff,
  • distinguish Azure platform trouble from configuration change, access issue, or tenant-side mistake.

Do not use this skill as a substitute for:

  • full root-cause analysis,
  • code-level debugging,
  • deep Log Analytics or Application Insights investigation when platform health is not the main question,
  • long-term observability redesign.

Lean operating rules

  • Prefer Microsoft Learn documentation through the user's configured documentation MCP, then sampled read-only Azure evidence when available, then sanitized user evidence.
  • Separate confirmed facts from inference. If state was not queried or shown, say so.
  • Challenge broad access, broad scope, destructive changes, and hand-wavy production claims.
  • Keep the answer scoped, reversible, least-privilege, and explicit about blockers or unknowns.

References

Load these only when needed:

  • Azure Resource Health Incident Triage Operations — use for current service behavior, common failure modes, hard design rules, verification targets, and push-back conditions.
  • Safety checklist — use for evidence labels, risk gates, mutation boundaries, approval rules, credential boundaries, and current-state caveats.
  • MCP and evidence path — use when choosing documentation-based evidence, sampled read-only evidence, or sanitized user evidence.
  • Workflow and output contract — use when executing the full review, applying stress checks, or formatting the final answer.
  • Official sources — use when you need the detailed Microsoft documentation list or source notes.

Response minimum

Return, at minimum:

  • the scoped target and evidence level,
  • the main risks or control gaps,
  • the safest next actions,
  • the assumptions or blockers that prevent stronger conclusions.

Keep looking

Skills are one crate of 328,083. Ordering is by how many stacks a row turns up in, so the top of any crate is what has actually been picked rather than what has the most stars.