agentsclimarketplace

Grafana

Skill addxai/enterprise-harness-engineering/skills/grafana

Enterprise-grade AI Agent Skills for software development, DevOps, SRE, security, and product teams. Compatible with Claude Code, Cursor, Windsurf, Gemini CLI, GitHub Copilot, and 30+ AI coding agents.

Install
npx -y skills add addxai/enterprise-harness-engineering --skill grafana

Assembled from the repository path, not quoted from the project. Check it against their README if it does not work.

What its author says it does

Copied from the file, not written here

Query and manage Grafana dashboards, alert rules, and data sources via HTTP API. Use when viewing dashboards, troubleshooting alerts, checking service metrics, finding data sources, or when Grafana, monitoring, alerts, dashboards, or observability is mentioned.

SKILL.md

3.6 KB, as published. Nobody here has run it

Grafana

Query and manage Grafana monitoring dashboards, alert rules, and data sources via HTTP API.

Setup

Configure your Grafana instance:

VariableDescriptionRequired
GRAFANA_URLYour Grafana server URL (e.g., https://grafana.example.com)Yes
GRAFANA_TOKENAPI Key (Settings → API Keys, viewer or editor role)Yes

Authentication: curl -s -H "Authorization: Bearer $GRAFANA_TOKEN" "$GRAFANA_URL/api/..."

Customize the following for your organization:

  • Folder structure: How dashboards are organized (by team, region, service, etc.)
  • Tag conventions: What tags are used for filtering (region, environment, service type)
  • Alert naming: Your alert naming pattern (e.g., {Service} {Metric} {Condition} [{env}])
  • Key dashboards: Which dashboards to check after deployments

Rules

Dashboard Operations

  • Search first, then detail: Use search API with tags/query to narrow scope, then fetch by uid
  • Never delete production dashboards — archive by moving to an Archive folder
  • Write operations require user confirmation before execution (create/modify dashboards, alert rules)

Common Workflows

Find a Dashboard:

curl -s -H "Authorization: Bearer $GRAFANA_TOKEN" \
  "$GRAFANA_URL/api/search?query=<keyword>&tag=<tag>&type=dash-db" \
  | jq '.[] | {uid, title, folderTitle}'

Get Dashboard Details:

curl -s -H "Authorization: Bearer $GRAFANA_TOKEN" \
  "$GRAFANA_URL/api/dashboards/uid/<uid>" \
  | jq '.dashboard.panels[] | {title, type}'

Check Active Alerts:

curl -s -H "Authorization: Bearer $GRAFANA_TOKEN" \
  "$GRAFANA_URL/api/alerts?state=alerting"

List Data Sources:

curl -s -H "Authorization: Bearer $GRAFANA_TOKEN" \
  "$GRAFANA_URL/api/datasources" \
  | jq '.[] | {id, name, type, url}'

Search by Folder:

curl -s -H "Authorization: Bearer $GRAFANA_TOKEN" \
  "$GRAFANA_URL/api/search?folderIds=<folder_id>&type=dash-db"

Post-Deployment Checks

After each deployment, check key dashboards for:

  • Error rate changes (before vs after deployment)
  • Latency P95/P99 trends
  • Consumer lag (if using message queues)
  • Connection counts (TCP/WebSocket)
  • Log level distribution (error/warn spikes)

Alert Troubleshooting

  1. GET /api/alerts?state=alerting — list all firing alerts
  2. Identify the dashboard and panel from the alert
  3. Read the PromQL expression from the panel
  4. Query Prometheus directly to understand the data
  5. Check recent deployments or config changes as potential cause

Examples

Bad

# Delete a production dashboard (never delete, archive instead)
curl -X DELETE "$GRAFANA_URL/api/dashboards/uid/abc123"

# Fetch all dashboards without filtering (use search first)
for uid in $(curl ... /api/search | jq -r '.[].uid'); do
  curl ... /api/dashboards/uid/$uid
done

Good

# Search dashboards by tag
curl -s -H "Authorization: Bearer $GRAFANA_TOKEN" \
  "$GRAFANA_URL/api/search?tag=production&type=dash-db" \
  | jq '.[] | {uid, title, folderTitle}'

# Check active alerts
curl -s -H "Authorization: Bearer $GRAFANA_TOKEN" \
  "$GRAFANA_URL/api/alerts?state=alerting"

# Get alert details for troubleshooting
curl -s -H "Authorization: Bearer $GRAFANA_TOKEN" \
  "$GRAFANA_URL/api/alerts/<alert_id>"

Keep looking

Skills are one crate of 328,083. Ordering is by how many stacks a row turns up in, so the top of any crate is what has actually been picked rather than what has the most stars.