Workflow monitor
Skill athola/claude-night-market/plugins/imbue/skills/workflow-monitor
23 Claude Code plugins: TDD enforcement hooks, git/PR workflows, spec-driven development, code review, project lifecycle, fix-from-error, maintenance automation, context optimization, research, and multi-LLM delegation. 186 skills, 128 commands, 54 agents.
npx -y skills add athola/claude-night-market --skill workflow-monitorAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
What its author says it does
Copied from the file, not written here
Detects workflow failures and inefficient patterns then files GitHub issues. Use when a workflow step repeatedly fails or produces inconsistent output.
SKILL.md
6.7 KB, ~1.5k tokens by cl100k_base, as published. Nobody here has run it
When NOT To Use
- A one-off failure worth debugging directly (use
superpowers:systematic-debugging) - Rewriting the workflow's assets (use
sanctum:workflow-improvement)
Table of Contents
- Philosophy
- Quick Start
- Detection Patterns
- Workflow
- Issue Template
- Configuration
- Guardrails
- Integration Points
- Output Format
Workflow Monitor
Monitor workflow executions for errors and inefficiencies, automatically creating issues on the detected git platform (GitHub/GitLab) for improvements. Check session context for git_platform: and use Skill(leyline:git-platform) for CLI command mapping.
Philosophy
Workflows should improve over time. When execution issues occur, capturing them systematically enables continuous improvement. This skill hooks into workflow execution to detect problems and propose fixes.
Quick Start
Manual Invocation
# After a failed workflow
/workflow-monitor --analyze-last
# Monitor a specific workflow execution
/workflow-monitor --session <session-id>
# Analyze efficiency of recent workflows
/workflow-monitor --efficiency-report
Automatic Monitoring (via hooks)
When enabled, workflow-monitor observes execution and flags:
- Command failures (exit codes > 0)
- Timeout events
- Repeated retry patterns
- Context exhaustion
- Inefficient tool usage
Detection Patterns
Error Detection
| Pattern | Signal | Severity |
|---|---|---|
| Command failure | Exit code > 0 | High |
| Timeout | Exceeded timeout limit | High |
| Retry loop | Same command >3 times | Medium |
| Context exhaustion | >90% context used | Medium |
| Tool misuse | Wrong tool for task | Low |
Efficiency Detection
| Pattern | Signal | Threshold |
|---|---|---|
| Verbose output | >1000 lines from command | 500 lines recommended |
| Redundant reads | Same file read >2 times | 2 reads max |
| Sequential vs parallel | Independent tasks run sequentially | Should parallelize |
| Over-fetching | Read entire file when snippet needed | Use offset/limit |
Workflow
Phase 1: Capture (workflow-monitor:capture-complete)
- Log execution events - Commands, outputs, timing
- Tag anomalies - Failures, timeouts, inefficiencies
- Store evidence - For reproducibility
Phase 2: Analyze (workflow-monitor:analysis-complete)
- Classify issues - Error type, severity, scope
- Identify root cause - What triggered the issue
- Suggest fix - What would prevent recurrence
Phase 3: Report (workflow-monitor:report-generated)
- Generate issue body - Structured format
- Assign labels - workflow, bug, enhancement
- Link evidence - Command outputs, session info
Phase 4: Create Issue (workflow-monitor:issue-created)
- Check for duplicates - Search existing issues
- Create if unique - Via gh CLI
- Link to session - For traceability
Issue Template
## Background
Detected during workflow execution on [DATE].
**Source:** [workflow name] session [session-id]
## Problem
[Description of the error or inefficiency]
**Evidence:**
[Command that failed or was inefficient] [Output excerpt]
## Suggested Fix
[What should change to prevent this]
## Acceptance Criteria
- [ ] [Specific fix criterion]
- [ ] Tests added for new behavior
- [ ] Documentation updated
---
*Created automatically by workflow-monitor*
Configuration
# .workflow-monitor.yaml
enabled: true
auto_create_issues: false # Require approval before creating
severity_threshold: "medium" # Only report medium+ severity
efficiency_threshold: 0.7 # Flag workflows below 70% efficiency
detection:
command_failures: true
timeouts: true
retry_loops: true
context_exhaustion: true
tool_misuse: true
efficiency:
verbose_output_limit: 500
max_file_reads: 2
parallel_detection: true
Guardrails
- No duplicate issues - Check existing issues before creating
- Approval required - Unless
auto_create_issues: true - Evidence required - Every issue must have reproducible evidence
- Rate limiting - Max 5 issues per session
Required TodoWrite Items
workflow-monitor:capture-completeworkflow-monitor:analysis-completeworkflow-monitor:report-generatedworkflow-monitor:issue-created(if issue created)
Integration Points
imbue:proof-of-work: Captures execution evidencesanctum:fix-workflow: Implements suggested fixes- Hooks: Can be triggered by session hooks for automatic monitoring
Output Format
Efficiency Report
## Workflow Efficiency Report
**Session:** [session-id]
**Duration:** 12m 34s
**Efficiency Score:** 0.72 (72%)
### Issues Detected
| Type | Count | Impact |
|------|-------|--------|
| Verbose output | 3 | Medium |
| Redundant reads | 2 | Low |
| Sequential tasks | 1 | Medium |
### Recommendations
1. Use `--quiet` flags for npm/pip commands
2. Cache file contents instead of re-reading
3. Parallelize independent file operations
### Create Issues?
- [ ] Issue 1: Verbose output from npm install
- [ ] Issue 2: Redundant file reads in validation
Related Skills
imbue:proof-of-work: Evidence capture methodologysanctum:fix-workflow: Workflow improvement command
Status: Skeleton implementation. Requires:
- Hook integration for automatic monitoring
- Efficiency scoring algorithm
- Duplicate detection logic
Exit Criteria
- All 4 TodoWrite phases completed:
capture-complete,analysis-complete,report-generated, and (if an issue is created)issue-created - Every issue filed contains reproducible evidence: the exact command that failed or was inefficient plus an output excerpt
- Duplicate check run via
gh issue list --searchbefore creating any issue; duplicate suppressed and existing issue URL reported instead - No more than 5 issues created per session regardless of how many anomalies are detected; rate limit enforced
What ships with it: 3 files
14.2 KB alongside SKILL.md
modules/
- detection-patterns.md3.9 KB
- efficiency-metrics.md6.0 KB
- issue-templates.md4.3 KB
Gives 0 of the 12 instructions most automation workflows skills give in ~1.5k tokens
Counted across 745 of the 1,008 authors here whose files we hold, read 2026-08-07
- Write conventional commit messagesin 36 of 745, across 35 files
- Delete branches after mergein 30 of 745, across 21 files
- Make atomic commitsin 25 of 745, across 15 files
- Write minimal code to pass testsin 22 of 745, across 10 files
- Re-snapshot after navigation or DOM changesin 21 of 745, across 13 files
- Use try-catch for error handlingin 20 of 745, across 8 files
- Run tests before committingin 20 of 745, across 12 files
- Write tests before implementationin 20 of 745, across 8 files
- Configure branch protection rulesin 19 of 745, across 5 files
- Explain the why in commit messagesin 19 of 745, across 9 files
- Refactor code while tests remain greenin 19 of 745, across 6 files
- Interact with elements using refsin 19 of 745, across 11 files
Said here and by no other author read
- create issue via CLI if unique
- rate limit to five issues per session
- include reproducible evidence in every issue
- log all execution events
- classify anomalies by error type and severity
- identify root cause for each anomaly
Grouped from the skills themselves: near-identical wordings counted once, and counted by distinct author, so one author publishing three of these counts once. Length counted with cl100k_base; the agent that loads this file may tokenize it differently.