Giab benchmark validator
Babysitter enforces obedience on agentic workforces and enables them to manage extremely complex tasks and workflows through deterministic, hallucination-free self-orchestration
npx -y skills add a5c-ai/babysitter --skill giab-benchmark-validatorAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
What its author says it does
Copied from the file, not written here
Genome in a Bottle benchmark validation skill for pipeline accuracy assessment
SKILL.md
1.4 KB, as published. Nobody here has run it
GIAB Benchmark Validator Skill
Purpose
Enable Genome in a Bottle benchmark validation for pipeline accuracy assessment.
Capabilities
- Truth set comparison
- hap.py/vcfeval execution
- Sensitivity/specificity calculation
- Stratified performance metrics
- Difficult region analysis
- Validation report generation
Usage Guidelines
- Use appropriate GIAB reference samples
- Compare against truth sets with hap.py
- Calculate sensitivity and specificity
- Stratify by region type and variant class
- Analyze performance in difficult regions
- Generate comprehensive validation reports
Dependencies
- hap.py
- vcfeval
- GIAB resources
Process Integration
- Analysis Pipeline Validation (pipeline-validation)
- Whole Genome Sequencing Pipeline (wgs-analysis-pipeline)