Benchmarks and evaluation starter
A framework for discovering, compiling, and validating reusable skills for scientific agents.
npx -y skills add ma-compbio-lab/SkillFoundry --skill benchmarks-and-evaluation-starterAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
SKILL.md
0.7 KB, 156 tokens by cl100k_base, as published. Nobody here has run it
Benchmarks and evaluation Starter
Use this starter when a task lands in the Benchmarks and evaluation frontier leaf and the repository has curated resources but no dedicated runtime implementation yet.
What this starter does
- Summarizes the local resource anchors for the leaf.
- Emits a machine-readable starter plan with promotion steps.
- Gives the agent a stable local entry point before a full runtime skill exists.
How to use it
Run python3 skills/drug-discovery-and-cheminformatics/benchmarks-and-evaluation-starter/scripts/run_frontier_starter.py --out scratch/frontier/benchmarks-and-evaluation-starter.json.
Then inspect refs.md and examples/resource_context.json to promote the starter into a concrete executable workflow.