Academic experiments
Skill joshua-zyy/academic-paper-writer/skills/academic-experiments
Audit, run, or verify experimental evidence for CS/AI/ML papers. Produces Evidence Inventory with evidence_type annotations (newly_run/preexisting_artifact/user_claim) and Protocol Risk assessments. Use when: checking if experiment results are reproducible, auditing existing experiment artifacts, running minimal reproducible commands, evaluating checkpoints without full retraining, documenting protocol risks like data leakage or missing baselines. Triggers on: 复核实验, run experiments, 实验结果, experiment evidence, verify results, 实验验证, evidence inventory, protocol risk, 跑实验, check results, reproduce experiments, 实验审计.From its SKILL.md
npx -y skills add joshua-zyy/academic-paper-writer --skill academic-experimentsAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
One thing to look at
- runs commandsInstructs the agent to run 2 commands, including `Read manifest.yaml` and 1 more.
SKILL.md
3.1 KB, 584 tokens by cl100k_base, as published. Nobody here has run it
Academic Experiments
将此 skill 视为"实验取证代理",目标是建立最短且可信的证据链,而不是尽量多跑实验。
Router Protocol
- Read
manifest.yaml. It declaresalways_loadfiles,axes, andreferences.on_demand. - Read every file listed under
always_load. These are the skill's binding rules — not reference material. - Apply the loaded material as constraints:
stance.mddefines non-negotiable rules, evidence type semantics, failure degradation, and scope.red-lines.mddefines absolute prohibitions. Do not negotiate these.output-contract.mddefines deliverables per mode and claim-readiness classification.anti-patterns.mddefines known failure modes and their correct alternatives.
- Detect the mode using the manifest's
modeaxis:experiment-evidence-pass,evidence-inventory-only, orminimal-reproducible-run. Align evidence type semantics to../shared/core/evidence-policy.md. - Echo the selected mode to the user before executing.
- Reach for
references/only when the manifest'sreferences.on_demandcondition is satisfied.
Modes
| Mode | Use when |
|---|---|
experiment-evidence-pass | Full audit: inventory + run + record + risk analysis |
evidence-inventory-only | Inventory existing artifacts only, no execution |
minimal-reproducible-run | Execute minimal reproducible command (e.g. eval existing checkpoint) |
Agent Dispatch
agents/experiment_agent.md is dispatched by academic-paper-writer orchestrator at Step 4. The agent may run experiments but must not modify project source code or data files, nor write paper prose independently.
Independent Use
| Input | Mode | Priority | Behavior |
|---|---|---|---|
repo_path + no run mode | experiment-evidence-pass | 2 (path trigger) | Full audit: inventory → env → minimal run → risk |
repo_path + "inspect only" | evidence-inventory-only | 1 (explicit) | Inventory only, no commands |
repo_path + specific command | minimal-reproducible-run | 1 (explicit) | Verify env → execute → record |
No repo_path | — | 3 (no input) | Ask path, or auto-detect entry files |
| Scenario | Recommended |
|---|---|
| Just auditing/reproducing evidence | This skill (standalone) |
| Writing results into paper prose | academic-paper-writer orchestrator |
| Draft results need verification | This skill → academic-reviser |
What ships with it: 10 files
20.6 KB alongside SKILL.md, 1 of them executable
agents/
- experiment_agent.md5.9 KB
references/
- evidence-inventory.md1.7 KB
- protocol-risks.md1.8 KB
- run-strategy.md1.4 KB
scripts/
- evidence_scanner.pyruns4.2 KB
static/
- core/anti-patterns.md571 B
- core/output-contract.md906 B
- core/red-lines.md477 B
- core/stance.md2.5 KB
- manifest.yaml1.2 KB