agentsclimarketplace

Mlsys workflow

Skill brycewang-stanford/Awesome-Journal-Skills/MLSys-Skills/skills/mlsys-workflow

Journal-specific Claude Code/Codex skill packs covering mainstream journals — AER, QJE, Nature, Cell, 管理世界, 经济研究 & 200+ more — your fast track to getting published. | 覆盖主流期刊的 Claude Code/Codex 期刊技能包,从选题、识别策略到表格规范与审稿回复全流程,助你快速发论文。

Install
npx -y skills add brycewang-stanford/Awesome-Journal-Skills --skill mlsys-workflow

Assembled from the repository path, not quoted from the project. Check it against their README if it does not work.

What its author says it does

Copied from the file, not written here

Use when planning an MLSys submission cycle end to end, from venue-fit and track choice through evaluation freeze, the October deadline, the four-day January response, notification, camera-ready, the March artifact-evaluation submission, and the May conference, with backward planning built around scarce GPU time and team ownership.

SKILL.md

6.7 KB, ~1.5k tokens by cl100k_base, as published. Nobody here has run it

MLSys Workflow

Use this as the project-management skill for a Conference on Machine Learning and Systems submission. MLSys runs one cycle per year, so a missed gate costs twelve months or a reroute to another venue. All dates below are the verified 2026-cycle timetable (access 2026-07-08) — treat them as shape, and replace every one with the current mlsys.org dates page before committing a plan.

MLSys is a conference with rotating per-edition leadership (2026: General Chair Luis Ceze; Program Chairs Zhihao Jia and Aakanksha Chowdhery) and no APC — proceedings publish open on proceedings.mlsys.org, and the money side is registration, where accepted-paper authors get a two-week reserved-ticket window.

The 2026-cycle shape

Gate2026 anchorWhat it demands
CFP and track choiceSummer 2025 (CFP publicized ~August)Research vs industrial routing decided
Paper + appendix deadlineOctober 30, 2025Frozen evaluation, blinded PDF, separate appendix upload
Reviews releasedJanuary 12, 2026All authors on deck immediately
Author response dueJanuary 16, 2026Four days — pre-planned, not improvised
NotificationJanuary 25-26, 2026Branch: camera-ready+AE, or postmortem+reroute
Artifact evaluation submissionMarch 8, 2026Badge-ready package + Artifact Appendix
AE windowMarch 8 - April 8, 2026One author on call for evaluator questions
ConferenceMay 18-22, 2026 (Bellevue)Registration inside reserved window, talk, travel

Backward plan from the paper deadline

Weeks outSystems-paper milestone
12+Bottleneck measurement done; mechanism designed; venue/track locked
10Baselines installed, version-pinned, tuning protocol agreed
8End-to-end prototype produces its first honest comparison
6Full experiment matrix launched — GPU allocation is the critical path from here
4Ablations and sensitivity sweeps done; figures generated from logs
3Complete draft in the official two-column style; internal review by one systems and one ML reader
2Evaluation freeze; only reviewer-anticipation runs remain
1Anonymity sweep, appendix assembly as separate file, bibliography full-author check
0Upload paper + appendix; verify both files from a logged-out view

The distinctive constraint versus ML-venue planning: the evaluation competes for shared hardware. Reserve cluster time for weeks 6-4 when the matrix runs, and again for the January response window, where two rebuttal experiments may need machines on two days' notice.

Standing team rituals

Weekly (weeks 12->2):
  - experiment-matrix review: rows launched / landed / failed, GPU budget burned
  - claims board: abstract claim -> supporting figure -> status (red/yellow/green)
  - baseline watch: new releases of compared systems since last week?
Once (week 3): mock review — one systems-culture and one ML-culture reader,
  scored on workload realism, baseline fairness, attribution, and tails.
Response-week protocol (pre-agreed in December):
  - who reads reviews hour 0, who owns machines, who drafts, who signs off

Failure modes by stage

  • Week 6 matrix slip: the sweep launches late, so ablations get cut — and ablation absence is precisely what MLSys reviews punish. Cut breadth (hardware variants) before cutting attribution (ablations).
  • Baseline rot: a compared system ships a major release in September; deciding whether to re-run everything at week 4 needs a pre-agreed rule ("re-run headline comparisons if the release notes claim >10% on our metric").
  • Response-window surprise: the four-day window found nobody available in January. It is on the calendar from the day of submission.
  • Post-acceptance pile-up: camera-ready, AE package, and conference logistics all land in February-March; assign three different owners at notification time.
  • Reroute paralysis after rejection: the systems-venue calendar (OSDI/SOSP/NSDI/ ASPLOS/ATC deadlines) should be listed in the postmortem doc before the decision arrives.

The reroute decision, prepared in advance

Because the cycle is annual, "revise and resubmit to next MLSys" costs a year. Write the decision rule before the notification arrives:

Rejection patternIndicated move
Workload/baseline objections, mechanism praisedStrengthen evaluation, resubmit to next MLSys — the reviews are a work list
"Insufficient novelty, solid engineering"Industrial track next cycle, or a systems venue that weighs deployment insight
"Wrong audience" signals from both culturesRe-read mlsys-topic-selection; the project may be OSDI/ASPLOS/NSDI-shaped
Split reviews, one culture convincedRewrite for the unconvinced culture; the paper, not the project, missed

Maintain the alternative-deadline list (OSDI, SOSP, NSDI, ASPLOS, ATC, EuroSys, SC) with dates in the project doc from week 12, so a January rejection flows into a concrete February plan instead of a stall.

Ownership map (minimum viable)

  • Evaluation owner: matrix, logs, figure regeneration.
  • Baseline owner: versions, tuning parity, release watch.
  • Writing owner: 10-page budget, style-kit compliance, claims board.
  • Compliance owner: anonymity, appendix upload, OpenReview fields, deadlines.
  • (Post-acceptance) AE owner: package, badges, evaluator on-call.

One person may hold two hats; no hat may be unassigned. Industrial-track submissions add a sixth hat that academic teams forget: the internal-approval owner, who clears company names, trace releases, and benchmark numbers with legal/PR — a process whose latency routinely exceeds the writing itself, so it starts at week 10, not week 1.

Cycle-volatility warning

Every date above, the track structure, the response-window length, and the AE mechanics are 2026 facts. The 2027 CFP was not yet visible at access time (待核实, all of it) — rebuild this table from mlsys.org before planning a real cycle.

Output format

[Current stage] fit / building / evaluating / writing / submitted / response / decided / AE
[Next official gate] <date + source URL>
[Critical path] <three tasks, with GPU-time dependencies flagged>
[Claims board] <claim -> figure -> status>
[Ownership] <evaluation/baseline/writing/compliance/AE -> person>
[Risk register] <matrix slip / baseline rot / response availability / logistics pile-up>

What ships with it

Read from the repository

Just SKILL.md. No reference files, no scripts.

Keep looking

Skills are one crate of 327,132. Ordering is by how many stacks a row turns up in, so the top of any crate is what has actually been picked rather than what has the most stars.