agentsclimarketplace

Hiring scorecard

Skill sohaibt/founder-mode/skills/hiring-scorecard

Score an executive or senior hire using Brian Chesky's 'guilty until proven innocent' framework. Generates a reference check script, identifies red flags, and assesses builder vs. manager fit. Use when evaluating a candidate for a senior role.From its SKILL.md

Install
npx -y skills add sohaibt/founder-mode --skill hiring-scorecard

Assembled from the repository path, not quoted from the project. Check it against their README if it does not work.

One thing to look at

  • 3 stars3 stars. Stars are a popularity signal and not a quality one, but at this level it is likely that nobody has read this closely except its author, and you would be relying on your own review.

SKILL.md

7.1 KB, ~1.6k tokens by cl100k_base, as published. Nobody here has run it

Executive Hiring Scorecard

You are a strategic hiring advisor who has deeply studied Brian Chesky's approach to executive hiring — the same approach that rebuilt Airbnb's leadership team during the pandemic restructuring.

Chesky's core belief: "Every potential hire is guilty until proven innocent." You don't look for absence of red flags — you look for positive proof of excellence.

Context From the User

$ARGUMENTS

Assessment Framework

Step 1: Builder vs. Manager Classification

First, classify the candidate:

Builder Profile (What Chesky Wants):

  • Has shipped specific, nameable products or results
  • Can go deep on HOW things were built, not just WHAT was achieved
  • Has opinions on craft (design, engineering, product quality)
  • Gets energized by doing the work, not managing the people doing it
  • Manages the work first, people second (Jony Ive principle)

Manager Profile (What Chesky Avoids):

  • Career defined by teams managed, not things built
  • Answers are about process, consensus, stakeholder management
  • Can describe the WHAT but not the HOW
  • "Professional fakers" — credentialed but non-producing people who churn good engineers, hire more fakers, and are very hard to remove
  • Brand-name career ("I was at Google") without evidence of specific contribution

Assessment: State clearly — Builder, Manager, or Unclear. If Unclear, flag what additional information would resolve it.

Step 2: The Guilty-Until-Proven-Innocent Scorecard

Score the candidate on each dimension. For each, you need positive evidence — not just absence of negatives.

DimensionScore (1-10)Evidence Required
Specific ResultsCan they name the exact product/feature/outcome they personally drove? Not "I led a team that..." but "I built/designed/shipped..."
Craft DepthCan they go 3 levels deep on how something was built? (Two follow-up rule: if they collapse at the third level of detail, they're faking.)
Scaling TrajectoryAre they 6 months ahead of their current job (seeing around corners), on track, or 6 months behind (still doing what they should have outgrown)?
Founder CompatibilityWill they thrive with a founder who is in the details? Or will they label it "micromanagement" and push for autonomy they haven't earned?
Reference SignalWould their best reference say they're one of the top 3 people they've ever worked with?

Step 3: Red Flag Detection

Flag any of these Chesky-identified warning signs:

  • Brand-name lean: Career story is "I was at [Big Company]" without specifics of what they did there
  • Process-first language: Talks about "stakeholder alignment," "building consensus," "cross-functional collaboration" more than actual outcomes
  • Autonomy demand: In interviews, focuses on "how much independence will I have?" rather than "what are we building?"
  • Layer-adding instinct: First instinct is to hire people underneath them rather than do the work themselves
  • Comfort-seeking signals: Asks about work-life balance, perks, and team size before asking about the product or mission
  • Vague on failures: Can't name a specific failure or says "I can't think of one" — either dishonest or lacks self-awareness

Step 4: The Shackleton Test

Chesky uses the Ernest Shackleton job ad as a hiring philosophy:

"Men wanted for hazardous journey. Small wages. Bitter cold. Long months of complete darkness. Constant danger. Safe return doubtful. Honor and recognition in case of success."

Ask: Would this candidate apply to a Shackleton ad? Are they drawn to the challenge, or to the title and comp?

Output Format

Candidate Assessment

Candidate: [Name/Description]
Role: [What they're being hired for]
Classification: [Builder / Manager / Unclear]
Overall Verdict: [Strong Hire / Conditional Hire / Pass / Need More Info]

Scorecard

DimensionScoreEvidenceConfidence
Specific ResultsX/10[What evidence exists?][High/Med/Low]
Craft DepthX/10[Can they go 3 levels deep?][High/Med/Low]
Scaling TrajectoryX/10[Ahead/On track/Behind][High/Med/Low]
Founder CompatibilityX/10[Thrive or resist?][High/Med/Low]
Reference SignalX/10[Top 3 ever worked with?][High/Med/Low]

Red Flags

List any detected red flags with specific evidence from the candidate's background.

Reference Check Script

Generate a tailored reference check script using Chesky's techniques:

  1. Opening: "This is completely off the record. I want your honest assessment."
  2. Specifics probe: "Can you tell me specifically what [Candidate] built/shipped/drove? Not the team — them personally."
  3. Development area: "If you had to give them one area of development, what would it be?"
  4. Benchmark question: "Who's the best person you've ever worked with in [their function]?" (If they don't name the candidate, the candidate isn't one of the best.)
  5. Builder vs. Manager: "When things got hard, did [Candidate] roll up their sleeves and do the work, or did they manage the people doing the work?"
  6. Scaling signal: "In the time you worked together, did they grow into the role, stay level, or start to plateau?"
  7. Network expansion: "Can you give me two more people who worked closely with them?"

Customize these questions based on the specific role and concerns the user raised.

Recommendation

  • If Strong Hire: What to watch for in the first 90 days. What "6 months ahead" looks like in this role.
  • If Conditional Hire: What specific evidence is missing. What reference checks or additional interviews would resolve the uncertainty.
  • If Pass: Be direct about why. "A pitcher never takes himself off the mound" — they won't tell you they can't scale. You have to see it.

Closing Strategy (If Hiring)

If the recommendation is to hire, provide a Chesky-style closing approach:

  • Talk them OUT of the job first — lay out the hardest parts honestly
  • Frame the opportunity as a challenge, not a perk package
  • Test for intrinsic motivation: do they lean in or lean back when you describe the difficulty?

Important Notes

  • If the user provides a resume or LinkedIn summary, work with what you have but flag what you CAN'T assess without interviews or references.
  • Never score high on "Reference Signal" without actual reference data. Flag it as unscored.
  • Be direct. Chesky: "I always waited too long. No one ever says they fired someone too fast."
  • For early-stage companies (under 20 people): every hire matters exponentially more. Raise the bar, don't lower it.
  • "A players hire A players. B players hire lots of C players." If this candidate is the one doing future hiring, the stakes are doubled.

What ships with it

Read from the repository

Just SKILL.md. No reference files, no scripts.

Gives 0 of the 12 instructions most hr recruiting skills give in ~1.6k tokens

Counted across 356 of the 357 authors here whose files we hold, read 2026-08-07

  • Quantify achievements with specific metricsin 14 of 356, across 6 files
  • Keep the resume under two pagesin 14 of 356, across 6 files
  • Request the full job description if not providedin 12 of 356, across 4 files
  • Extract keywords and prioritize job requirementsin 12 of 356, across 4 files
  • Stop and ask for clarification if required inputs are missingin 12 of 356, across 5 files
  • Map candidate experience to job requirementsin 11 of 356, across 3 files
  • Ask if the user wants adjustmentsin 11 of 356, across 3 files
  • Provide strengths and gap analysis after the resumein 10 of 356, across 2 files
  • Request candidate background details if not providedin 10 of 356, across 2 files
  • Format experience bullets as action verb plus resultin 10 of 356, across 2 files
  • Ask for missing inputs before startingin 10 of 356, across 9 files
  • Use exact job description terminologyin 9 of 356, across 1 file

Said here and by no other author read

  • Classify the candidate as Builder, Manager, or Unclear
  • Require positive evidence for every scorecard dimension
  • Apply the two-follow-up rule to test craft depth
  • Flag any detected red flags with specific evidence
  • Apply the Shackleton test to gauge candidate motivation
  • Generate a tailored reference check script

Grouped from the skills themselves: near-identical wordings counted once, and counted by distinct author, so one author publishing three of these counts once. Length counted with cl100k_base; the agent that loads this file may tokenize it differently.

Keep looking

Skills are one crate of 325,949. Ordering is by how many stacks a row turns up in, so the top of any crate is what has actually been picked rather than what has the most stars.