← ClaudeAtlas

health-score-designerlisted

When the user wants to design, weight, calibrate, backtest, audit, or fix a customer health score. Also use when the user mentions 'health score is garbage', 'nobody trusts it', 'score that actually predicts', 'design a health score', 'build a customer health score', 'our health score is broken', 'nobody trusts the health score', 'what should go into our health score', 'health score weights', 'green accounts keep churning', 'is our health score any good', 'backtest the health score', 'calibrate our health score', 'red yellow green thresholds', 'the score isn't predicting anything', 'audit our scoring model', or 'rescore the book'. Use this whenever someone is deciding what makes an account healthy, or arguing about a score's inputs, even if they never say 'health score' — a request for 'one number that tells me who to work on' is this skill. For scoring one account today, see churn-risk. For the data inventory the score is built on, see cs-context. For per-account renewal probability, see renewal-forecast.
gaintrace/customer-success-skills · ★ 1 · AI & Automation · score 75
Install: claude install-skill gaintrace/customer-success-skills
# Health Score Designer You own the health score as a **measurement instrument**, not a dashboard widget. The standard is a score a CFO would accept as an input to the renewal forecast: it predicts a named commercial event over a named horizon, and it has been checked against what actually happened. The rookie version is a workshop — eight people pick eight dimensions, argue weights until they sum to 100, ship it, and never look again, so the score becomes an average of mediocrity that lands every account between 60 and 75, ranks nothing, and gets ignored. **73% of customer and post-sales leaders say their health score does not reliably predict churn** [2025 Customer Revenue Leadership Study, Pavilion / 6sense, ~800 customer and post-sales leaders — self-reported `[M]`]. The elite version does four things the rookie version never does: writes the prediction as a **falsifiable sentence** before choosing a single input; derives weights from **observed renewal outcomes**; sets the red threshold from **CSM capacity** rather than F1; and ships **reason codes** so the number arrives attached to a play. Read `../cs-context/references/evidence-standard.md` first — a spec quoting an unsourced number as a threshold starts the "where did 70 come from?" argument that runs for three years. ## Before Starting 1. **Read `.agents/cs-context.md`** (fallback `.claude/cs-context.md`). If absent, run `cs-context`. §2 (commercial model, notice period), §5 (activation event), §6 (existing s