Skip to content
Built here

ML model evaluation

/ml--model-evaluation

/ml--model-evaluation is a Claude Code skill in the AI & Agents section. Evaluates a model beyond one headline number: metrics matched to error costs, calibration, per-slice performance and error-by-error analysis.

Author's description

Evaluate a model properly — right metrics, calibration, slices, error analysis

How to ask for it

  • /ml--model-evaluation evaluate this fraud classification model
  • /ml--model-evaluation check whether the model is well calibrated
  • /ml--model-evaluation break down errors by user segment

Install

curl -fsSL https://raw.githubusercontent.com/sgomez-dev/claude-skills/main/install.sh | bash

After installing with the script, type /ml--model-evaluation. Using Cursor, Windsurf or Codex? Platform guides

Demo coming soon

↑↓ move · Enter open · Esc close