Built here
ML model evaluation
/ml--model-evaluation
/ml--model-evaluation is a Claude Code skill in the AI & Agents section. Evaluates a model beyond one headline number: metrics matched to error costs, calibration, per-slice performance and error-by-error analysis.
Author's description
Evaluate a model properly — right metrics, calibration, slices, error analysis
How to ask for it
/ml--model-evaluation evaluate this fraud classification model/ml--model-evaluation check whether the model is well calibrated/ml--model-evaluation break down errors by user segment
Install
curl -fsSL https://raw.githubusercontent.com/sgomez-dev/claude-skills/main/install.sh | bashAfter installing with the script, type /ml--model-evaluation. Using Cursor, Windsurf or Codex? Platform guides
Demo coming soon