Adoption

Agent Skills are supported by leading AI development tools.

VS Code Gemini CLI GitHub Goose Amp Cursor Claude Code Letta OpenCode Claude OpenAI Codex Factory VS Code Gemini CLI GitHub Goose Amp Cursor Claude Code Letta OpenCode Claude OpenAI Codex Factory

foryourhealth111-pixel/evaluating-machine-learning-models

Name: evaluating-machine-learning-models
Author: foryourhealth111-pixel

bundled/skills/evaluating-machine-learning-models/SKILL.md

npx skillsauth add foryourhealth111-pixel/vco-skills-codex evaluating-machine-learning-models

Clean

TrivyContainer and dependency vulnerability scanner

Clean

SemgrepStatic code analysis for vulnerabilities

Clean

mcp-scan (Snyk)Model Context Protocol security validation

Skipped

Snyk (dep)Open source security scanning

Skipped

Socket.devSupply chain security analysis

Skipped

VirusTotalMulti-engine malware detection

Skipped

CrowdStrikeAdvanced threat intelligence

Skipped

OSV-ScannerOpen Source Vulnerability database check

Skipped

OWASP Dep-Check

Model Evaluation Suite

Use this skill when the model exists and the question is whether it is good enough.

Overview

This skill focuses on choosing and interpreting the right evaluation metrics for the problem, then comparing candidate models or thresholds.

When to Use This Skill

Comparing candidate models with consistent metrics
Reviewing precision/recall/F1/AUC, regression error, calibration, or ranking quality
Stress-testing validation strategy before deployment or publication

Not For / Boundaries

Building the training pipeline itself: use scikit-learn for classical modeling or ml-pipeline-workflow for end-to-end workflow ownership
Engineering features: use preprocessing-data-with-automated-pipelines
Checking train/test contamination: use ml-data-leakage-guard

Typical Outputs

Metric suite recommendations
Model comparison tables
Notes on threshold tradeoffs, calibration, and validation weaknesses

Related Skills

scikit-learn for class-level error breakdowns and confusion matrices
scientific-reporting when the evaluation must become a deliverable

foryourhealth111-pixel/evaluating-machine-learning-models

bundled/skills/evaluating-machine-learning-models/SKILL.md

Evaluate trained machine learning models with the right metrics and comparison logic. Use for benchmark review, threshold selection, calibration, validation, and model comparison; not for feature engineering or leakage auditing.

2,393 stars

testing

Updated Jul 19, 2026

$ install --global

skillsauth

npx skillsauth add foryourhealth111-pixel/vco-skills-codex evaluating-machine-learning-models

Install this skill globally with one command. Works with Claude Code, Cursor, and Windsurf.

Security Scan Results

3 of 9 scanners reported clean

Some scanners were skipped, did not run, or reported a non-clean status. Review each row below.

Scanners Passed

Scanners in report

Clean

TrivyContainer and dependency vulnerability scanner

95%

Clean

SemgrepStatic code analysis for vulnerabilities

95%

Clean

mcp-scan (Snyk)Model Context Protocol security validation

95%

Skipped

Snyk (dep)Open source security scanning

50%

Skipped

Socket.devSupply chain security analysis

50%

Skipped

VirusTotalMulti-engine malware detection

50%

Skipped

CrowdStrikeAdvanced threat intelligence

50%

Skipped

OSV-ScannerOpen Source Vulnerability database check

50%

Skipped

OWASP Dep-Check

50%

Last scanned: Jul 19, 2026, 3:21 AM245.4s8 files scanned

SKILL.md

name:: evaluating-machine-learning-models
description:: |
allowed-tools:: Read, Write, Edit, Grep, Glob, Bash(cmd:*)
version:: 1.0.0
author:: Jeremy Longshore <[email protected]>
license:: MIT

Model Evaluation Suite

Use this skill when the model exists and the question is whether it is good enough.

Overview

This skill focuses on choosing and interpreting the right evaluation metrics for the problem, then comparing candidate models or thresholds.

When to Use This Skill

Comparing candidate models with consistent metrics
Reviewing precision/recall/F1/AUC, regression error, calibration, or ranking quality
Stress-testing validation strategy before deployment or publication

Not For / Boundaries

Building the training pipeline itself: use scikit-learn for classical modeling or ml-pipeline-workflow for end-to-end workflow ownership
Engineering features: use preprocessing-data-with-automated-pipelines
Checking train/test contamination: use ml-data-leakage-guard

Typical Outputs

Metric suite recommendations
Model comparison tables
Notes on threshold tradeoffs, calibration, and validation weaknesses

Related Skills

scikit-learn for class-level error breakdowns and confusion matrices
scientific-reporting when the evaluation must become a deliverable

Related Skills

foryourhealth111-pixel/zarr-python

development

VerifiedTrustedCommunity

Chunked N-D arrays for cloud storage. Compressed arrays, parallel I/O, S3/GCS integration, NumPy/Dask/Xarray compatible, for large-scale scientific computing pipelines.

2,438SKILL.mdUpdated Jul 23, 2026

foryourhealth111-pixel/zarr-python

foryourhealth111-pixel/yeet

tools

VerifiedTrustedCommunity

Use only when the user explicitly asks to stage, commit, push, and open a GitHub pull request in one flow using the GitHub CLI (`gh`).

2,438SKILL.mdUpdated Jul 23, 2026

foryourhealth111-pixel/yeet

foryourhealth111-pixel/xlsx

tools

VerifiedTrustedCommunity

Spreadsheet toolkit (.xlsx/.csv). Create/edit with formulas/formatting, analyze data, visualization, recalculate formulas, for spreadsheet processing and analysis.

2,438SKILL.mdUpdated Jul 23, 2026

foryourhealth111-pixel/xlsx

foryourhealth111-pixel/xan

tools

VerifiedTrustedCommunity

High-performance CSV processing with xan CLI for large tabular datasets, streaming transformations, and low-memory pipelines.

2,438SKILL.mdUpdated Jul 23, 2026

foryourhealth111-pixel/xan

Download

For Claude Desktop. Download once, then upload the file in the app — no terminal needed.

Need help? View full Cowork setup guide →

Install manually

Choose your platform

# Clone the repo
git clone https://github.com/foryourhealth111-pixel/vco-skills-codex.git

# Copy into Claude Code skills folder (global)
cp -r vco-skills-codex/bundled/skills/evaluating-machine-learning-models ~/.claude/skills/

Claude Code Skills — official skills path docs.

Repository

foryourhealth111-pixel/vco-skills-codex

2,393 stars

Compatible with

Claude Code

OpenAI Codex CLI

ChatGPT