model-evaluation skillA
model-evaluation is agent-read markdown (skill) from aperivue/medsci-skills: Compute and report task-correct held-out metrics for a trained medical-imaging model — segmentation (Dice plus a boundary metric such as HD95 or NSD, per structure), classification (AUROC plus AUPRC and sensitivity/specificity with bootstrap CIs at the deployment prevalence), detection (FROC or mAP with a stated IoU criterion), interactive/promptable segmentation (the interaction-count, convergence, and per-case-time axes a static Dice omits), or generative/synthesis image evaluation (similarity.
Indexed from public GitHub and served as immutable, content-addressed versions. Install it pinned to an exact SHA-256 with the mdr CLI, and every file is verified against the hash recorded here before it reaches your agent. The deterministic audit below grades the latest version, and the same file always earns the same grade.
How to install
mdr add aperivue/medsci-skills/model-evaluation@git:20260705.6badf48mdr add aperivue/medsci-skills/model-evaluation@sha256:b39293e54402fa79Pin to a label to follow the author's releases, or to a sha256 to freeze the exact bytes forever. Either way the resolved hash is written to mdr.lock, and mdr install reproduces it on any machine.
[](https://markdownregistry.com/a/art_urn5unkyqww3i4oh)
0 badge views in 30 days
Versions
| version | committed | commit | size | audit | |
|---|---|---|---|---|---|
| git:20260705.6badf48 latest | 2026-07-05 | 6badf48 | 7,380 B | A | view · diff |
| git:20260705.142a317 | 2026-07-05 | 142a317 | 7,414 B | A | view · diff |
| git:20260705.c91e838 | 2026-07-05 | c91e838 | 6,494 B | A | view · diff |
| git:20260629.657d615 | 2026-06-29 | 657d615 | 5,780 B | A | view · diff |
| git:20260628.530cf75 | 2026-06-28 | 530cf75 | 5,031 B | A | view |
Audit of the latest version
- pass: Frontmatter block present
- pass: Frontmatter declares a name
- pass: Frontmatter declares a description
- pass: Size between 200 bytes and 200 KB (7380 bytes)
- pass: No zero-width or bidi control characters
- pass: No instruction hidden inside an HTML comment
- pass: No link to an exfiltration or paste host
- pass: No credential-shaped string
- pass: No instruction to send local credentials anywhere
- pass: No text hidden with inline styles
- pass: No prompt-injection phrasing
- pass: No curl or wget piped into a shell
- pass: No recursive delete of root, home or parent
- pass: No instruction to read or print local credentials
- pass: No base64 blob over 200 characters
- pass: No link to a raw IP address
- pass: No script tag
Source
aperivue/medsci-skills · 306 stars · license MIT · pushed 2026-09-15 · branch main
API
GET https://markdownregistry.com/api/v1/artifacts/art_urn5unkyqww3i4oh GET https://markdownregistry.com/api/v1/resolve?ref=aperivue/medsci-skills/model-evaluation GET https://markdownregistry.com/api/v1/blob/b39293e54402fa79ac73de827201f4cd98bb2cc70fd1cd4fc8da4e940d429b3b
Agents talk at modelranch.com: hand yours the instructions there and it joins the network that reads files like this one.