prompt-evaluation-runner skillA
This file is byte-identical to the first copy the registry indexed. Same content hash, same audit grade.
prompt-evaluation-runner is agent-read markdown (skill) from yeaight7/agent-powerups: Use when evaluating prompts, LLM outputs, red-team suites, or model behavior with local eval configs and safe provider/cost controls..
Indexed from public GitHub and served as immutable, content-addressed versions. Install it pinned to an exact SHA-256 with the mdr CLI, and every file is verified against the hash recorded here before it reaches your agent. The deterministic audit below grades the latest version, and the same file always earns the same grade.
What the file says
# Prompt Evaluation Runner ## When to use Use when you need to evaluate an LLM app, test a prompt systematically, or run red-team/vulnerability scans against a target model or application. ## Requirements / Checks 1. Check if an evaluation tool is defined in project deps, scripts, lockfiles, or local toolchain (e.g., `promptfoo`, `evals`, `braintrust`). 2. Do not run unvetted remote runners without checking the project's toolchain first (e.g., avoid `npx promptfoo@latest` if `promptfoo` is already installed locally). 3. If no runner exists, ask before adding a dev dependency or using an ephemeral runner. 4. Confirm expected cost, provider, API keys, and network target before any execution. ## Workflow 1. **Define risk** — state target behavior, failure mode, provider(s), and budget limits before writing any config. 2. **Choose assertions** — prefer deterministic checks first: | Assertion type | When to use | |---|---| | `contains` / `not-contains` | Output must include/exclude specific text | | `regex` | Structured output pattern (e.g., JSON key present) | | `json-schema` | Output must conform to a schema | …
Read the whole file at its exact version.
How to install
mdr add yeaight7/agent-powerups/prompt-evaluation-runner@git:20260606.2be894fmdr add yeaight7/agent-powerups/prompt-evaluation-runner@sha256:676d8ea20a244d66Pin to a label to follow the author's releases, or to a sha256 to freeze the exact bytes forever. Either way the resolved hash is written to mdr.lock, and mdr install reproduces it on any machine.
[](https://markdownregistry.com/a/art_muc7llpkyahgx7yn)
1 badge views in 30 days
Versions
| version | committed | commit | size | audit | |
|---|---|---|---|---|---|
| git:20260606.2be894f latest | 2026-06-06 | 2be894f | 3,329 B | A | view · diff |
| git:20260515.bcf8b2b | 2026-05-15 | bcf8b2b | 3,324 B | A | view · diff |
| git:20260505.ce9dc7a | 2026-05-05 | ce9dc7a | 2,086 B | A | view |
Audit of the latest version
- pass: Frontmatter block present
- pass: Frontmatter declares a name
- pass: Frontmatter declares a description
- pass: Size between 200 bytes and 200 KB (3329 bytes)
- pass: No zero-width or bidi control characters
- pass: No instruction hidden inside an HTML comment
- pass: No link to an exfiltration or paste host
- pass: No credential-shaped string
- pass: No instruction to send local credentials anywhere
- pass: No text hidden with inline styles
- pass: No prompt-injection phrasing
- pass: No curl or wget piped into a shell
- pass: No recursive delete of root, home or parent
- pass: No instruction to read or print local credentials
- pass: No base64 blob over 200 characters
- pass: No link to a raw IP address
- pass: No script tag
Source
yeaight7/agent-powerups · 6 stars · license Apache-2.0 · pushed 2026-09-21 · branch main
API
GET https://markdownregistry.com/api/v1/artifacts/art_muc7llpkyahgx7yn GET https://markdownregistry.com/api/v1/resolve?ref=yeaight7/agent-powerups/prompt-evaluation-runner GET https://markdownregistry.com/api/v1/blob/676d8ea20a244d66e9f5707a50cb733b08378eaf361ce1971ac665ad3b462309
Your agent does the legwork. You hear about the deals worth your word. Hand yours the standing instructions at modelranch.com and it joins the network that reads files like this one.
More from yeaight7/agent-powerups
Every file in yeaight7/agent-powerups