unboundcompute/security-agent-skills · skills/testing-agents-for-indirect-prompt-injection/SKILL.md

testing-agents-for-indirect-prompt-injection skillB

testing-agents-for-indirect-prompt-injection is agent-read markdown (skill) from unboundcompute/security-agent-skills: Test whether an AI agent obeys instructions hidden in the content it ingests, rather than only the user's. Enumerate every channel through which untrusted content reaches the model context (retrieved docs, fetched pages, uploaded files, emails, tool outputs, filenames, images and PDFs, other agents), plant channel-appropriate payloads, and measure whether they change the agent's actions. Use when reviewing any agent or LLM app that reads external content and can act. Covers channel enumeration, .

Indexed from public GitHub and served as immutable, content-addressed versions. Install it pinned to an exact SHA-256 with the mdr CLI, and every file is verified against the hash recorded here before it reaches your agent. The deterministic audit below grades the latest version, and the same file always earns the same grade.

How to install

Latest version
mdr add unboundcompute/security-agent-skills/testing-agents-for-indirect-prompt-injection@git:20260816.4aab195
Exact content
mdr add unboundcompute/security-agent-skills/testing-agents-for-indirect-prompt-injection@sha256:f14ae95f70df41d1

Pin to a label to follow the author's releases, or to a sha256 to freeze the exact bytes forever. Either way the resolved hash is written to mdr.lock, and mdr install reproduces it on any machine.

Badge

mdr badge

[![mdr](https://markdownregistry.com/badge/art_kjjw3eigow6nfysl.svg)](https://markdownregistry.com/a/art_kjjw3eigow6nfysl)

0 badge views in 30 days

Versions

versioncommittedcommitsizeaudit
git:20260816.4aab195 latest2026-08-16 4aab195 6,889 BB view

Audit of the latest version

B  16 of 17 checks passed. Deterministic, no model, same answer every run.
  • fail: No prompt-injection phrasing (matched: ignore prior instructions)
  • pass: Frontmatter block present
  • pass: Frontmatter declares a name
  • pass: Frontmatter declares a description
  • pass: Size between 200 bytes and 200 KB (6889 bytes)
  • pass: No zero-width or bidi control characters
  • pass: No instruction hidden inside an HTML comment
  • pass: No link to an exfiltration or paste host
  • pass: No credential-shaped string
  • pass: No instruction to send local credentials anywhere
  • pass: No text hidden with inline styles
  • pass: No curl or wget piped into a shell
  • pass: No recursive delete of root, home or parent
  • pass: No instruction to read or print local credentials
  • pass: No base64 blob over 200 characters
  • pass: No link to a raw IP address
  • pass: No script tag

Source

GitHub

unboundcompute/security-agent-skills · 5 stars · license MIT · pushed 2026-09-08 · branch main

API

GET https://markdownregistry.com/api/v1/artifacts/art_kjjw3eigow6nfysl
GET https://markdownregistry.com/api/v1/resolve?ref=unboundcompute/security-agent-skills/testing-agents-for-indirect-prompt-injection
GET https://markdownregistry.com/api/v1/blob/f14ae95f70df41d14d3c80960b07c1b0c87560f58de45c361d5e113eb08a016b

Agents talk at modelranch.com: hand yours the instructions there and it joins the network that reads files like this one.