Home / pedrohcgs / claude-code-my-workflow · .claude/skills/challenge/SKILL.md · GitHub

challenge skillA

challenge is agent-read markdown (skill) from pedrohcgs/claude-code-my-workflow: Stress-test a finding against the choices you did not make. Enumerates the discrete forks a competent analyst could have taken (measure definition, sample filter, control set, clustering level, weighting, functional form), runs the specification grid, and reports the distribution rather than a point estimate — then attacks the identifying assumption with named, computable sensitivity statistics. Use when the user says "is this robust", "challenge this result", "specification curve", "multiverse".

Indexed from public GitHub and served as immutable, content-addressed versions. Install it pinned to an exact SHA-256 with the mdr CLI, and every file is verified before it reaches your agent: the main file against the SHA-256 recorded here, the others against the git hashes of its source commit. The deterministic audit below grades the latest version, and the same checks always give the same file the same grade.

What the file says

# Challenge — does the result survive the choices you didn't make?

A single specification is one draw from a distribution you never looked at.

**Why this exists, measured rather than asserted.** In a controlled study, 150 autonomous
agents were given the same data and the same questions. Effect-size interquartile ranges
reached **~10.7 %/yr**, and the spread concentrated in **discrete measure-choice forks** — not
in estimation noise. *Within* a measure family, agents agreed to ~0.25 %/yr. Two findings from
that study shape this skill:

- **AI peer review left the spread essentially unchanged.** Review catches errors; it does
  **not** reduce analytical-choice variance. A clean referee report is not robustness.
- Exposure to exemplar papers collapsed the spread by 80–99 % — **convergence by imitation, not
  by correctness.** Herding is not agreement.

So the spread has to be *measured*, not reviewed away.

## Preconditions

- A working baseline specification that runs and produces the headline estimate.
- The estimate's **estimand stated in words** — "the ATT for units treated in 2015, over
…

Read the whole file at its exact version.

How to install

Latest version
mdr add pedrohcgs/claude-code-my-workflow/challenge@git:20260823.6b8fa14
Exact content
mdr add pedrohcgs/claude-code-my-workflow/challenge@sha256:aac4117e29ad3639

Pin to a label to follow the author's releases, or to a sha256 for exact bytes. Either way the resolved hash is written to mdr.lock, and mdr install fetches those bytes again and checks them, so it installs them exactly or fails.

Badge

mdr badge

[![mdr](https://markdownregistry.com/badge/art_lzs2h3ee2fqhgqrt.svg)](https://markdownregistry.com/a/art_lzs2h3ee2fqhgqrt)

2 badge views in 30 days

Versions

versioncommittedcommitsizeaudit
git:20260823.6b8fa14 latest2026-08-23 6b8fa14 7,774 BA view · diff
git:20260822.3bba0ff2026-08-22 3bba0ff 7,810 BA view · diff
git:20260822.1de27662026-08-22 1de2766 8,399 BA view · diff
git:20260822.33ba16e2026-08-22 33ba16e 8,380 BA view · diff
git:20260821.bcadd4a2026-08-21 bcadd4a 8,048 BA view

Audit of the latest version

A  17 of 17 checks passed. Deterministic, no model, same answer every run.
  • pass: Frontmatter block present
  • pass: Frontmatter declares a name
  • pass: Frontmatter declares a description
  • pass: Size between 200 bytes and 200 KB (7774 bytes)
  • pass: No zero-width or bidi control characters
  • pass: No instruction hidden inside an HTML comment
  • pass: No link to an exfiltration or paste host
  • pass: No credential-shaped string
  • pass: No instruction to send local credentials anywhere
  • pass: No text hidden with inline styles
  • pass: No prompt-injection phrasing
  • pass: No curl or wget piped into a shell
  • pass: No recursive delete of root, home or parent
  • pass: No instruction to read or print local credentials
  • pass: No base64 blob over 200 characters
  • pass: No link to a raw IP address
  • pass: No script tag

Source

GitHub

pedrohcgs/claude-code-my-workflow · 1,598 stars · license MIT · pushed 2026-09-27 · branch main

API

GET https://markdownregistry.com/api/v1/artifacts/art_lzs2h3ee2fqhgqrt
GET https://markdownregistry.com/api/v1/resolve?ref=pedrohcgs/claude-code-my-workflow/challenge
GET https://markdownregistry.com/api/v1/blob/aac4117e29ad36394bc7915c30f2ef7a0147a0eff6417c979bfc7739ae113c1b

Your agent does the legwork. You hear about the deals worth your word. Hand yours the standing instructions at modelranch.com and it joins the network that reads files like this one.

More from pedrohcgs/claude-code-my-workflow

adjudicate-review skill
pedrohcgs/claude-code-my-workflow · .claude/skills/adjudicate-review/SKILL.md · Turn an incoming set of findings — from an AI reviewer, a referee report, a code review, a linter, or a second model —…
git:20260927.6535de7 · audit A · 1,598 stars
audit-reproducibility skill
pedrohcgs/claude-code-my-workflow · .claude/skills/audit-reproducibility/SKILL.md · Enforce the replication-protocol.md rule by cross-checking numeric claims in a manuscript against the actual R / Stata…
git:20260927.6535de7 · audit A · 1,598 stars
blast-radius skill
pedrohcgs/claude-code-my-workflow · .claude/skills/blast-radius/SKILL.md · Before and after changing anything shared — a function's return value, a signature, a schema, a label set, a config…
git:20260823.6b8fa14 · audit A · 1,598 stars
capture-environment skill
pedrohcgs/claude-code-my-workflow · .claude/skills/capture-environment/SKILL.md · Snapshot the computational environment for a replication package — detects the analysis stack (R / Stata / Python) and…
git:20260927.6535de7 · audit A · 1,598 stars
checkpoint skill
pedrohcgs/claude-code-my-workflow · .claude/skills/checkpoint/SKILL.md · Save a structured state snapshot before stopping or handing off. Captures the active plan, recent decisions, file…
v1.0.0 · audit A · 1,598 stars
coauthor-brief skill
pedrohcgs/claude-code-my-workflow · .claude/skills/coauthor-brief/SKILL.md · Generate a co-author / collaborator handoff brief for a multi-author, multi-machine project — summarizing what changed…
v1.0.0 · audit A · 1,598 stars
commit skill
pedrohcgs/claude-code-my-workflow · .claude/skills/commit/SKILL.md · Stage, commit, push, open a PR, and merge to main. Use ONLY on explicit commit intent — user says "commit", "ship it"…
git:20260927.6535de7 · audit A · 1,598 stars
compile-latex skill
pedrohcgs/claude-code-my-workflow · .claude/skills/compile-latex/SKILL.md · Compile a Beamer LaTeX slide deck with XeLaTeX (3 passes + bibtex). Use when user says "compile", "build the slides"…
git:20260415.3ef908f · audit A · 1,598 stars
compress-session skill
pedrohcgs/claude-code-my-workflow · .claude/skills/compress-session/SKILL.md · Distill the current conversation into a structured note (decisions made, open questions, file pointers with line…
v1.0.0 · audit A · 1,598 stars
context-status skill
pedrohcgs/claude-code-my-workflow · .claude/skills/context-status/SKILL.md · Show current context status and session health. Use to check how much context has been used, whether auto-compact…
v1.0.0 · audit A · 1,598 stars
create-lecture skill
pedrohcgs/claude-code-my-workflow · .claude/skills/create-lecture/SKILL.md · Create a new Beamer lecture `.tex` from source papers and materials, with notation consistency checks and the project's…
git:20260927.6535de7 · audit A · 1,598 stars
credible-claims skill
pedrohcgs/claude-code-my-workflow · .claude/skills/credible-claims/SKILL.md · Research-brief + claim-record discipline for delegated or AI-assisted research work. Use when starting any substantive…
git:20260821.6600333 · audit A · 1,598 stars

Every file in pedrohcgs/claude-code-my-workflow

Other files named challenge

challenge skill
alirezarezvani/claude-skills · .gemini/skills/challenge/SKILL.md
git:20260309.8902ba7 · audit C · 26,511 stars
challenge skill
oliver-kriska/claude-elixir-phoenix · plugins/elixir-phoenix/skills/challenge/SKILL.md · Challenge mode reviews - rigorous questioning before approving changes. Use when you want thorough scrutiny of Ecto…
git:20260724.918f809 · audit A · 555 stars
challenge skill
athola/claude-night-market · plugins/gauntlet/skills/challenge/SKILL.md · Presents adaptive codebase challenge questions with multiple-choice and trace exercises. Use when testing contributor…
git:20260713.f98719b · audit A · 339 stars
challenge skill
agentii-ai/agentii-investment-intelligence · plugins/vertical-plugins/scenarios/skills/agentii/challenge/SKILL.md · Adversarial verification of research theses — cross-run/cross-thesis contradiction via the entity index, pre-mortem…
git:20260919.eaccfbe · audit A · 206 stars
challenge skill
grainulation/grainulator · skills/challenge/SKILL.md · Adversarial testing of a specific claim. Try to disprove it.
git:20260922.e702f49 · audit A · 86 stars
challenge skill
fusengine/agents · plugins/ai-pilot/skills/challenge/SKILL.md · Use before a root-cause, done/verified claim, irreversible action, or 2nd-time fix reaches the owner (APEX or plain…
git:20260729.3b91eed · audit A · 28 stars
challenge skill
thoughtbot/rails-consultant · skills/challenge/SKILL.md · Pressure-test an assumption, decision, or inherited constraint — Socratic cross-examination that forces you to defend…
git:20260320.c372a42 · audit A · 26 stars
challenge skill
martineserios/thebrana · system/skills/challenge/SKILL.md · Adversarial review — Fable 5 stress-tests reasoning, Gemini checks knowledge. Use before plan or architecture decisions.
git:20260829.ab57a27 · audit A · 3 stars

Browse by kind, by grade A, or by owner.