Home / ai-analyst-lab / ai-analyst · .claude/skills/experiment/SKILL.md · GitHub

experiment skillA

experiment is agent-read markdown (skill) from ai-analyst-lab/ai-analyst: The analysis and lifecycle owner for experiments. Full experiment lifecycle: design, power analysis, statistical analysis, interpretation, reporting, and monitoring of A/B tests. Invoke as /experiment. Trigger on "A/B test", "experiment", "treatment vs control", "sample size", "MDE", "statistical significance", "ship decision", "test readout", "is this result significant?". Runs the SRM gate first..

Indexed from public GitHub and served as immutable, content-addressed versions. Install it pinned to an exact SHA-256 with the mdr CLI, and every file is verified against the hash recorded here before it reaches your agent. The deterministic audit below grades the latest version, and the same file always earns the same grade.

What the file says

# Skill: /experiment — OpenXP Experimentation Platform

## Purpose
Multi-mode skill for the full experiment lifecycle — from design through analysis to ship/no-ship decision. Orchestrates experiment agents and calls coded statistical helpers from `helpers/stats/experiment_stats/` instead of improvising Python.

## When to Use
Invoke as `/experiment [mode]` or trigger on experiment-related intents:
- "I want to run an experiment"
- "Analyze this A/B test"
- "Did this experiment work?"
- "What's the power for this test?"

## Modes

### `/experiment design`
**Purpose:** Create a pre-registered experiment config.
**Agent:** `agents/experiments/experiment-designer.md`
**Flow:**
1. Run Experiment Brief skill to capture hypothesis, north star, guardrails
2. Invoke Experiment Designer agent
3. Output: `experiments/{slug}/experiment.yaml` (from `templates/experiment.yaml`)
**Checkpoint:** Config review (Type B — skippable with --just-do-it)

### `/experiment power`
**Purpose:** Power analysis + duration estimation.
**Flow:**
1. Read `experiments/{slug}/experiment.yaml` for metric type, baseline, MDE
2. Call `helpers/stats/experiment_stats/power.py`:
…

Read the whole file at its exact version.

How to install

Latest version
mdr add ai-analyst-lab/ai-analyst/experiment@git:20260902.b370de6
Exact content
mdr add ai-analyst-lab/ai-analyst/experiment@sha256:4eeb325633fac25f

Pin to a label to follow the author's releases, or to a sha256 to freeze the exact bytes forever. Either way the resolved hash is written to mdr.lock, and mdr install reproduces it on any machine.

Badge

mdr badge

[![mdr](https://markdownregistry.com/badge/art_3bj6wprpsrejp3d6.svg)](https://markdownregistry.com/a/art_3bj6wprpsrejp3d6)

1 badge views in 30 days

Versions

versioncommittedcommitsizeaudit
git:20260902.b370de6 latest2026-09-02 b370de6 8,023 BA view · diff
git:20260827.7ff2e252026-08-27 7ff2e25 8,034 BA view

Audit of the latest version

A  17 of 17 checks passed. Deterministic, no model, same answer every run.
  • pass: Frontmatter block present
  • pass: Frontmatter declares a name
  • pass: Frontmatter declares a description
  • pass: Size between 200 bytes and 200 KB (8023 bytes)
  • pass: No zero-width or bidi control characters
  • pass: No instruction hidden inside an HTML comment
  • pass: No link to an exfiltration or paste host
  • pass: No credential-shaped string
  • pass: No instruction to send local credentials anywhere
  • pass: No text hidden with inline styles
  • pass: No prompt-injection phrasing
  • pass: No curl or wget piped into a shell
  • pass: No recursive delete of root, home or parent
  • pass: No instruction to read or print local credentials
  • pass: No base64 blob over 200 characters
  • pass: No link to a raw IP address
  • pass: No script tag

Source

GitHub

ai-analyst-lab/ai-analyst · 302 stars · license MIT · pushed 2026-09-22 · branch main

API

GET https://markdownregistry.com/api/v1/artifacts/art_3bj6wprpsrejp3d6
GET https://markdownregistry.com/api/v1/resolve?ref=ai-analyst-lab/ai-analyst/experiment
GET https://markdownregistry.com/api/v1/blob/4eeb325633fac25f255dff418804d7a4782961515fe92e648233e5a9f7d829d2

Your agent does the legwork. You hear about the deals worth your word. Hand yours the standing instructions at modelranch.com and it joins the network that reads files like this one.

More from ai-analyst-lab/ai-analyst

always-compare skill
ai-analyst-lab/ai-analyst · .claude/skills/always-compare/SKILL.md · Never present a metric or number in isolation; anchor every number to a comparison (prior period, benchmark, or another…
git:20260827.7ff2e25 · audit A · 302 stars
analysis-design skill
ai-analyst-lab/ai-analyst · .claude/skills/analysis-design/SKILL.md · Takes a vague analytical hunch, stakeholder request, or business question and produces a rigorous, stakeholder-ready…
git:20260902.b370de6 · audit A · 302 stars
analyst-core skill
ai-analyst-lab/ai-analyst · .claude/skills/analyst-core/SKILL.md · Operating rules for every data analysis. Apply for ANY data-analysis intent: "analyze", "investigate", "why did X…
git:20260921.652dcb5 · audit A · 302 stars
archaeology skill
ai-analyst-lab/ai-analyst · .claude/skills/archaeology/SKILL.md · Retrieve proven SQL patterns, table cheatsheets, and join patterns from .knowledge/query-archaeology/ so past work gets…
git:20260827.7ff2e25 · audit A · 302 stars
archive-analysis skill
ai-analyst-lab/ai-analyst · .claude/skills/archive-analysis/SKILL.md · Save completed analyses to the knowledge system's analysis archive for future reference. Use this skill after…
git:20260902.b370de6 · audit A · 302 stars
auth-preflight skill
ai-analyst-lab/ai-analyst · .claude/skills/auth-preflight/SKILL.md · Verify Google Workspace MCP authentication at the start of any session that needs Google APIs (Docs, Slides, Drive)…
git:20260902.b370de6 · audit A · 302 stars
business skill
ai-analyst-lab/ai-analyst · .claude/skills/business/SKILL.md · Browse, search, and explore your organization's business context system — glossary terms, product catalog, metric…
git:20260902.b370de6 · audit A · 302 stars
causal skill
ai-analyst-lab/ai-analyst · .claude/skills/causal/SKILL.md · Causal inference toolkit for when experiments are not possible: estimate treatment effects from observational data with…
git:20260828.01c26bc · audit A · 302 stars
chart-to-drive skill
ai-analyst-lab/ai-analyst · .claude/skills/chart-to-drive/SKILL.md · Standardized workflow for uploading local chart PNGs to Google Drive and making them available for insertion into…
git:20260827.7ff2e25 · audit A · 302 stars
close-the-loop skill
ai-analyst-lab/ai-analyst · .claude/skills/close-the-loop/SKILL.md · Ensure every analysis that includes a recommendation ends with a clear, actionable follow-up plan. CRITICAL RULE - Only…
git:20260902.b370de6 · audit A · 302 stars
codex-review skill
ai-analyst-lab/ai-analyst · .claude/skills/codex-review/SKILL.md · Independently validate the current analysis with a second model (OpenAI Codex). Codex re-derives the same answer from…
git:20260902.b370de6 · audit A · 302 stars
compare-datasets skill
ai-analyst-lab/ai-analyst · .claude/skills/compare-datasets/SKILL.md · Compare metrics, findings, and patterns across two or more connected datasets. Helps identify cross-dataset patterns…
git:20260902.b370de6 · audit A · 302 stars

Every file in ai-analyst-lab/ai-analyst

Other files named experiment

experiment skill
sethgammon/citadel · skills/experiment/SKILL.md · Automated optimization loop with scalar fitness function. Proposes changes in isolated worktrees, measures with a…
git:20260611.0f3f8e7 · audit A · 923 stars
experiment skill
insajin/autopus-adk · .omp/skills/experiment/SKILL.md · Experiment loop for iterative metric-driven code optimization using XLOOP
git:20260826.34a9b20 · audit A · 111 stars
experiment skill
simota/agent-skills · experiment/SKILL.md · Designing A/B tests: hypothesis docs, sample size, feature flags, significance analysis, CUPED, SRM detection…
git:20260918.e307415 · audit A · 80 stars
experiment skill
ariaxhan/kernel-claude · skills/experiment/SKILL.md · Rules as hypotheses: falsifiable tests, confidence updates, graduate or kill. Triggers: experiment, hypothesis, prove…
git:20260910.20bff72 · audit A · 13 stars

Browse by kind, by grade A, or by owner.