Home / yeaight7 / agent-powerups · skills/flaky-test-investigation/SKILL.md · GitHub

flaky-test-investigation skillA

This file is byte-identical to the first copy the registry indexed. Same content hash, same audit grade.

flaky-test-investigation is agent-read markdown (skill) from yeaight7/agent-powerups: Use when tests pass and fail intermittently without code changes, or a test passes alone but fails in the full suite..

Indexed from public GitHub and served as immutable, content-addressed versions. Install it pinned to an exact SHA-256 with the mdr CLI, and every file is verified against the hash recorded here before it reaches your agent. The deterministic audit below grades the latest version, and the same file always earns the same grade.

What the file says

## Purpose

Flaky tests erode trust in CI. Do not just re-run them and hope for the best — isolate the flake vector, fix it, and prove the fix with a stress loop.

## When to Use

- A test fails intermittently in CI but passes locally (or vice versa)
- A test passes alone but fails in the full suite
- A re-run "fixed" a failure and nobody knows why

## Inputs

- The flaky test's name/path and the runner command for it
- Recent failing runs, if available, to estimate the failure rate

## Workflow

1. **Isolate the test.** Run the specific failing test by itself. If it passes alone, the flake is likely an **order dependency** or **state leakage** from a previous test — run the suite up to and including it to confirm.

2. **Stress test.** Run the test in a tight loop to establish the failure rate before changing anything:

   ```bash
   for i in {1..100}; do npm test -- -t "My Test" || echo "FAIL on run $i"; done
   ```

   (Adapt the inner command to the project's runner; some runners have repeat flags built in.)

3. **Check the common vectors:**
   - **Time** — does the test rely on `Date.now()` or `setTimeout`? Mock the clock.
…

Read the whole file at its exact version.

How to install

Latest version
mdr add yeaight7/agent-powerups/flaky-test-investigation@git:20260606.62e2213
Exact content
mdr add yeaight7/agent-powerups/flaky-test-investigation@sha256:66adc4d3a5e1e4f7

Pin to a label to follow the author's releases, or to a sha256 to freeze the exact bytes forever. Either way the resolved hash is written to mdr.lock, and mdr install reproduces it on any machine.

Badge

mdr badge

[![mdr](https://markdownregistry.com/badge/art_jfk74tvwz63rn4xk.svg)](https://markdownregistry.com/a/art_jfk74tvwz63rn4xk)

1 badge views in 30 days

Versions

versioncommittedcommitsizeaudit
git:20260606.62e2213 latest2026-06-06 62e2213 2,542 BA view · diff
git:20260602.b0c6a442026-06-02 b0c6a44 1,172 BA view

Audit of the latest version

A  17 of 17 checks passed. Deterministic, no model, same answer every run.
  • pass: Frontmatter block present
  • pass: Frontmatter declares a name
  • pass: Frontmatter declares a description
  • pass: Size between 200 bytes and 200 KB (2542 bytes)
  • pass: No zero-width or bidi control characters
  • pass: No instruction hidden inside an HTML comment
  • pass: No link to an exfiltration or paste host
  • pass: No credential-shaped string
  • pass: No instruction to send local credentials anywhere
  • pass: No text hidden with inline styles
  • pass: No prompt-injection phrasing
  • pass: No curl or wget piped into a shell
  • pass: No recursive delete of root, home or parent
  • pass: No instruction to read or print local credentials
  • pass: No base64 blob over 200 characters
  • pass: No link to a raw IP address
  • pass: No script tag

Source

GitHub

yeaight7/agent-powerups · 6 stars · license Apache-2.0 · pushed 2026-09-21 · branch main

API

GET https://markdownregistry.com/api/v1/artifacts/art_jfk74tvwz63rn4xk
GET https://markdownregistry.com/api/v1/resolve?ref=yeaight7/agent-powerups/flaky-test-investigation
GET https://markdownregistry.com/api/v1/blob/66adc4d3a5e1e4f75b434bd106495ad3579db33b6b3d47e4d7d7daa77984832a

Your agent does the legwork. You hear about the deals worth your word. Hand yours the standing instructions at modelranch.com and it joins the network that reads files like this one.

More from yeaight7/agent-powerups

AGENTS.md agents
yeaight7/agent-powerups · AGENTS.md
git:20260606.7ba0c89 · audit A · 6 stars
CLAUDE.md claude
yeaight7/agent-powerups · CLAUDE.md
git:20260602.7ef7075 · audit A · 6 stars
AGENTS.md@agents-md/dbt-project agents
yeaight7/agent-powerups · agents-md/dbt-project/AGENTS.md
git:20260428.06dc83e · audit A · 6 stars
AGENTS.md@agents-md/ml-project agents
yeaight7/agent-powerups · agents-md/ml-project/AGENTS.md
git:20260428.c8f26bf · audit A · 6 stars
AGENTS.md@agents-md/open-source-maintainer agents
yeaight7/agent-powerups · agents-md/open-source-maintainer/AGENTS.md
git:20260428.259b8b4 · audit A · 6 stars
AGENTS.md@agents-md/python-library agents
yeaight7/agent-powerups · agents-md/python-library/AGENTS.md
git:20260428.ee0647d · audit A · 6 stars
AGENTS.md@agents-md/typescript-app agents
yeaight7/agent-powerups · agents-md/typescript-app/AGENTS.md
git:20260428.8b5127b · audit A · 6 stars
prompt-evaluation-runner skill
yeaight7/agent-powerups · plugins/agent-evaluation-lab/skills/prompt-evaluation-runner/SKILL.md · Use when evaluating prompts, LLM outputs, red-team suites, or model behavior with local eval configs and safe…
git:20260606.f062e44 · audit A · 6 stars
red-team-eval-authoring skill
yeaight7/agent-powerups · plugins/agent-evaluation-lab/skills/red-team-eval-authoring/SKILL.md · Use when creating or reviewing red-team eval plugins, attack templates, grader rubrics, safety fixtures, or model-risk…
git:20260606.7a1fed0 · audit B · 6 stars
skill-evaluation-workbench skill
yeaight7/agent-powerups · plugins/agent-evaluation-lab/skills/skill-evaluation-workbench/SKILL.md · Use when designing, running, debugging, or hardening deterministic eval suites for agent skills, prompts, tool…
git:20260606.7a2d45f · audit A · 6 stars
agent-harness-design skill
yeaight7/agent-powerups · plugins/agentic-systems/skills/agent-harness-design/SKILL.md · Use when designing tool definitions for a new agent or subagent, an agent shows high retry rates, ambiguous tool…
git:20260606.362908d · audit A · 6 stars
canonical-advisor-routing skill
yeaight7/agent-powerups · plugins/agentic-systems/skills/canonical-advisor-routing/SKILL.md · Use when routing a prompt to a local provider CLI for a second opinion, review, or plan -- you are about to call a…
git:20260606.2fcb67e · audit A · 6 stars

Every file in yeaight7/agent-powerups

Other files named flaky-test-investigation

flaky-test-investigation skill
yeaight7/agent-powerups · plugins/debugging-diagnostics/skills/flaky-test-investigation/SKILL.md · Use when tests pass and fail intermittently without code changes, or a test passes alone but fails in the full suite.
git:20260606.4d49a49 · audit A · 6 stars

Browse by kind, by grade A, or by owner.