Search everything...

Stats

Actions

Available In

qa-flake-triage

Name: qa-flake-triage
Author: testland

By testland

Flake triage: 2 skills (flaky-test-quarantine, flake-pattern-reference) and 5 agents (e2e-flake-bisector, parallel-isolation-checker, regression-bisector, ai-flake-detector, e2e-test-trend-reporter).

npx claudepluginhub testland/qa --plugin qa-flake-triage

Popularity

Stars

Med: 0·Avg: 285

Installs

Med: 0·Avg: 1

What's Inside

Agents5

ai-flake-detector

/ai-flake-detector

Reads historical CI test results (JUnit XML or vendor JSON) and predicts which currently-green tests are likely to go flaky next, using signals from the 8-pattern catalog (test size correlation, async waits with fixed sleeps, parallel-execution heuristics). Returns a ranked watchlist with rationale per test. Use proactively as a weekly screen across a large suite to focus prevention effort before the test starts failing.

e2e-flake-bisector

/e2e-flake-bisector

Runs a target end-to-end test N times under varied conditions (worker isolation, test order, viewport, network throttling, parallelism) to identify the axis along which the flake reproduces. Returns a probable root cause classified against the 8 flake patterns plus a numeric reproduction rate per axis. Use when a test has been flagged flaky and the team needs to know which condition triggers the failure.

e2e-test-trend-reporter

/e2e-test-trend-reporter

Generates a periodic (weekly / monthly) test-suite health report from CI history - total runs, suite duration, flakiness rate, top failing tests, time-to-green per PR, week-over-week deltas. Emits a markdown summary suitable for a team Slack channel or wiki page. Use as a scheduled CI job to keep test health visible.

parallel-isolation-checker

/parallel-isolation-checker

Inspects a test suite that flakes under parallel execution and identifies the specific shared state - DB rows, env vars, files, ports, lockfiles, or global module state - that workers are colliding on. Runs targeted instrumentation around suspect resources, correlates each test's writes with another worker's reads, and reports the colliding resource with file:line evidence. Use after `e2e-flake-bisector` has implicated parallel execution.

regression-bisector

/regression-bisector

Orchestrates `git bisect` against a target test or build script to identify the introducing commit of a regression. Wraps the bad/good marking, the `git bisect run` script, the 125 exit code for unbuildable revisions, and the final culprit report. Use when a test that previously passed has started failing 100% of the time on the trunk.

Skills4

flake-dashboard-author

/flake-dashboard-author

Builds a persistent flakiness infrastructure dashboard from JUnit XML or JSON CI run history: defines the flake-rate metric (failures per test over a configurable window), authors the data model, generates a Grafana time-series panel JSON or configures a Datadog CI Visibility view, derives the quarantine-candidate query, and wires trend alerts. Use when a team needs a long-lived observability surface for test reliability that outlasts any single weekly report.

flake-pattern-reference

/flake-pattern-reference

Reference catalog of flake patterns - async/timing, test ordering, shared parallel state, resource leaks, network, locator drift, environment variance, randomness - with detection heuristics and remediation per pattern. Use when triaging an unknown flake to identify the category before bisecting.

flake-remediation-guide

/flake-remediation-guide

Provides concrete code-level fixes for each of the eight recurring flake patterns cataloged in flake-pattern-reference - replacing fixed sleeps with framework auto-waits, isolating state in beforeEach fixtures, adopting stable role-based locators, mocking network and clock, seeding RNG, and closing leaked resources. Use when a flake has been classified by pattern and the engineer needs the specific code change to apply.

flaky-test-quarantine

/flaky-test-quarantine

Builds a quarantine workflow for flaky tests - marks the test with the framework's skip/fixme/retry annotation, records the failure-rate observation and a bisect link in the annotation body, sets an auto-expiry date, and produces a CI report listing every quarantined test that has expired and needs re-evaluation. Use when a flaky test is blocking the trunk and must be removed from the gating path without losing track of it.

Stats

Version1.1.0

LanguagePython

Stars0

MaintenanceExcellent

LicenseMIT

Last CommitJun 4, 2026

AddedJun 9, 2026

Actions

View on GitHub View README Plugin Marketplace JSON Homepage

Own this plugin?

Verify ownership to unlock analytics, metadata editing, and a verified badge. GitHub access is read-only (username + org membership).

Available In

testland-qa

Safety Signals

Caution

Uses power tools

Uses Bash, Write, or Edit tools

README

testland-qa

A rigorously curated quality-engineering plugin marketplace for Claude Code. 77 plugins, 695 components, every one rating-gated before merge.

Why testland-qa

6-dimension quality rubric (D1–D6) before merge, with a hard-reject for uncited claims (citation theater) via the d6 floor
CI-validated composition: every agent's preloaded skills are reference-checked, no dangling deps
Differentiation required: every component must articulate how it differs from its nearest neighbors. Generic, persona-shaped scopes that can't name a trigger condition get sent back for reshaping
Reviewer-calibrated: two-evaluator rubric, A/C/F-grade exemplars in docs/REVIEWER_TRAINING.md

See Quality bar and docs/REVIEWER_CHECKLIST.md.

How it works

The marketplace ships three kinds of building block:

Plugin — an installable bundle scoped to one QA area (e.g. qa-api-testing, qa-load-testing). You install only the plugins your stack needs.
Skill — an atomic, self-contained capability inside a plugin, usually wrapping one tool or one technique (e.g. great-expectations, oauth-flow-test-author). Claude loads a skill when your request matches its trigger; you can also ask for it by name.
Agent — a task-scoped subagent that runs one focused job (e.g. schema-diff-reviewer reviews a migration diff and returns a findings table). An agent may preload one or more skills to do its work.

Installed components stay dormant until a matching task comes up, so adding a plugin doesn't add noise — it adds capability that activates on demand.

Install

Claude Code marketplace (recommended)

/plugin marketplace add testland/qa
/plugin install <plugin-name>@testland-qa

For example:

/plugin install qa-data-quality@testland-qa

Direct URL

/plugin marketplace add https://github.com/testland/qa

Manual / hermetic environments

git clone https://github.com/testland/qa ~/.claude/marketplaces/testland-qa

Before you install: plugins run inside your Claude Code session and ship agent instructions and tool wrappers. Anthropic doesn't vet marketplace contents — review a plugin's components before installing it into a sensitive project. Every component here is rating-gated (see Quality bar), but you remain in control of what runs.

Start here

New to the marketplace? Install one or two plugins for your role rather than everything — components activate on demand, so a focused set keeps things sharp.

If you're a…	Try first
Manual / exploratory tester	qa-manual-testing · qa-bdd · qa-bug-repro
Test automation engineer	qa-web-e2e · qa-api-testing · qa-unit-tests-js
Performance engineer	qa-load-testing · qa-chaos-resilience
Security tester	qa-sast · qa-secrets · qa-dast
Lead / manager / head of quality	qa-roles · qa-test-management · qa-process

The full catalog is below; for versions and component counts see CATALOG.md.

Using an installed plugin

Once a plugin is installed, its skills and agents are available to Claude Code — invoke them by describing the task in plain language. Example with qa-data-quality:

/plugin install qa-data-quality@testland-qa

Ask "add Great Expectations checks to my orders pipeline" → the great-expectations skill scaffolds an ExpectationSuite + Checkpoint and wires the results into a CI gate.
On a database change, ask "review this migration's schema diff" → the schema-diff-reviewer agent returns a Critical / Warning / Info findings table covering breaking-vs-additive changes and downstream impact.

Each plugin's README.md lists its skills and agents and what each one does.

Plugin catalog

View full README on GitHub

qa-flake-triage

Popularity

What's Inside

Confidence

README

testland-qa

Why testland-qa

How it works

Install

Claude Code marketplace (recommended)

Direct URL

Manual / hermetic environments

Start here

Using an installed plugin

Plugin catalog

Similar Plugins

unity-dev-toolkit

dotnet-skills

r-skills

creative-writing

everything-claude-code

claude-seo

More by testland

qa-visual-regression

qa-contract-testing

qa-bug-repro

qa-api-testing

qa-data-quality

testland-qa

Why testland-qa

How it works

Install

Claude Code marketplace (recommended)

Direct URL

Manual / hermetic environments

Start here

Using an installed plugin

Plugin catalog

Popularity

Health & Quality

More by testland

qa-visual-regression

qa-contract-testing

qa-bug-repro

qa-api-testing

qa-data-quality

Similar Plugins

unity-dev-toolkit

dotnet-skills

r-skills

creative-writing

everything-claude-code

claude-seo