Installable Claude Code plugins from community marketplaces — bundles of commands, agents, hooks and MCP servers. 111 curated, ranked by GitHub signal.
111 plugins in Testing & Debugging
openai
Integration with OpenAI Codex CLI for complex code generation and debugging
libukai
VSCode Extensions Toolkit - 配置 VSCode 扩展的完整工具集,包含 httpYac API 测试、Port Monitor 端口监控、SFTP 静态网站部署
samber
AI Agent Skills for production-ready Go projects. 40+ atomic, cross-referencing Go skills covering code quality, architecture, concurrency, testing, performance, and tooling. Maintained externally by samber.
Owl-Listener
Prototyping and testing skills: wireframe specs, usability heuristics, heuristic evaluations, accessibility audits, A/B test design, and benchmark analysis.
tanweai
WooYun business logic vulnerability methodology — a security testing knowledge plugin distilled from 22,132 public cases, adding evidence-based references, quantitative statistics, and data-driven prioritization.
xu-xiang
The most comprehensive Claude Code plugin — 14+ agents, 56+ skills, 33+ commands, and production-ready hooks for TDD, security scanning, code review, and continuous learning
rshankras
161 Apple platform development skills across 23 categories (iOS, macOS, watchOS, visionOS): code generators, product workflow, App Store/ASO, testing and TDD, design, performance, security, and more. Surfaced as 23 category skills that route to the full library.
karanb192
Remembers approaches you tried and reverted (with reason and estimated token cost) and warns before you retry them — via a prompt-submit card and an ask-before-edit guard.
karanb192
Stamps a provenance receipt — prompts, estimated spend, tests run with exit codes, and a truthful agent-authored line split — into your PR body on `gh pr create`.
mckinsey
Skills and agents for working with Ark - testing, setup, analysis, and research
mozilla
Control Firefox for browsing, web testing, and debugging. Fill forms, capture network and console activity, take screenshots, run scripts, and profile performance. Supports Android devices.
mozilla
Diagnose and fix Firefox DevTools MCP setup issues. Use this if the firefox-devtools-mcp plugin is not working.
safaiyeh
Evaluates code against Apple's App Store Review Guidelines for iOS, macOS, tvOS, watchOS, and visionOS apps
dariushoule
Skills for x64dbg debugger automation — state snapshots, memory analysis, and more
pzep1
Build and manage iOS/macOS apps using native xcodebuild and xcrun simctl CLI tools. Replaces XcodeBuildMCP with direct CLI commands and XCUITest for UI automation.
tmdgusya
Enforce measurement-driven optimization and reproduce-first debugging. Activates automatically on performance and bug-related queries.
rfxlamia
Professional skill and subagent creation with dual-mode workflow: 12-step fast mode and 15-step full mode with behavioral pressure testing and TDD integration.
rfxlamia
Validate implementation plans against DRY, YAGNI, TDD principles and best practices. Identifies gaps, anti-patterns, and improvement opportunities before execution. Prevents over-engineering and ensures test-first approach.
WolframResearch
A full Wolfram Language development environment with code evaluation, documentation search, symbol inspection, static analysis, and test execution.
lgbarn
Ship software systematically: project lifecycle, TDD, parallel agents, code review, security auditing, and infrastructure validation
lucasfcosta
Drive a goal to completion autonomously while enforcing backpressure (lint, tests, benchmarks, reviews, manual testing, PR monitoring) at every step.
tonone-ai
Infrastructure Specialist Team — Chaos: Chaos engineering — failure injection design, game days, resilience testing, blast radius control
tonone-ai
Chaos skill: chaos-recon
tonone-ai
Evaluate model performance — check for accuracy drops, data drift, and error patterns. Use when asked about "model accuracy dropped", "evaluate the model", "check for drift", or "model performance".
tonone-ai
Audit embedding infrastructure — model drift, index freshness, query latency, coverage gaps.
tonone-ai
Design eval harnesses — task schemas, metrics, dataset versioning, eval-as-code patterns.
tonone-ai
Build automated regression suites — golden sets, threshold alerting, CI integration for model changes.
tonone-ai
AI Operations Team — Evals: Eval harness design, benchmark suites, automated regression, human eval orchestration.
tonone-ai
Data quality and pipeline health check — freshness, schema drift, null rates, orphaned records, pipeline status. Use when asked about "data quality check", "pipeline health", "is our data fresh", or "schema drift".
tonone-ai
Diagnose runtime infrastructure issues — cold starts, timeouts, scaling problems, network failures. Use when asked about "infra is slow", "cold starts", "network issues", "why is this timing out", "scaling problem", "latency spikes", or "service is down".
tonone-ai
|
tonone-ai
|
tonone-ai
Audit guardrail coverage — bypass vectors, false positive rates, policy gap analysis, red-team scenarios.
Jamie-BitFlight
tonone-ai
A/B test design — produce an experiment spec with hypothesis, primary metric, MDE, sample size, run time, and decision rule. Also determines when NOT to A/B test and what to do instead. Use when asked to "design an A/B test", "should we test this", "experiment design", "how do we know if this works", "what's the sample size", or "set up an experiment".
tonone-ai
Frontend audit — bundle size, dependencies, accessibility, performance, component quality. Use when asked for "frontend review", "performance audit", "accessibility check", or "bundle size".
tonone-ai
Audit prompt library — duplication, quality, coverage gaps, version drift, eval alignment.
tonone-ai
QA & testing engineer — test strategy, E2E suites, integration testing, test infrastructure, flaky test triage
tonone-ai
Build API test suites — endpoint testing, contract testing, load testing for REST/GraphQL/gRPC APIs. Use when asked to "test this API", "API tests", "endpoint testing", "contract tests", or "load test".
tonone-ai
Audit test suite health — find flaky tests, slow tests, coverage gaps, and testing anti-patterns. Use when asked to "audit tests", "fix flaky tests", "why are tests slow", "test health", or "improve test suite".
tonone-ai
Build E2E test specs for critical user journeys — Playwright or Cypress, page objects, setup/teardown, CI config. Use when asked to "write E2E tests", "end-to-end testing", "browser tests", "UI tests", or "Playwright tests".
tonone-ai
Testing reconnaissance — inventory all tests, frameworks, coverage, CI integration, and assess testing maturity for project takeover. Use when asked to "understand the tests", "testing assessment", "what's tested", or "test inventory".