Testing & QA plugins
Installable bundles that ship commands, skills, agents and MCP config together.
33 of 340 entries
ponytail
★ 83,490Lazy senior dev mode. Forces the simplest, shortest solution that actually works: YAGNI, stdlib first, no unrequested abstractions.
DietrichGebertupdated 19d agoMITruflo-loop-workers
★ 64,474Cache-aware /loop workers and CronCreate background automation — wraps 5 hooks_worker-* MCP tools (list/dispatch/status/detect/cancel) and exposes 12 background worker triggers (ultralearn, optimize, consolidate, predict, audit, map, preload, deepdive, document, refactor, benchmark, testgaps)
ruvnetupdated 14d agoMITruflo-neural-trader
★ 64,474Neural trading via npx neural-trader — self-learning strategies, Rust/NAPI backtesting, 112+ MCP tools, swarm coordination, and portfolio optimization
ruvnetupdated 14d agoMITruflo-testgen
★ 64,474Test gap detection, coverage analysis, and automated test generation — drives the testgaps background worker via hooks_worker-dispatch; SPARC Refinement-phase canonical owner
ruvnetupdated 14d agoMITruflo-browser
★ 64,474Session-as-skill browser automation: Playwright + RVF cognitive containers + ruvector trajectories + AgentDB selector memory + AIDefence PII/injection gates
ruvnetupdated 14d agoMITagentic-bundle-qa-testing
★ 43,285Editorial "QA & Testing" bundle for Claude Code from Agentic Awesome Skills.
sickn33updated 14d agoMITagentic-bundle-aas-qa-test-automation
★ 43,285Editorial "AAS QA & Test Automation" bundle for Claude Code from Agentic Awesome Skills.
sickn33updated 14d agoMITpr-review
★ 39,867Complete PR review workflow with security, testing, and docs
luongnv89updated 18d agoMITquantitative-trading
★ 37,927Quantitative analysis, algorithmic trading strategies, financial modeling, portfolio risk management, and backtesting
wshobsonupdated 14d agoMITbackend-development
★ 37,927Backend API design, GraphQL architecture, workflow orchestration with Temporal, and test-driven backend development
wshobsonupdated 14d agoMITperformance-testing-review
★ 37,927Performance analysis, test coverage review, and AI-powered code quality assessment
wshobsonupdated 14d agoMITtdd-workflows
★ 37,927Test-driven development methodology with red-green-refactor cycles and code review
wshobsonupdated 14d agoMITshell-scripting
★ 37,927Production-grade Bash scripting with defensive programming, POSIX compliance, and comprehensive testing
wshobsonupdated 14d agoMITapi-testing-observability
★ 37,927API testing automation, request mocking, OpenAPI documentation generation, observability setup, and monitoring
wshobsonupdated 14d agoMITship-mate
★ 37,927Your AI development teammate. Turns a story file into a shipped, reviewed, and tested feature via orchestrator → architect → developer → PR reviewer → QA → Playwright.
wshobsonupdated 14d agoMITdeveloper-essentials
★ 37,927Essential developer skills including Git workflows, SQL optimization, error handling, code review, E2E testing, authentication, debugging, and monorepo management
wshobsonupdated 14d agoMITunit-testing
★ 37,927Unit and integration test automation for Python and JavaScript with debugging support
wshobsonupdated 14d agoMITfull-stack-orchestration
★ 37,927End-to-end feature orchestration with testing, security, performance, and deployment
wshobsonupdated 14d agoMITplaywright
★ 32,159Browser automation and end-to-end testing MCP server by Microsoft. Enables Claude to interact with web pages, take screenshots, fill forms, click elements, and perform automated browser testing workflows.
anthropicsupdated 14d agoApache-2.0fakechat
★ 32,159Localhost iMessage-style web chat for Claude Code — test surface with file upload and edits. No tokens, no access control.
anthropicsupdated 14d agoApache-2.0skill-creator
★ 32,159Create new skills, improve existing skills, and measure skill performance. Use when users want to create a skill from scratch, update or optimize an existing skill, run evals to test a skill, or benchmark skill performance with variance analysis.
anthropicsupdated 14d agoApache-2.0pr-review-toolkit
★ 32,159Comprehensive PR review agents specializing in comments, tests, error handling, type design, code quality, and code simplification
anthropicsupdated 14d agocybersecurity-skills
★ 25,606817 cybersecurity skills covering web security, pentesting, DFIR, threat intelligence, cloud security, malware analysis, and more.
mukul975updated 1mo agoApache-2.0chief-customer-officer-advisor
★ 22,612Chief Customer Officer advisory for startups: retention decomposition analyzer (honest GRR vs NRR + 7-category churn taxonomy), customer segmentation designer (4-tier framework + ICP fit scoring + kill list), CS coverage calculator (pooled vs named CSM models + ratio math + 12-month hiring plan). 4 in-depth references: retention decomposition, customer segmentation strategy, CS coverage model, CS team org evolution (CSM vs Support vs AM vs IM vs CS Ops). Stdlib-only. Standalone-installable; also bundled in c-level-skills. Strategic only - does not duplicate business-growth tactical CS skills.
alirezarezvaniupdated 15d agoMITprompt-governance
★ 22,612Use when managing prompts in production at scale: versioning prompts, running A/B tests on prompts, building prompt registries, preventing prompt regressions, or creating eval pipelines for production
alirezarezvaniupdated 15d agoMITdossier
★ 22,612Decision-grade entity research skill — produces a hypothesis-tested dossier on a specific company, person, nonprofit, or government org, not a generic profile. Forcing intake makes the user state their hypothesis upfront (what they already believe and want to verify or disprove) so the dossier tests it rather than confirms it. Output is an editable Word document (.docx) with verdict on the hypothesis, identity facts, 12-month activity timeline, network signals, reputation signals, red flags, 3-5 conversation hooks tied to specific findings, and source-provenance audit log. Uses WebSearch + WebFetch + free APIs (SEC EDGAR, GitHub, ProPublica Nonprofit Explorer) as workhorses; optional BYOK MCPs (LinkedIn, Crunchbase, Apollo, Pitchbook, SimilarWeb) enhance coverage. Triggers: 'research [company]', 'dossier on [person/company]', 'background check on [entity]', 'prep me for a meeting with [person/company]', 'due diligence on [company]', 'what should I know about [entity]', 'research [person] before I [meet/hire/invest]', 'competitor research on [company]', 'investor diligence [company]', 'interview prep for [company]'. Honors sensitivity exclusions for journalism + personal-vetting contexts.
alirezarezvaniupdated 15d agoMITpw
★ 22,612Production-grade Playwright testing toolkit. Generate tests from specs, fix flaky failures, migrate from Cypress/Selenium, sync with TestRail, run on BrowserStack. 55+ ready-to-use templates, 3 specialized agents, smart reporting that plugs into your existing workflow.
alirezarezvaniupdated 15d agoMITexecutive-mentor
★ 22,612Adversarial thinking partner for founders and executives. Stress-tests plans, prepares for board meetings, navigates hard decisions, and forces honest post-mortems.
alirezarezvaniupdated 15d agoMITdemo-video
★ 22,612Create polished demo videos from screenshots and scene descriptions. Orchestrates playwright, ffmpeg, and edge-tts to produce product walkthroughs, feature showcases, and marketing teasers with story structure, scene design system, and narration guidance.
alirezarezvaniupdated 15d agoMITroast
★ 22,612Pressure-test a business idea before you build it. Convenes a 5-angle adversarial panel — The Critic (what kills this?), The Champion (the 10x upside?), The Analyst (does the logic hold?), The Investigator (what does the market say?), and The Customer (would I actually pay?) — fired in parallel as independent reviewers, then a Judge synthesizes one GO / RESHAPE / KILL verdict with explicit confidence and the cheapest 48-hour test to de-risk it. Never averages the scores: a weighted synthesizer with demand/fatal-flaw/logic veto gates produces the call, backed by deterministic stdlib tools. The opposite of Claude's default agreeableness.
alirezarezvaniupdated 15d agoMITengineering-skills
★ 22,61232 production-ready engineering skills: architecture, frontend, backend, fullstack, QA, DevOps, security, AI/ML, data engineering, Playwright, self-improving agent, security suite (adversarial-reviewer, ai-security, cloud-security, incident-response, red-team, threat-detection), Stripe integration, TDD guide, Google Workspace CLI, a11y audit, Snowflake development, and more. v2.8.1 augments senior-fullstack / senior-frontend / senior-backend with karpathy-coder + Matt Pocock discipline: each ships a 7-question forcing-question library, 4 customization profiles (JSON, swappable), a deterministic decision engine, a composition map into POWERFUL specialists, plus cs-fullstack-engineer / cs-frontend-engineer / cs-backend-engineer orchestrator agents (context: fork) invokable by other agents via /cs:fullstack-review, /cs:frontend-review, /cs:backend-review, /cs:engineer-grill. Agent skill and plugin for Claude Code, Codex, Gemini CLI, Cursor, OpenClaw.
alirezarezvaniupdated 15d agoMITstatistical-analyst
★ 22,612Hypothesis testing, A/B experiment analysis, sample size calculation, and confidence intervals. 3 stdlib-only Python tools with Z-test, t-test, chi-square, effect sizes, power analysis, and Wilson score intervals.
alirezarezvaniupdated 15d agoMITcompliance-team-iso42001
★ 22,612ISO/IEC 42001:2023 AI Management System (AIMS) specialist for compliance teams: AIMS gap analyzer (Clauses 4-10 coverage scoring + remediation priority), AI risk register builder (Annex A 38 controls + risk-to-treatment map per ISO 23894), AIMS audit scheduler (Clause 9.2 internal audit cadence + 12-month plan + auditor independence checks). 4 in-depth references: ISO 42001 Clauses 4-10 walkthrough, Annex A controls A.1-A.10, AIMS implementation maturity model, cross-framework mapping (42001 ↔ EU AI Act ↔ NIST AI RMF ↔ ISO 23894). Stdlib-only. Standalone-installable; also bundled in ra-qm-skills. Built for compliance officers running internal AIMS audits, not for executive AI strategy decisions (see chief-ai-officer-advisor for those).
alirezarezvaniupdated 15d agoMIT