BrowseFull catalogOutcomesSolve a specific problemRolesStack by teamTrustFilter by risk tier
← Back to the Claude Observatory

red-team-blue-team-agent-fabric

Skill Security Usable
Works inClaude Code
Usable Scanned — metadata only

342-test security harness for autonomous AI agents. Pre-certified against NIST AI 800-2, AIUC-1, and covering MCP/A2A protocols. 97.9% production-validated.

342-test security harness for autonomous AI agents. MCP, A2A, x402/L402, AIUC-1 pre-cert, NIST AI 800-2 aligned. 97.9% production-validated. pip install agent-security-harness

23 starsApache-2.0 (commercial OK)FreeQuick setup
Usable rating — This tool is functional but has notable gaps. Review the evaluation notes below before deploying.

342-test security harness for autonomous AI agents. Pre-certified against NIST AI 800-2, AIUC-1, and covering MCP/A2A protocols. 97.9% production-validated.

Security teams evaluating AI agent deployments who need a structured red-team/blue-team testing framework.

Claude Code Claude Cowork Claude Chat

https://github.com/msaleme/red-team-blue-team-agent-fabric

By msaleme

How to Get It

Option 1: Claude Desktop App (Code Mode)Click the + button next to the prompt box → PluginsAdd plugin. Search and click Install. Skills work in Claude Code only.
Option 2: Paste into Claude CodeCopy the command below and paste it into your conversation. Claude will install it.
Command
pip install agent-security-harness

Tip: Paste this into a Claude Code conversation. Verify command matches your Claude Code version.

Auto-generated from the tool's public listing — not hands-on verified. Cross-check against the source repo's README before running.

First thing to try

After installing, paste this into Claude:

Help me run structured red-team and blue-team tests on agent deployments

Trust Signals Auto-scanned

Stars23Contributors6Last updated2026-08-04LicenseApache-2.0 (OK for commercial use)Known CVEsNone foundSources: GitHub Advisory Database + OSV.dev · Scanned 2026-08-12 · scanner vattempted-no-data

Community Pulse Active

Discussed on Hacker News, Reddit

1 mentions across 1 sources

Reviewer notes

Auto-scanned review. These are observations, not a security certification.

catalog_hygiene stale-eval refresh: Scored from trust signals (evidence-eval-v1): 20 GitHub stars; 5 contributors; last commit 1d ago; license Apache-2.0.

Things to check

  • Framework is agent-specific; effectiveness depends on correct MCP endpoint configuration and may require custom rule adaptation for non-standard agent architectures. Production validation rate (97.9%) is self-reported and not independently verified.

How to evaluate tools before deploying →

Data shown here comes from public APIs and automated scanning. Reviewer notes reflect one person's experience. This is not a security certification or legal recommendation. Always evaluate tools according to your own organization's policies.

Evaluation

Ease of Use
3/5
Versatility
2/5
Reliability
4/5
Security
3/5
Overall score3.00 / 5.00 UsableEvaluatedJul 2026
catalog_hygiene stale-eval refresh: Scored from trust signals (evidence-eval-v1): 20 GitHub stars; 5 contributors; last commit 1d ago; license Apache-2.0.

← Back to the Claude Observatory

Rolling Claude out in your org? Let's talk.

Start a conversation →