BrowseFull catalogOutcomesSolve a specific problemRolesStack by teamTrustFilter by risk tier
← Back to the Claude Observatory

harness-evolver

Skill Development Usable
Works inClaude Code
Usable Scanned — metadata only

Reduces manual tuning of Claude agent configurations by automatically diagnosing failure modes and proposing optimized prompts, routing rules, and retrieval …

Automated harness evolution for AI agents. A Claude Code plugin that iteratively optimizes system prompts, routing, retrieval, and orchestration code using full-trace counterfactual diagnosis. Based on Meta-Harness (Lee et al., 2026).

42 starsMIT (commercial OK)FreeQuick setup
Usable rating — This tool is functional but has notable gaps. Review the evaluation notes below before deploying.

Reduces manual tuning of Claude agent configurations by automatically diagnosing failure modes and proposing optimized prompts, routing rules, and retrieval strategies. Cuts iteration cycles from weeks to hours.

Teams deploying multi-step Claude agents who need faster optimization of system prompts and orchestration logic without hand-tuning each component.

Claude Code Claude Cowork Claude Chat

https://github.com/raphaelchristi/harness-evolver

By raphaelchristi

How to Get It

Option 1: Claude Desktop App (Code Mode)Click the + button next to the prompt box → PluginsAdd plugin. Search and click Install. Skills work in Claude Code only.
Option 2: Paste into Claude CodeCopy the command below and paste it into your conversation. Claude will install it.
Command
claude plugins install raphaelchristi/harness-evolver

Tip: Paste this into a Claude Code conversation. Verify command matches your Claude Code version.

Auto-generated from the tool's public listing — not hands-on verified. Cross-check against the source repo's README before running.

First thing to try

After installing, paste this into Claude:

Help me auto-tune routing logic when agent picks wrong tool for task
CostFree

Trust Signals Auto-scanned

Stars42Contributors1Last updated2026-04-18LicenseMIT (OK for commercial use)Known CVEsNone foundSources: GitHub Advisory Database + OSV.dev · Scanned 2026-07-25 · scanner v1

Community Pulse Growing

Discussed on Reddit

3 mentions across 1 sources

Reviewer notes

Auto-scanned review. These are observations, not a security certification.

Scored from trust signals (evidence-eval-v1): 42 GitHub stars; 1 contributors; last commit 98d ago; license MIT.

Things to check

  • Scanned, not hands-on tested — this entry was auto-scanned from public metadata (GitHub metrics, license, security flags). No reviewer has run it, and no tool-specific limitations have been documented yet.
  • Single maintainer. Consider the risk if this person stops maintaining the project.

How to evaluate tools before deploying →

Data shown here comes from public APIs and automated scanning. Reviewer notes reflect one person's experience. This is not a security certification or legal recommendation. Always evaluate tools according to your own organization's policies.

Evaluation

Ease of Use
3/5
Versatility
2/5
Reliability
3/5
Security
3/5
Overall score2.75 / 5.00 UsableEvaluatedJul 2026
Scored from trust signals (evidence-eval-v1): 42 GitHub stars; 1 contributors; last commit 98d ago; license MIT.

← Back to the Claude Observatory

Rolling Claude out in your org? Let's talk.

Start a conversation →