assay

metahub-ai/assay
★ 2 stars TypeScript AI/LLM Updated today
An open, reproducible framework for evaluating AI artifacts — skills, MCP servers, agents and plugins. Reads what is inside them, optionally runs them in a sandbox and judges what they did, and publishes a report a stranger can verify.
View on GitHub → 🔍 Audit Wallet Slippage →

Quick Install

Copy the config for your editor. Some servers may need additional setup — check the README.

Add to claude_desktop_config.json:

{
  "mcpServers": {
    "assay": {
      "command": "npx",
      "args": [
        "-y",
        "metahub-ai/assay"
      ]
    }
  }
}

Topics

aiai-agentsai-toolsbenchmarkclaudeclaude-skillsevalevaluationllmllm-evaluationmcpmcp-serverreproducible-researchtesting