benchmark

8 servers · 1,492★ total

1,446★

ADR secures enterprise AI agents through observability, security benchmarking, and threat detection. Deployed at Uber.

Private, local-first AI assistant for Windows with permissioned tools, durable memory, and evidence-driven small-model improvements.

Agent control plane for governed AI coding: validate changes, enforce policy gates, track findings, proofs, and evals based on your habits.

Assay — a canary-oracle benchmark for MCP security: a frozen task set scored by a recomputable HMAC-canary oracle (structural-zero false positives), not an LLM judge. Maintained by Verosek.

Measure the context-window tax an MCP server charges your AI agent — the real token cost of its tool schemas + responses. Ground truth (Anthropic count_tokens) or keyless estimate. Cross-platform single-file CLI.

Head-to-head benchmark comparing the official MCP to the MCP auto-created by Hintas.

An automated benchmark and public leaderboard for Model-Context Protocol (MCP) clients. Test your client and see how it ranks!