benchmark
8 servers · 1,492★ total
ADR secures enterprise AI agents through observability, security benchmarking, and threat detection. Deployed at Uber.
Private, local-first AI assistant for Windows with permissioned tools, durable memory, and evidence-driven small-model improvements.
Agent control plane for governed AI coding: validate changes, enforce policy gates, track findings, proofs, and evals based on your habits.
Assay — a canary-oracle benchmark for MCP security: a frozen task set scored by a recomputable HMAC-canary oracle (structural-zero false positives), not an LLM judge. Maintained by Verosek.
Measure the context-window tax an MCP server charges your AI agent — the real token cost of its tool schemas + responses. Ground truth (Anthropic count_tokens) or keyless estimate. Cross-platform single-file CLI.
Head-to-head benchmark comparing the official MCP to the MCP auto-created by Hintas.
An automated benchmark and public leaderboard for Model-Context Protocol (MCP) clients. Test your client and see how it ranks!