Skip to content
#

security-benchmark

Here are 56 public repositories matching this topic...

Adversarial security benchmark for agent authorization: does a compromised agent's policy-violating proposal become an unauthorized external effect? 61 trials, nine families, an independent oracle, per-mechanism ablation, confidence intervals. 0 unauthorized effects in 61 attack trials (95% CI [0.0%, 5.9%]). Reproduction is partial.

  • Updated Sep 21, 2026
  • Elixir

Vendor-neutral benchmark measuring how MCP security proxies/gateways DEFEND against 22+ attack vectors — crosswalked to NIST AI RMF & OWASP LLM/Agentic Top 10. CI-gated, reproducible, DOI-cited. Submit your tool to the leaderboard.

  • Updated Sep 30, 2026
  • JavaScript

An AI assistant that can't be hijacked by the documents it reads. The part that decides what to do never sees untrusted text, and every value is tagged with its origin, checked at each action. In a 1440-run offline benchmark: a normal agent was hijacked 73/73, this one 0/73.

  • Updated Sep 28, 2026
  • Python

Add this topic to your repo

To associate your repository with the security-benchmark topic, visit your repo's landing page and select "manage topics."

Learn more