The Eval Index / Red Teaming & Safety / #112

gy15901580825/Argus

by gy15901580825 · Red Teaming & Safety · updated 18d ago

Black-box, open-source red-team testing for AI agents. Point it at any HTTP, gRPC, or browser-using agent endpoint; run 205 probes mapped to OWASP LLM Top 10 / MITRE ATLAS / NIST AI RMF, incl. payment and MCP tool-abuse families whose verdicts come from captured evidence, not LLM opinion. Reports state what was NOT tested. SARIF, CLI, GH Action.

58
momentum
205
stars
28
forks
#112
rank
View on GitHub →