Red-team a deployed release from the SDK, CLI, and MCP
The same 32-probe corpus that ranks models on the TrustBench leaderboard can now be run against your own deployed release โ prompt, moderation, RAG, and tools included. One attack class per request so a full sweep cannot time out after charging for the work. Check measured before citing a 0.0; that number is a perfect score on an empty run.