Skip to content

Red-team a deployed release from the SDK, CLI, and MCP

Copy page

Platform

The same 32-probe corpus that ranks models on the TrustBench leaderboard can now be run against your own deployed release โ€” prompt, moderation, RAG, and tools included. One attack class per request so a full sweep cannot time out after charging for the work. Check measured before citing a 0.0; that number is a perfect score on an empty run.

Read the docs โ†’


โ† All changes ยท See it in context