Offers a generative AI red-teaming test suite that probes diverse large language models for hallucinations, prompt injection, jailbreaks, and other vulnerabilities to assess overall model safety.