AILuminate
MLCommons (AI Risk & Reliability working group)
AILuminate is MLCommons' family of safety and security benchmarks for generative AI systems, organized around 12 hazard categories. The v1.0 safety benchmark (December 2024) tests general-purpose chat systems with more than 24,000 human-written prompts and grades responses with an ensemble of evaluator models; a jailbreak benchmark measures how safety degrades under attack.