ai-testing
AI Safety Evaluation Frameworks: Red-Teaming, Robustness Testing, and Adversarial Probing
AI safety evaluation is not a checkbox — it is an engineering discipline with concrete tools, measurable benchmarks, and adversarial test suites. This guide covers the NIST AI Risk Management Framework as an organizational lens, red-teaming LLMs with Garak and manual probing, adversarial robustness testing with the Adversarial Robustness Toolbox,