For Heads-of · Practitioner

Adversarial robustness testing

Structured red-teaming and adversarial testing of an AI system against known attack techniques before deployment and on a recurring cadence after.

  • preventive
  • adversarial-ml
  • red-teaming
  • testing

What it does

Runs a defined battery of adversarial tests, evasion attempts, prompt injection probes, extraction queries, backdoor triggers, against a model before it ships and periodically once it is in production.

Where it fits

The direct mitigation for the adversarial-testing-coverage-gap risk, and the empirical evidence behind any claim of model robustness.

Risks this mitigates

The risks this control addresses, ranked by effectiveness.

Framework and clause references

FrameworkClauseTitle
NIST AI Risk Management Framework (AI RMF 1.0)MeasureMeasure
MITRE ATLASAML.T0015Evade AI Model