AI red teaming
Deliberately attacking an AI system — with prompt injection, jailbreaks, and adversarial input — to find its weaknesses before an attacker does.
Deliberately attacking an AI system — with prompt injection, jailbreaks, and adversarial input — to find its weaknesses before an attacker does.