AI Model Release Tracker - Timeline of Major AI Models from 2022-2026

Prompt injection robustness

Gray Swan IPIk = 1

Attackers hide malicious instructions inside content the AI reads — a web page, an email, a document — and try to hijack what it does. Gray Swan's indirect prompt injection benchmark measures how often such an attack succeeds when the attacker gets a single try. Lower is better.

Rankings

Lower is better
← All benchmarks