MITRE ATLAS 2026 Top AI Detection: Adversarial Backdoor Trigger Pattern in Infer
Detects adversarial backdoor triggers (such as pixel patches, watermark-like artifacts, or malicious token sequences) injected into AI inference inputs. The rule identifies these inputs by observing a combination of suspicious trigger patterns and significant, anomalous jumps in prediction confidence directed at a specific target class, indicative of a model integrity compromise.
Sigma

