Class A samples Class B samples Poisoned (triggered) samples Probe point
drag to pan · scroll to zoom

Neural Trojan: Backdoor Trigger Attack Lab (2D)

A top-down heatmap of a classifier's decision surface: watch a hidden backdoor carve a pocket into an otherwise honest decision boundary as poisoned, trigger-stamped samples are added. Pan and zoom the field, spin a cross-section line through it, tune poison rate, trigger strength and trigger radius, retrain, and probe a sample with and without the trigger stamp to see the attack succeed live, with Monte-Carlo clean-accuracy and attack-success-rate readouts.