Jev Classifier Performance

This project tests the effectiveness of the Jev classifier in preventing AI coding agents from executing dangerous shell commands, comparing its performance against another agent named Laya.
Can a classifier stop an AI coding agent from running dangerous or malicious shell commands?
I tested @typesafeai's Jev against Laya, run locally on CPU and GPU.
Laya on GPU: 4.5× faster. Attacks let through: 44 of 111. Jev: 6. At 856 attempts: 114 vs 9.
Faster isn't safer. x.com/lovelylogicss/status/210250173113736
