A block of prompt text did more for agent security than a better model
Baseline scores across the eight models tested ran from 35% to 92%. With one block of text appended to the system prompt, every one of them landed between 95% and 99%. 1Password has released SCAM, an open-source benchmark that gauges how...