Get the latest tech news
Can open-source prompt-injection detectors catch realistic AI agent attacks?
🛡️ Regex catches 0%, Meta's Prompt Guard 2 catches 1% of 629 realistic AgentDojo injection attacks when they're buried in tool output. Reproducible benchmark. - rudratoshs/buried-injections
None
Or read this on Hacker News