🔒OpenAPPA Preview: Zero Attacks in Security Benchmarks
New Security Engine Stops All Attacks in Benchmarks
TL;DR
OpenAPPA, an open-source security engine, achieves a perfect score in Bench-Corp and AgentThreatBench with zero successful attacks. It's a game-changer for data exfiltration prevention.
OpenAPPA, an open-source security engine, has achieved a perfect score in security benchmarks Bench-Corp and AgentThreatBench, with zero successful attacks reported. This matters because it offers a robust solution for preventing data exfiltration caused by prompt injection or model hallucination. OpenAPPA runs outside the agent's prompt and execution loop, ensuring security without compromising utility. Benchmarks show a 0% attack success rate and an 89% task completion rate, outperforming competitors like Claude Code and Microsoft FIDES.

Key Points
OpenAPPA stops all attacks in Bench-Corp and AgentThreatBench benchmarks.
Achieves 0% attack success rate and 89% task completion rate.
Runs outside the agent's prompt and execution loop for security.
Security policies in appa.toml file detail data sources, audiences, trust levels.
Disposable Child Branches enable on-demand confinement for untrusted data.
Why It Matters
If you're running complex enterprise workflows with security policies, OpenAPPA's perfect score in Bench-Corp and AgentThreatBench benchmarks means it's a strong candidate for preventing data exfiltration. Its unique approach to security and utility balance could be a game-changer for teams dealing with prompt injection or model hallucination.
Comments
Be the first to comment
Enjoyed this article?
Get it daily. 7am. Free. Reads in 5 minutes.
Join 3,550 builders reading daily.