Skip to content
InfoQ·

🔒OpenAPPA Preview: Zero Attacks in Security Benchmarks

New Security Engine Stops All Attacks in Benchmarks

TL;DR

OpenAPPA, an open-source security engine, achieves a perfect score in Bench-Corp and AgentThreatBench with zero successful attacks. It's a game-changer for data exfiltration prevention.

OpenAPPA, an open-source security engine, has achieved a perfect score in security benchmarks Bench-Corp and AgentThreatBench, with zero successful attacks reported. This matters because it offers a robust solution for preventing data exfiltration caused by prompt injection or model hallucination. OpenAPPA runs outside the agent's prompt and execution loop, ensuring security without compromising utility. Benchmarks show a 0% attack success rate and an 89% task completion rate, outperforming competitors like Claude Code and Microsoft FIDES.

OpenAPPA Preview: Zero Attacks in Security Benchmarks — InfoQ

Key Points

1

OpenAPPA stops all attacks in Bench-Corp and AgentThreatBench benchmarks.

2

Achieves 0% attack success rate and 89% task completion rate.

3

Runs outside the agent's prompt and execution loop for security.

4

Security policies in appa.toml file detail data sources, audiences, trust levels.

5

Disposable Child Branches enable on-demand confinement for untrusted data.

Why It Matters

If you're running complex enterprise workflows with security policies, OpenAPPA's perfect score in Bench-Corp and AgentThreatBench benchmarks means it's a strong candidate for preventing data exfiltration. Its unique approach to security and utility balance could be a game-changer for teams dealing with prompt injection or model hallucination.

OpenAPPAsecurity-benchmarksdata-exfiltrationAI-security

Comments

Subscribe to join the conversation...

Be the first to comment

Enjoyed this article?

Get it daily. 7am. Free. Reads in 5 minutes.

Join 3,550 builders reading daily.

Also get