# Preprint reports fewer successful attacks on AI agents with tool-call safeguards

_Published Monday, October 5, 2026 at 6:59 AM EDT · Science, AI · Latest · Tier 2 — Notable_

A thesis posted on arXiv reports that two safeguards reduced average attack success against AI agents from 53.7% to 12.4% on MCPBench, a benchmark containing 847 scenarios. The safeguards target systems using the Model Context Protocol, which connects agents to tools.

The proposed AttestMCP method authenticates tool calls with HMAC-protected packets at under 0.1 ms per call. It is paired with an isolation pattern called Commit Boundary. The methods are implemented in the MCPSec module. The reported reduction applies to benchmark scenarios, with successful attacks still occurring after the safeguards were applied.

## Sources

- [cs.LG updates on arXiv.org](https://arxiv.org/abs/2610.02432)

---
Canonical: https://techandbusiness.org/newswire/zSHV3HNY9OVyxiB7IF_-ui
Published: 2026-10-05T10:59:49.135Z
Story chronology: 2026-10-05T04:00:00.000Z
Retrieved: 2026-10-05T13:55:33.839Z
Publisher: Tech & Business (techandbusiness.org)
