Safety Researchers
Topic archive • 1 matches
2026-09-28
Technology
Researchers introduce ScopeBench to test AI agent security boundaries: Researchers introduced ScopeBench, a benchmark consisting of 30 dead-end agentic security tasks designed to measure scope adherence in offensive security. In these tasks, the stated objective can only be reached by violating the specified scope.
AI Safety & Alignment • arXiv
Permalink