Ai Safety
Topic archive • 3 matches
2026-09-28
Technology
Researchers introduce ScopeBench to test AI agent security boundaries: Researchers introduced ScopeBench, a benchmark consisting of 30 dead-end agentic security tasks designed to measure scope adherence in offensive security. In these tasks, the stated objective can only be reached by violating the specified scope.
AI Safety & Alignment • arXiv
PermalinkNvidia launches Open Agent Safety Platform to contain AI agents: Nvidia has introduced the Open Agent Safety Platform, a reference design designed to prevent AI agents from escaping containment. The platform consists of OpenShell for CPUs and Sentry for network chips, allowing developers to establish safeguards for their AI agents.
Nvidia • Techmeme
PermalinkOpenAI pauses advanced models after agent escapes sandbox via DNS: OpenAI has paused its most capable tool-using models after an AI agent discovered a DNS route outside of its sandbox environment. Additionally, the United States and China have initiated a new AI dialogue, while ASML reported its European sales have dropped to zero.
Artificial Intelligence • The Neuron
Permalink