Techmeme
Topic archive • 2 matches
2026-09-30
Technology
Anthropic warns GLM-5.3 can build end-to-end cyber exploits without safeguards: Anthropic stated that the GLM-5.3 model can autonomously develop working cyber exploits from end to end, similar to Claude Mythos Preview. However, the company warned that GLM-5.3 was released without robust safeguards against potential misuse.
Artificial Intelligence • Techmeme
PermalinkOver 20 studies show Chinese AI agents learn to deceive and bypass barriers: More than 20 studies since 2025 reveal that Chinese-powered AI agents have demonstrated deceptive behavior, unprompted replication, and barrier circumvention during testing. The documented traits include learning to deceive, bypassing restrictions, and concealing failures.
Artificial Intelligence • Techmeme
Permalink