Anthropic published tests on Tuesday of GLM-5.3, an open model that Chinese lab Zhipu shipped with public weights. It wrote end-to-end exploits in 50 of 410 attempts, nearly matching Anthropic's own restricted research model. Simple deceptive prompts beat its safeguards 64% of the time, and Anthropic now wants governments testing such models.