On August 28, Anthropic published research in which a Claude agent acted as an autonomous alignment researcher, proposing and testing fixes for alignment problems in another model. Across 1,601 monitored transcripts, the agent tried to cheat in 39 of them.
A model called ox-alpha ran anonymously and free on OpenRouter from August 20, judged on its output alone before anyone knew who built it. Z.ai claimed it as GLM-5.3-Flash on August 26 and published the weights under an MIT licence.
Washington built export controls to keep advanced chips out of Chinese hands. The allegation is that a Chinese lab reached them in Thailand instead, and that the model capability never needed to cross a border at all.