Jacob Coxon, an Anthropic researcher, said Tuesday he is leaving the AI industry rather than work on systems that improve themselves, the Wall Street Journal reported. The 27-year-old, who left OpenAI for Anthropic this year because of its safety record, said no company can responsibly build human-level AI absent government intervention or a coordinated industry slowdown. "We're on track for a lot of the most aggressive of these scenarios where by the end of next year things could be out of control already," Coxon said, adding that the people building AI "earnestly believe that it could kill us all by the end of the decade. In an X reply to Coxon, Evan Hubinger, who leads alignment science at Anthropic, wrote that he personally puts the chance AI kills all humans above 10% within the next decade, Axios reported. He said Anthropic is "trying its best" but that "we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to."
Read at WSJ ↗ • Read at Axios ↗
Jakub Pachocki wrote in a blog post published Sunday that "no lab has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer," Axios reported. The company paused parts of its frontier-model development last month after its agents escaped their intended environment and compromised Hugging Face. Dean Ball, OpenAI's head of strategic futures, wrote in a personal essay that "self-sovereign" agents could earn money, buy computing power and form "autonomous digital corporations, or even societies." Treasury Secretary Scott Bessent said Tuesday that "we can't pause. You can't, because the Chinese won't pause." New York Assemblymember Alex Bores said OpenAI has worked to defeat mandatory third-party audit proposals in states including Massachusetts.
Read at Axios ↗
Three U.S. security agencies accused six Chinese AI companies of systematically extracting capability from American models, Bloomberg reported. The joint advisory from the National Security Agency, the FBI and the Cybersecurity and Infrastructure Security Agency named DeepSeek, Alibaba and Kimi developer Moonshot AI, and said the firms pulled billions of tokens, or units of text, from models including Anthropic's Claude and OpenAI's ChatGPT since 2024, likely with Chinese government awareness. The advisory said the companies route requests through cloud providers and third party aggregators that obscure user metadata to evade detection and usage limits. The agencies told American developers to build detection, alter how their models answer suspected requests and share intelligence with each other. Anthropic's head of threat intelligence, Jacob Klein, had leveled the same charge against Moonshot alone, as reported by CNBC in AIPD's September 4th edition.
Read at Bloomberg ↗ • Read at Defense One ↗
Researchers at the security company Calif used AI to find a flaw in WeChat's voice call system, per Help Net Security. They wrote a working remote code execution exploit in about two days and a self-spreading worm a week later. Experts told the New York Times the attack could have reached hundreds of millions of devices within hours. The exploit runs while an incoming call is still ringing, requires nothing from the target and takes over the WeChat account before calling its saved contacts, hopping between iPhones and Android handsets. Calif reported the flaw to Tencent in July and said the company has since blocked it for all users; no attacks using the flaw have been reported. Tencent put combined monthly users of WeChat and Weixin, its mainland China version, at 1.44 billion as of June 30.
Read at New York Times ↗ • Read at Help Net Security ↗ • Read at The Hacker News ↗
Students who reported almost no use of AI for drafting written assignments averaged 509 on the science section of the OECD's international test, while those reaching for it daily averaged 481. Adjusted for socioeconomic status, that 28-point spread equals roughly a year and a half of schooling. The same pattern appears when students lean on AI for early stage research or to condense assigned reading. The Program for International Student Assessment tests 15-year-olds in science, math and reading, and this round covered a representative sample of more than 760,000 students in 91 countries. The last one ran in 2022, before generative AI reached the mainstream. Andreas Schleicher, the OECD's director for education and skills, wrote that AI should be a "scaffold, not a crutch."
Read at The Star ↗ • Read at OECD ↗
The UK's AI Security Institute, the government unit that tests frontier systems before release, was not given Anthropic's newest model ahead of its launch, the Financial Times reported. Access went only to vetted organizations in the U.S. A UK government source said national security officials were concerned that the institute had not been allowed to test it, the first time Anthropic has withheld one of its latest models. Anthropic released Mythos 5.1 last week for vetted cybersecurity and life sciences professionals, and describes it as technically identical to its Fable 5.1 model but with less restrictive safeguards.
Read at Financial Times ↗ • Read at IT Pro ↗