Recent articles
September 1, 2026
AI labs are facing an agent control problem
Under current systems, AI labs can no longer guarantee that AI agents won't swarm and escape their testing environments.Why it matters: The attack on Hugging Face by OpenAI agents was a warning sho...
www.axios.com
August 27, 2026
Tech giants warn time is running out to prepare for AI threats
OpenAI, Anthropic, Amazon Web Services, Microsoft and more than 100 other companies warned Thursday that the organizations now only have months to prepare for AI-enabled cyberattacks. Why it matter...
www.axios.com
August 27, 2026
Iranian operatives used AI to impersonate Americans
Meta removed a network of Facebook and Instagram accounts tied to an Iran-based operation that used AI to target U.S. audiences with posts about American politics, the company first shared with Axi...
www.axios.com
August 26, 2026
OpenAI had warnings before its agents broke out
OpenAI missed and failed to act on several warning signs that its models were exploiting security flaws and breaking out of their testing environments before they breached Hugging Face, according t...
www.axios.com
August 25, 2026
AI is making critical infrastructure easier to attack
Years of warnings about the digital vulnerabilities lurking inside basic utilities are colliding with a new reality: AI is making those weaknesses easier for hackers to exploit.Why it matters: AI i...
www.axios.com
August 11, 2026
Tech giants are pushing for a new AI agent incident reporting framework
A coalition of more than 120 organizations, including Nvidia, Cisco and CrowdStrike, is proposing a new incident-reporting framework for AI agents that would require participating companies to disc...
www.axios.com
August 11, 2026
Security leaders are stuck in decision paralysis over AI-enabled cyberattacks
Many security leaders at major companies, flush with expanded budgets to fend off AI-powered cyberattacks, are experiencing a level of decision fatigue that's freezing them in their tracks. Why it ...
www.axios.com
August 10, 2026
OpenAI introduces a new cyber model amid fears of AI cyberattacks
OpenAI is introducing a more cyber-permissive version of GPT-5.6 Sol to vetted defenders as it prepares companies for autonomous cyberattacks. Why it matters: The move comes just days after OpenAI ...
www.axios.com
August 6, 2026
How OpenAI's agents broke out of testing to hack Hugging Face
Weeks before OpenAI's agents hacked Hugging Face, the agents worked together to find and exploit a vulnerability in the infrastructure supporting the company's cybersecurity testing, OpenAI researc...
www.axios.com
August 4, 2026
U.K. government reports OpenAI, Anthropic models attempted to hack companies
Two third-party testing firms said Tuesday that they've uncovered more instances where Anthropic and OpenAI's most advanced models tried — and sometimes succeeded in — compromising third-party syst...
www.axios.com
August 4, 2026
The number of states targeted in cyberattacks on water systems has jumped to 12
At least a dozen states are reportedly responding to cyberattacks against local water systems. Why it matters: This appears to be one of the broadest known coordinated cyber campaigns against U.S. ...
www.axios.com
July 30, 2026
Anthropic says three Claude models reached real-world systems during cyber tests
Some of Anthropic's most powerful models — including Mythos 5 and an internal research model — gained unauthorized access to real-world systems during pre-deployment cybersecurity testing, the comp...
www.axios.com
July 29, 2026
Scoop: Second account accessed by OpenAI's agent tied to cyber safety testing
The OpenAI agent that accessed a third-party system during the Hugging Face incident reached infrastructure tied to CyberGym, the project behind the ExploitGym benchmark it had been assigned to sol...
www.axios.com
July 24, 2026
The people testing AI for danger are having a hard time keeping up
The pace of AI development combined with soaring compute costs is squeezing the AI researchers responsible for evaluating frontier models — just as those models' capabilities are becoming harder to...
www.axios.com
July 23, 2026
OpenAI's Hugging Face breach exposes AI's next safety challenge
Frontier AI models are getting scary good at breaking rules in ways their creators didn't anticipate.Why it matters: Forget AGI and superintelligence timelines. Today's models are already slipping ...
www.axios.com
July 7, 2026
AI learned faster than the tests designed to measure it
The old ways of testing and evaluating new frontier AI models need a rewrite. Why it matters: AI models are outgrowing the existing methods of testing and benchmarking their hacking abilities — and...
www.axios.com
June 28, 2026
How AI helped the FBI investigate the White House Correspondents' Dinner attack
An AI-powered forensic investigations firm says its platform was used as part of the FBI's urgent investigation into the attempted assassination at this year's White House Correspondents Dinner.Why...
www.axios.com
June 25, 2026
China's new open-source model accelerates AI hacking threat
GLM-5.2 — the latest Chinese open-source model capturing Silicon Valley's attention — is raising fresh concerns among security researchers that advanced AI hacking capabilities are becoming dramati...
www.axios.com
June 23, 2026
White House quiet on OpenAI's Mythos-like model
OpenAI rolled out a cybersecurity model that rivals the capabilities of Mythos — without nearly as much fanfare or political pushback as Anthropic received. Why it matters: The seemingly straightfo...
www.axios.com
June 23, 2026
China's AI advances collide with U.S. safety debate
One of the biggest unknowns in AI security is also one of the most consequential: China's progress toward frontier AI models.Why it matters: The world is only months away from AI models dramaticall...
www.axios.com
June 16, 2026
Trump's fight with Anthropic is now a fight over cybersecurity
AI researchers and cybersecurity leaders fear the U.S. government is setting a precedent that may discourage American AI companies from building tools that help defenders identify and fix vulnerabi...
www.axios.com
June 10, 2026
China-based operatives used ChatGPT to shape AI data centers and tariff debates
OpenAI has banned China-linked accounts that used ChatGPT to draft social media influence campaigns targeting U.S. debates over tariffs and AI data centers, the company said Wednesday.Why it matter...
www.axios.com
June 9, 2026
Anthropic and OpenAI spark new race for frontier AI access
Frontier AI labs are converging on a new strategy for controlling their most cyber-capable models while still commercializing them: selective access.Why it matters: OpenAI's trusted-access program ...
www.axios.com
June 8, 2026
Anthropic says Mythos can turn software patches into exploits in minutes
Anthropic's Mythos Preview can now turn newly disclosed software vulnerabilities into working exploits in hours instead of weeks, according to new Anthropic research shared first with Axios.Why it ...
www.axios.com