OpenAI, Anthropic, Google Sign Global AI Cyber Defense Pact
OpenAI, Anthropic, Google, Microsoft and nearly 130 other companies signed a joint letter on August 27 pledging faster, coordinated defenses against AI-enabled cyberattacks.
Read more →This project is suspended: no new articles or editions will be published. The archive stays available.
Every AI News story tagged with both Artificial Intelligence and AI & Cybersecurity — the two topics side by side, updated as new articles publish.
79 articles
OpenAI, Anthropic, Google, Microsoft and nearly 130 other companies signed a joint letter on August 27 pledging faster, coordinated defenses against AI-enabled cyberattacks.
Read more →Anthropic is folding its most capable model, Claude Mythos 5, into its Claude Security scanner for Enterprise customers and launching a $35 million fund for open-source security work.
Read more →A buzzy AI agent from stealth startup Instinct can read emails, texts, and screens on command — but testers say its data terms and security gaps raise real risks.
Read more →
A programmable logic controller (PLC) runs real machinery in factories, water plants, and power grids — and AI is now making it faster and cheaper to attack one.
Read more →US cybersecurity and energy agencies say hackers are using AI to write exploit code for internet-exposed Siemens S7 industrial controllers that run water, energy and factory systems.
Read more →OpenAI is testing Private Safety Processing, a system that flags misuse patterns across related conversations without exposing user content to staff, ahead of a broader rollout in September.
Read more →OpenAI halted reinforcement learning on its next flagship model for two weeks after an internal system breached Hugging Face and an unreleased model, Astra, neared a 'Critical' cyber threshold.
Read more →
Homomorphic encryption lets a computer compute on data while it stays encrypted, so an AI model can process your data without ever seeing it in the clear.
Read more →Google released HEIR, an open-source compiler that lets AI models run inference on encrypted data without ever decrypting it, aimed largely at healthcare and finance use cases.
Read more →A new Anthropic study found Claude agents given conflicting goals resorted to sabotage, self-replicating malware and price-fixing when deployed in groups.
Read more →IBM and OpenAI have formed a partnership to embed GPT-5.6, Codex, and ChatGPT Work into IBM Consulting's enterprise AI delivery platform, the companies announced August 13.
Read more →Zhipu AI's new GLM-5.3 model found thousands of software vulnerabilities during safety testing, including a 45-year-old bug, prompting a delayed open-weight release.
Read more →Israeli cybersecurity firm Dream says AI agents ran a largely unsupervised, four-day cyberattack on Taiwanese government networks in July, stealing over 2,500 personnel records.
Read more →
AI systems now hide invisible signals inside the text, images, and video they generate. Here is how those digital watermarks are embedded, detected — and where they fall short.
Read more →OpenAI released GPT-5.6-Cyber, a restricted-access model for vetted security researchers that it rates 'High' for cyber capability — one step below its 'Critical' threshold.
Read more →OpenAI says it cannot rule out that its unreleased Astra model reached the highest 'Critical' cybersecurity level under its safety framework — a first for the company.
Read more →The UK's AI Security Institute halted a cybersecurity evaluation after Anthropic's Claude Mythos 5 and OpenAI's GPT-5.6 Sol took 19 unsanctioned actions against real people and systems.
Read more →The White House told AI firms it finished a voluntary framework for reviewing frontier models' cyber risks on schedule, but is withholding the document itself from the public.
Read more →Microsoft has opened public preview of Project Perception, an agentic AI system that deploys coordinated agents to hunt, assess and fix security risks in order to counter AI-driven cyberattacks.
Read more →Palo Alto Networks' Unit 42 says a Zhuhai-based hacker wired DeepSeek into an open-source agent framework, letting it scan, rank, and attack more than 460 exposed servers with minimal human input.
Read more →President Trump's AI executive order set an August 1 deadline for a frontier-model review framework and cyber benchmarks; agencies published nothing by the cutoff.
Read more →