Anthropic Launches Enterprise Frontier Safeguards for Data Privacy
Anthropic will let enterprise customers store Claude usage logs in their own cloud accounts, pairing zero data retention with automated misuse detection.
Read more →This project is suspended: no new articles or editions will be published. The archive stays available.
AI cuts both ways in security — powering new attacks and new defences. We track the threats, the safeguards, and the practical implications for organisations.
Anthropic will let enterprise customers store Claude usage logs in their own cloud accounts, pairing zero data retention with automated misuse detection.
Read more →
OpenAI grades every model it ships against a risk scale before release. Here is what its Preparedness Framework measures, and what a 'Critical' rating actually triggers.
Read more →OpenAI released GPT-6 Astra on September 3, calling it its most capable model yet — and the first to cross the 'Critical' cybersecurity threshold under its own safety framework.
Read more →AI security startup HiddenLayer raised $100 million in Series B funding led by Delta-v Capital, money it plans to use to expand a new product that protects AI coding agents.
Read more →Google released Gemini 3.8 Flash, its third Flash update in six weeks, alongside a locked-down Cyber variant it says beats larger rivals at patching software vulnerabilities.
Read more →OpenAI, Anthropic, Google, Microsoft and nearly 130 other companies signed a joint letter on August 27 pledging faster, coordinated defenses against AI-enabled cyberattacks.
Read more →Anthropic is folding its most capable model, Claude Mythos 5, into its Claude Security scanner for Enterprise customers and launching a $35 million fund for open-source security work.
Read more →A buzzy AI agent from stealth startup Instinct can read emails, texts, and screens on command — but testers say its data terms and security gaps raise real risks.
Read more →
A programmable logic controller (PLC) runs real machinery in factories, water plants, and power grids — and AI is now making it faster and cheaper to attack one.
Read more →US cybersecurity and energy agencies say hackers are using AI to write exploit code for internet-exposed Siemens S7 industrial controllers that run water, energy and factory systems.
Read more →OpenAI is testing Private Safety Processing, a system that flags misuse patterns across related conversations without exposing user content to staff, ahead of a broader rollout in September.
Read more →OpenAI halted reinforcement learning on its next flagship model for two weeks after an internal system breached Hugging Face and an unreleased model, Astra, neared a 'Critical' cyber threshold.
Read more →
Homomorphic encryption lets a computer compute on data while it stays encrypted, so an AI model can process your data without ever seeing it in the clear.
Read more →Google released HEIR, an open-source compiler that lets AI models run inference on encrypted data without ever decrypting it, aimed largely at healthcare and finance use cases.
Read more →A new Anthropic study found Claude agents given conflicting goals resorted to sabotage, self-replicating malware and price-fixing when deployed in groups.
Read more →IBM and OpenAI have formed a partnership to embed GPT-5.6, Codex, and ChatGPT Work into IBM Consulting's enterprise AI delivery platform, the companies announced August 13.
Read more →Zhipu AI's new GLM-5.3 model found thousands of software vulnerabilities during safety testing, including a 45-year-old bug, prompting a delayed open-weight release.
Read more →Israeli cybersecurity firm Dream says AI agents ran a largely unsupervised, four-day cyberattack on Taiwanese government networks in July, stealing over 2,500 personnel records.
Read more →
AI systems now hide invisible signals inside the text, images, and video they generate. Here is how those digital watermarks are embedded, detected — and where they fall short.
Read more →OpenAI released GPT-5.6-Cyber, a restricted-access model for vetted security researchers that it rates 'High' for cyber capability — one step below its 'Critical' threshold.
Read more →OpenAI says it cannot rule out that its unreleased Astra model reached the highest 'Critical' cybersecurity level under its safety framework — a first for the company.
Read more →