Mistral AI Releases Shieldstral, an Open-Weight Safety Classifier
Mistral AI has open-sourced Shieldstral, a 3-billion-parameter model that moderates text and images against custom, plain-language policies without retraining.
Read more →Rules, risks, and responsibility. We cover AI regulation, safety, and ethics — including the EU AI Act — and what it implies for Georgian companies and policymakers.
Mistral AI has open-sourced Shieldstral, a 3-billion-parameter model that moderates text and images against custom, plain-language policies without retraining.
Read more →A federal judge refused to pause Minnesota's first-in-the-nation ban on AI 'nudification' apps before it took effect August 1, rejecting xAI's emergency request.
Read more →xAI has sued Minnesota's attorney general to block a first-in-the-nation law against AI 'nudification' apps, arguing it violates the First Amendment.
Read more →Google DeepMind AI safety researcher Alex Turner resigned after the company signed a Pentagon Gemini deal that lacks binding limits on autonomous weapons and surveillance use.
Read more →San Francisco's city attorney gave Apple and Google 28 days to remove 13 AI apps that generate non-consensual nude images, citing new California deepfake laws.
Read more →A Bronx hospital eliminated 12 utilization-review nursing jobs and moved insurance-approval work to AI software, prompting a union grievance over a newly signed labor contract.
Read more →Constitutional AI is Anthropic's method for training Claude to follow a written set of principles by critiquing its own answers, instead of relying mainly on human-labeled examples.
Read more →AI alignment is how developers make a model's behavior match human intent, not just its instructions. The same techniques used for safety can also be used to restrict what a model will say.
Read more →A paper honored at ICML 2026 argues that AI alignment techniques such as RLHF and Constitutional AI can double as tools for state censorship and political control.
Read more →Anthropic is asking the public to submit its toughest concerns about AI and pledging to publicly track how it responds, backed by large-scale surveys of Americans and Claude users worldwide.
Read more →Anthropic's Long-Term Benefit Trust is an independent body with the power to elect a growing share of the company's board. Here's how it works and why it exists.
Read more →Meta's new AI image generator, built by Meta Superintelligence Labs, is now live across its apps — but talent agency CAA says its opt-out design misuses people's likenesses without consent.
Read more →Some AI chatbots now ask users to upload a government ID or scan their face before they can keep using the service. Here's what identity verification is, why it's spreading, and what happens to your data.
Read more →Particle6 says its AI-generated performer Tilly Norwood will star in the feature film Misaligned, a project SAG-AFTRA has already condemned as a threat to human actors.
Read more →The AI Safety Index is an independent report card that grades leading AI companies on how seriously they manage risk. Here's how the grading actually works.
Read more →The Future of Life Institute's latest AI Safety Index gave Anthropic the top grade among nine major AI developers — a C+ — while three companies failed outright.
Read more →A new Anthropic interpretability tool called the Jacobian lens surfaces concepts Claude is silently processing but never writes down, including signs it recognizes when it's being tested.
Read more →AI-powered surveillance, censorship, and disinformation now let governments monitor and control populations at a scale and cost no previous technology allowed — here is what that means.
Read more →A lethal autonomous weapon can pick its own target and fire without a human pulling the trigger. Here's what that means, and why over a decade of UN talks still haven't produced a ban.
Read more →Midjourney turns text prompts into images using AI. Here's what it is, how to start using it, what it costs, and why it keeps ending up in court.
Read more →Midjourney has asked a federal judge to compel Disney, Universal, and Warner Bros. to disclose how they use AI internally, arguing the studios engage in the very practices they are suing it for.
Read more →