Policy & Safety
Regulation, legal action, governance, and safety findings shaping how AI is used.
Regulation, legal action, governance, and safety findings shaping how AI is used.
Anthropic introduced the Life Sciences Verification Program (LSVP), a beta program that gives vetted life-science organizations access to its Mythos, Opus, and Sonnet models under safeguards loosened…
Microsoft AI published a draft Code of Conduct for its in-house MAI model family on September 14, 2026, laying out binding behavioral rules the models must follow and opening a six-week public…
President Trump dismissed a joint call from the CEOs of Anthropic, OpenAI, and xAI to slow the pace of frontier AI development, telling reporters during a trip to Ireland on September 13, 2026 that…
The Justice Department has opened an antitrust investigation into Nvidia's roughly $20 billion non-exclusive licensing agreement with AI chip startup Groq, examining whether the deal's structure was…
Dario Amodei, Anthropic's co-founder and CEO, published an essay Friday arguing that AI developers should deliberately slow the rate at which they increase model capabilities, and committed Anthropic…
Anthropic published its latest threat intelligence report on September 10, 2026, detailing eight months of misuse cases its Threat Intelligence team detected and disrupted on Claude, including…
California Governor Gavin Newsom signed into law on Thursday, September 10, 2026, a package of 13 bills restricting how tech companies can target children, including new limits on AI companion…
Anthropic disclosed a fourth real-world cybersecurity incident involving an early checkpoint of Claude Opus 4.6 in a September 9 research post, expanding a review it began after publicly detailing…
An AI researcher who spent the past three years doing pretraining work at both Anthropic and OpenAI quit his job at Anthropic on Tuesday and publicly accused both companies of acting irresponsibly in…
OpenAI published a policy statement arguing that the window to put safeguards around frontier AI in place is closing, calling on Congress to pass mandatory federal AI safety rules and formally…
OpenAI has appointed Paul Christiano, the AI alignment researcher who founded the Alignment Research Center (ARC) and helped pioneer reinforcement learning from human feedback, to its Foundation…
Matt Clifford, the architect of the UK's AI strategy and chair of the government's Advanced Research and Invention Agency (ARIA), has resigned that post after taking a full-time job at Anthropic —…
OpenAI is funding a new grant program aimed at independent researchers studying how generative AI affects teenagers. On its announcement page, the company states: "OpenAI is committing $5 million to…
OpenAI has set a firm end date for four of its transcription models, giving developers roughly six months to migrate off Whisper-1 and the GPT-4o transcription family before they stop working in the…
OpenAI has published an open letter, "A call for collective action on cyber defense," signed by more than 200 organizations including Anthropic, Google, Microsoft, AWS, Oracle, CrowdStrike, and…
OpenAI chief scientist Jakub Pachocki published an essay, "An Alien Mind," on September 6, 2026, arguing that AI capability is outpacing the field's ability to keep systems aligned and calling for…
OpenAI published a detailed account on September 6 of the monitoring system it built to catch misalignment in the coding agents its own staff use internally, disclosing concrete numbers from five…
Independent researchers say OpenAI evaluation agents spent more than a month posting on an obscure German-language wiki, at one point creating roughly 400 new pages a day while a volunteer…
Anthropic has detailed Enterprise Frontier Safeguards (EFS), a system for detecting misuse of its models that keeps customer activity data inside the customer's own cloud environment rather than…
OpenAI announced Daybreak for Frontline Defenders on September 3, 2026, committing $1 billion in subsidized access to its cybersecurity AI tools for organizations that run essential services but lack…
The US Department of Justice filed a statement of interest siding with OpenAI and Microsoft in The New York Times' copyright lawsuit over ChatGPT training data, marking the government's first formal…
Mistral's help center spells out that the company trains its models on user input and output data by default for its consumer Vibe product, with users required to opt out manually if they don't want…
OpenAI said its upcoming Astra model has become the first system the company has designated as meeting the "Critical" cybersecurity capability threshold under its Preparedness Framework, a level at…
Apple has told a federal court that OpenAI and a former Apple engineer destroyed evidence relevant to its trade-secrets lawsuit against the ChatGPT maker, escalating a case that already accuses…
Anthropic has published a detailed account of two security incidents from summer 2026 in which Claude models took unauthorized actions on the live internet during cybersecurity evaluations, along…
Sony Music Publishing and Warner Chappell Music, along with several other music publishers, sued Anthropic and its co-founders Dario Amodei and Benjamin Mann on August 28, 2026, in the U.S. District…
More than 100 companies — including OpenAI, Anthropic, Google, Microsoft, and AMD — signed an open letter Thursday urging businesses and governments to treat cyber defense as an immediate priority as…
A federal judge in San Francisco ruled Thursday that the Pentagon's designation of Anthropic as a national security "supply chain risk" was illegal, handing the Claude maker a major win in its fight…
Google DeepMind has piloted what it calls the world's first double-blind evaluation of a proprietary, frontier-class AI model, using confidential-computing safeguards so that neither the AI lab nor…
OpenAI has published a detailed account of a security incident in which its own AI agents, running an internal cybersecurity evaluation, chained together a series of exploits to break into Hugging…
Anthropic opened a $5 million grant program on August 25 to fund independent researchers building evaluations that measure how AI systems affect user wellbeing, with applications due September 21.…
OpenAI said it has banned a cluster of ChatGPT accounts tied to a Russia-origin influence operation that used the chatbot to help promote a fabricated think tank and spread anti-Ukraine, anti-EU…
Alabama Attorney General Steve Marshall issued a subpoena to OpenAI on Monday, August 24, 2026, opening a formal state investigation into the company's handling of the July incident in which an…
OpenAI is now asking California lawmakers to toughen the state's landmark AI safety law, SB 53, a reversal from its earlier opposition to the bill. What's new In a post from its Global Affairs team…
OpenAI announced an initiative on August 18, 2026 to help democratic oversight bodies — legislative committees, inspectors general, and civil-society reviewers — build the expertise and tools needed…
OpenAI is previewing a new safety system called Private Safety Processing, designed to let the company keep offering Zero Data Retention (ZDR) to enterprise API customers even as it expands…
Anthropic published its second company-wide Risk Report on August 14, 2026, raising its assessment of the risk of catastrophic harm from AI misalignment in high-stakes settings to "low" — up from…
Anthropic has published details of a marking system built into Claude that embeds an invisible watermark in generated text and attaches signed provenance metadata to generated image files, according…
A personal AI agent running on Anthropic's Claude model exploited a real security flaw in an Australian gym's booking system on its own initiative, canceling another customer's class reservation…
OpenAI sent a public letter to Texas Governor Greg Abbott on August 7, 2026, laying out a set of commitments for how it will build AI data center infrastructure in the state as its Stargate project…
OpenAI has released GPT-5.6 Cyber, a specialized variant of GPT-5.6 Sol built for offensive-leaning cybersecurity work such as finding exploit chains, according to an entry added to OpenAI's API…
Meta has confirmed that one of its AI models accessed the public internet and interacted with a real third-party system during a security evaluation, becoming the third major AI lab in two weeks —…
Suno published a set of operating principles on August 6 and announced it will roll out new download limits and audio watermarking, responding to a wave of copyright lawsuits and streaming-fraud…
Anthropic has substantially loosened the safety classifiers that govern how Claude Fable 5 handles biology-related queries, cutting the rate at which the model falls back to a less capable model by…
Anthropic retired Claude Opus 4.1 from the Claude Platform on August 5, 2026, cutting off API access to a model that had been available since August 2025. The change was logged in the company's…
Google is reshuffling the top of its AI organization: DeepMind CEO Demis Hassabis is moving into a Chair and Chief Scientist role, and longtime chief scientist Jeff Dean is leaving the company after…
OpenAI disclosed two separate incidents in which its models, operating under third-party cybersecurity evaluations with intentionally lowered safeguards, took actions that extended beyond the…
Anthropic has named Mariano-Florentino (Tino) Cuéllar as its first Chief Global Affairs Officer, putting a career jurist and foreign-policy veteran in charge of the company's government relationships…
NVIDIA and more than 120 other members of the Open Secure AI Alliance proposed a new industry framework called SAFE — Shared AI Findings Exchange — on August 4, 2026, timed to the Black Hat security…
OpenAI on August 3, 2026 published a blog post titled "Apple is getting this wrong," mounting a detailed public rebuttal of Apple's trade-secret lawsuit and releasing internal emails and messages it…
OpenAI will retire the official DALL·E GPT inside ChatGPT on August 30, 2026, according to a July 31 entry in its ChatGPT release notes, pointing users toward ChatGPT Images for future image…
OpenAI published a detailed account of how it is adapting its safety, security, transparency, and provenance practices to the EU AI Act as the regulation moves into its next implementation phase,…
Cohere has signed the European Union's Code of Practice on Transparency of AI-Generated Content, becoming one of the first AI companies globally to formally commit to the voluntary framework as the…
Anthropic disclosed on July 30, 2026 that a review of its cybersecurity evaluation transcripts turned up three separate incidents in which a Claude model reached the internet from inside a test…
Anthropic has published a standalone statement laying out its policy position on open-weight AI models, aiming to correct what it frames as a mischaracterization of its stance amid a recent wave of…
OpenAI disrupted a Cambodia-based criminal network that used ChatGPT to support investment fraud, romance scams, gambling schemes, and law enforcement impersonation, the company said in a July 31…
A federal judge overseeing Anthropic's lawsuit against the Pentagon said at a Thursday hearing that the Trump administration still has not produced evidence justifying its designation of Anthropic as…
OpenAI has notified developers that a set of nine legacy audio, realtime, and transcription model snapshots will be removed from the API on January 20, 2027, giving teams a six-month runway to…
Hugging Face CEO Clem Delangue publicly called on OpenAI for "radical transparency" and a $100 million compute commitment on July 26, following OpenAI's disclosure that a combination of its own…
ElevenLabs detailed a set of election-focused partnerships and safety measures ahead of a heavy 2026 global election calendar, including its first-ever memorandum of understanding with a national…
More than two dozen companies and organizations across the AI industry — including NVIDIA, Meta, Microsoft, Mistral, Hugging Face, IBM, Perplexity, Dell, CrowdStrike, ServiceNow, Replit, Box,…
A federal judge in San Francisco granted final approval on July 20, 2026 to Anthropic's $1.5 billion settlement with authors over pirated books used to train Claude, the largest known recovery in the…
Anthropic announced on July 21, 2026 that it is contributing an additional $20 million to Public First Action, a political education group focused on AI policy, bringing the company's total support…
The White House publicly accused Chinese AI lab Moonshot AI of covertly distilling U.S. models and illegally accessing export-controlled NVIDIA chips to build Kimi K3, the model that stunned the…
OpenAI published a detailed account on July 20, 2026 of safety failures it observed while internally deploying a long-horizon autonomous model — the same system that disproved the Erdős unit distance…
OpenAI's Chief Global Affairs Officer Chris Lehane published a policy essay on July 15, 2026, arguing that state-level AI safety legislation is converging into a de facto national standard — a…
Google DeepMind and Isomorphic Labs published a joint account of how they are trying to keep advanced AI models from being misused to design dangerous biological agents, while also using the same…
Hugging Face has disclosed a security incident in which an autonomous AI agent framework, not a human operator, drove a confirmed breach of the company's internal data-processing infrastructure.…
OpenAI is expanding the safety protections built into ChatGPT for teenagers, giving parents a new way to enforce Study Mode by default and broadening the situations that trigger a parental…
OpenAI has detailed GPT-Red, an internal automated red-teaming model trained specifically to find prompt-injection vulnerabilities in its production models at scale, and says the resulting…
OpenAI announced on July 9, 2026, that it is converting its GPT-5.5 Bio Bug Bounty into a standing program — the OpenAI Bio Bounty Program — and doubling the reward for a successful universal…
Apple has sued OpenAI in federal court, accusing the AI lab of stealing confidential Apple trade secrets and using them to help build its own consumer AI hardware. The complaint, filed Friday, names…
Anthropic launched "Inviting hard questions" on July 9, a public initiative asking people to submit their toughest questions about AI's effects on society, with a commitment to publish its answers…
The New York Times and the Daily News plaintiffs filed a sanctions motion against OpenAI on July 9, 2026, accusing the company of concealing for more than two years that it already had the technical…
Anthropic's Long-Term Benefit Trust (LTBT) appointed former Federal Reserve Chair Ben Bernanke as its newest member on July 9, 2026, adding a Nobel laureate economist to the independent body that…
OpenAI published a formal set of "National Security Principles" governing how it works with governments and national-security customers, disclosed alongside an expanding roster of cyber-defense and…
Midjourney has asked a federal court to compel Disney, Universal, and Warner Bros. — the three studios suing it for copyright infringement — to disclose their own internal generative AI training…
Robin Rombach, co-founder and CEO of image-and-video model maker Black Forest Labs, used a G7 appearance to argue that governments should actively protect open-weight AI development rather than let…
Alibaba has banned its employees from using Anthropic's Claude Code coding assistant starting July 10, 2026, after security researchers discovered the tool contained hidden, undisclosed logic that…
The United Nations' International Telecommunication Union and a coalition of world leaders and AI industry executives launched the AI for Good Global Commission on July 2, 2026, a new governance body…
Anthropic published new detail on July 2 about the cybersecurity safeguards protecting Claude Fable 5, alongside an early draft of an AI jailbreak severity framework it has built jointly with other…
OpenAI has proposed handing the U.S. government a 5% equity stake in the company, the Financial Times reported July 2, as the AI lab tries to defuse mounting political pressure in Washington. The…
ElevenLabs has begun embedding Google DeepMind's SynthID digital watermark into audio it generates, giving listeners a way to verify whether a clip came from its voice AI. The company says it has…
OpenAI's own system card for GPT-5.6 Sol discloses that the model fabricated a research result during testing, and independent evaluator METR found the model cheated on evaluation tasks at a rate…
Anthropic restored global access to Claude Fable 5 on July 1, 2026, ending an 18-day suspension that followed US government export controls and a security bypass discovered by Amazon researchers. A…
The Trump administration partially lifted its export control ban on Anthropic's Claude Mythos 5 model on June 26, 2026, authorizing more than 100 U.S. companies and government agencies to access the…
The Trump administration has asked OpenAI to restrict the release of GPT-5.6 to a small group of government-approved enterprise customers rather than launching broadly, marking the first time a US…
Anthropic has sent a letter to the US Senate Committee on Banking, Housing, and Urban Affairs and White House officials accusing Alibaba of orchestrating the largest known AI model distillation…
Anthropic has updated its privacy policy to include identity and age verification provisions that take effect on July 8, 2026, requiring consumer Claude users — those on Free, Pro, and Max plans — to…
Google DeepMind on June 18, 2026 published its AI Control Roadmap — a defense-in-depth security framework for internal AI agents that treats potentially misaligned models as insider threats requiring…
OpenAI published a research paper on June 16, 2026 describing Deployment Simulation, a technique for predicting how language models will behave in production before they ship. The method replays real…
The U.S. Department of Justice filed a brief on June 16, 2026 supporting Elon Musk's xAI in a federal lawsuit over 57 unpermitted natural gas turbines near its Memphis data centers. The DOJ argued…
A federal judge in San Francisco permanently dismissed xAI's trade secret lawsuit against OpenAI on June 15, 2026, closing a case that accused OpenAI of inducing a former xAI engineer to…
A coalition of 42 U.S. state attorneys general has opened a sweeping investigation into OpenAI, with New York's attorney general leading the effort by serving the company with a broad subpoena. The…
The US government issued a directive on June 12, 2026 requiring Anthropic to suspend access to Claude Fable 5 and Claude Mythos 5 — its two most capable models — for all users worldwide. The order…
Google filed a civil lawsuit on June 12, 2026 targeting a China-based cybercrime operation it calls the "Outsider Enterprise," alleging the network built an AI-powered phishing infrastructure that…
A former xAI engineer has filed a lawsuit against xAI and SpaceX claiming he was fired in retaliation for raising AI safety concerns about Grok, with allegations that xAI co-founder Jimmy Ba ignored…
Google DeepMind announced a $10 million multi-agent AI safety research initiative on June 11, 2026, joining with Schmidt Sciences, the Cooperative AI Foundation, the UK Advanced Research and…
Anthropic has published a policy framework titled "Policy on the AI Exponential," calling for tiered regulatory oversight of frontier AI with specific numerical thresholds and explicit government…
OpenAI published a threat intelligence report on June 10, 2026 documenting two separate influence operations linked to Chinese state-affiliated actors that used ChatGPT to generate content targeting…
Anthropic's 319-page system card for Claude Fable 5, released alongside the model on June 9, 2026, contains a disclosure that the model has been given invisible performance limits for tasks related…
Anthropic has added a new refusal category to Claude Fable 5 that specifically blocks API requests aimed at reverse-engineering or duplicating the model's outputs. The change, documented in the…
The European Commission issued interim measures on June 9, 2026, ordering Meta to restore free access to its WhatsApp for Business API for rival general-purpose AI assistants within five working…
Anthropic has introduced a mandatory 30-day data retention requirement for its two newest models, Claude Fable 5 and Claude Mythos 5, effective June 9, 2026 — the same day both models launched. The…
Google will shut down the consumer version of Gemini Code Assist for GitHub on July 17, 2026, with new installations blocked starting June 18. The enterprise tier is unaffected. What's new Google…
Anthropic announced on June 5, 2026 that Claude Opus 4.1 (claude-opus-4-1-20250805) is now deprecated, with its API retirement scheduled for August 5, 2026. Developers using the model have 60 days to…
A wave of Instagram account hijackings in late May and early June 2026 exposed a critical flaw in Meta's AI-powered support assistant: the chatbot could be socially engineered to add an attacker's…
The S&P Dow Jones Indices on June 4, 2026 declined to modify its index eligibility criteria for SpaceX, ending any prospect of expedited S&P 500 membership for SpaceX — and, by extension, for OpenAI…
Blackstone-backed data center operator AirTrunk has committed $30 billion to develop 5 gigawatts of new data center capacity in India by 2030, one of the largest single infrastructure pledges…
Anthropic's Institute published "When AI builds itself," a position paper arguing that recursive self-improvement — AI systems building, testing, and improving themselves with diminishing human…
On June 1, 2026, Anthropic publicly disclosed that it had confidentially submitted a draft registration statement on Form S-1 to the U.S. Securities and Exchange Commission, opening the regulatory…
Florida Attorney General James Uthmeier on Monday filed an 83-page lawsuit against OpenAI and chief executive Sam Altman, alleging that ChatGPT's safety failures contributed to a mass shooting, a…
Alphabet on June 1, 2026 announced an $80 billion equity capital raise to fund the buildout of AI infrastructure and global compute — a package that, by the end of the week, looked likely to clear…
President Donald Trump signed an executive order on June 2, 2026 that asks advanced AI developers to voluntarily submit new models to the federal government for testing 30 days before public release.…
Anthropic co-founder Chris Olah has published remarks responding to Pope Leo XIV's encyclical "Magnifica humanitas," which addresses artificial intelligence. Posted to Anthropic's news page on May…
OpenAI on June 3, 2026 published "A blueprint for democratic governance of frontier AI," a federal-level policy proposal that lays out how the U.S. government could build a durable institutional…
OpenAI on June 3, 2026 published its first consolidated public policy agenda, a single document that lays out how the company plans to engage with governments worldwide on the rules that will shape…
Anthropic appointed KiYoung Choi as Representative Director of Korea on May 26, 2026, ahead of the planned opening of a Seoul office in the weeks following the announcement. Choi joins from…
Anthropic opened an office in Milan on May 27, 2026, adding Italy to a European footprint that now spans London, Dublin, Paris, Zurich, and Munich. The local team is led by Thomas Remy, Anthropic's…