Anthropic Seeks to Control AI Safety Through Dominance
Anthropic, valued at nearly $1 trillion, believes AI safety requires it to lead the field and set global standards.
365 articles
AI safety, alignment, and governance — model risk, red-teaming, regulation, and policy. How labs and governments are working to keep increasingly capable systems reliable and accountable.
Anthropic, valued at nearly $1 trillion, believes AI safety requires it to lead the field and set global standards.
Meta has replaced about half of human moderation tasks with AI models in 2025, planning to reach 90% for some content by year-end, according to the Financial Times.
Authors Guild test shows Pangram and Grammarly correctly identified all human-written texts, while Sidekicker and ZeroGPT failed on every sample.
Bristol police created a sprawling crime prediction system using data on 500,000 residents, but some models were abandoned due to poor performance.
A major AI conference in Beijing highlighted the need for global collaboration to manage risks from advanced models, with experts warning of potential chaos if nations fail to work together.
The Trump administration has engaged in multiple calls with Anthropic, led by cofounder Tom Brown, after replacing CEO Dario Amodei in discussions about re-releasing the Claude Fable 5 AI model.
Pangram CEO Max Spero claims language models expose themselves by repeating similar arguments, a key flaw in AI-generated text detection.
OpenAI announced Patch the Planet, a project offering free security consulting to open-source projects, with over 30 projects already participating.
The US banned foreign access to Anthropic's latest models, citing concerns over AI risks, with critics blaming the company's frequent warnings about potential harms.
The Trump administration ordered Anthropic to shut down its Fable 5 and Mythos 5 models, citing national security concerns. The order, issued on June 17, 2026, prompted immediate action from the company.
Signal President Meredith Whittaker warned that AI chatbots like ChatGPT and Claude are not your friends, citing privacy risks in a June 20, 2026 interview.
Eurocommerce, representing major retailers, seeks to exclude AI-generated ads from EU AI Act transparency rules, citing non-deceptive intent.