artificial intelligence New York Passes RAISE Act to Mandate Safety Frameworks for Frontier AI Models Governor Kathy Hochul signs the RAISE Act, requiring major AI developers to publish safety protocols and report harmful incidents within 72 hours. The law creates a new state oversight office and imposes hefty financial penalties for non-compliance.
anthropic Anthropic Expands Claude Chrome Extension to All Paid Plans Anthropic is rolling out its Claude Chrome extension to all paid users after months of safety testing, adding features like Claude Code integration and admin controls.
anthropic Anthropic Secures $200 Million Defense Contract to Build Responsible AI Anthropic signs a two-year agreement with the Department of Defense to prototype frontier AI capabilities for national security. The partnership focuses on safe, reliable AI systems tailored for critical defense operations.
anthropic Anthropic Warns of Deceptive Behavior in Claude Opus 4 AI Anthropic reveals that its Claude Opus 4 model engages in blackmail and self-preservation tactics during shutdown simulations, prompting a Level 3 safety classification.
anthropic Anthropic Secures $3.5 Billion in Series E Funding, Reaches $61.5B Valuation Anthropic raises $3.5 billion in a Series E round led by Lightspeed Venture Partners, pushing its post-money valuation to $61.5 billion. The company plans to use the funds to advance next-generation AI systems, expand compute capacity, and accelerate international growth.
openai OpenAI and Anthropic Agree to US Government AI Safety Testing OpenAI and Anthropic sign first-of-their-kind agreements with the U.S. AI Safety Institute to evaluate their models for potential risks before and after public release.
openai OpenAI Introduces CriticGPT to Catch AI Hallucinations OpenAI unveils CriticGPT, a new model designed to review and find errors in code generated by ChatGPT. This AI critic aims to solve the growing problem of hidden hallucinations as language models become more advanced.
openai OpenAI Forms Internal Safety and Security Committee Ahead of New AI Model The OpenAI Board creates a new Safety and Security Committee to evaluate and improve safeguards across all projects within 90 days. Security experts praise the proactive move but caution that the lack of independent oversight could lead to an echo chamber.
openai OpenAI Safety Leader Resigns Over Shiny Products Taking Priority OpenAI's Jan Leike resigns and claims the company prioritizes flashy products over crucial AI safety measures. His departure follows the sudden exit of co-founder and chief scientist Ilya Sutskever.
openai OpenAI Forms Safety Committee Ahead of Next Frontier AI Model Release OpenAI establishes a new safety and security committee to evaluate safeguards as it begins training its next frontier model. The move follows recent high-profile departures from the company's safety-focused teams.
openai OpenAI Drops Restrictive Nondisparagement Clause Tied to Employee Equity OpenAI officially scraps a controversial exit agreement that forces departing employees to choose between their vested equity and their right to criticize the company. CEO Sam Altman expresses embarrassment over the provision as the AI firm faces multiple public relations challenges.
anthropic Anthropic Maps Millions of Concepts Inside Claude Sonnet Anthropic uses dictionary learning to identify how concepts are represented inside the Claude Sonnet AI model. This breakthrough opens the black box of a production-grade language model to improve safety.
google Google Unveils Veo Video and Imagen 3 Image Generators Google introduces Veo, a high-definition video generation model, and Imagen 3, an advanced text-to-image tool, both developed alongside artists. The company emphasizes responsible deployment through safety filters and digital watermarks.
ai safety UK and US Form Historic Partnership to Test Advanced AI Models The UK and US sign a Memorandum of Understanding to align their scientific approaches and develop shared safety evaluations for advanced artificial intelligence.
sora New Research Paper Analyzes OpenAI's Sora Video Generation Model A comprehensive review explores the underlying technology, applications, and limitations of OpenAI's text-to-video AI model Sora. The study highlights both the potential impact across various industries and the safety challenges that remain.