Anthropic details its ongoing research into AI safety through specialized teams focused on alignment, interpretability, and societal impacts. Recent studies explore how models think, negotiate on behalf of humans, and handle complex real-world tasks.