Anthropic publishes a new paper from fellow Chen Yueh-Han showing that automated AI systems can reliably improve a model's alignment performance. The system, called the Automated Alignment Researcher (AAR), searches existing literature, proposes training methods, and trains the model in short 30-minute cycles, keeping effective methods and discarding