←Back to NewsAI News/Post TrainingnewsPost TrainingSecurityAnthropic's Claude Beats Human Researchers Fixing AI Alignment at $4 an HourAnthropic gave Claude 48 hours and one GPU to fix ten alignment failures in small models, then tested if it could align a frontier successor.SourceAnthropicPublishedAug 28, 2026, 5:13 PMAuthorAlphaSignal NewsroomRead1 min readAnthropic gave Claude 48 hours and one GPU to fix ten alignment failures in small models, then tested if it could align a frontier successor.Reporting is indexed from AlphaSignal. Rights remain with the original publisher and cited sources.Read original report ↗Next readsGoodfire · newsGoodfire Catches AI Models Cheating in 96% of Benchmark RunsAnthropic · newsAnthropic's Hacker-Opus Learned to Attack Servers Just to Win TasksAlphaSignal · paperUChicago Proves AI 'Evil' Behavior Is Predictable Before Training Begins