Anthropic and OpenAI Propose Embedded AI Safety Evaluators: What Changed and What It Costs
Leading AI labs Anthropic and OpenAI have proposed embedding independent safety evaluators into their model development pipelines, but experts question whether these assessors can truly be impartial. Over 100 AI researchers have signed a letter demanding more rigorous oversight.