Ex-OpenAI Researchers Grade The Labs
22AUG
A new standards nonprofit checked whether labs can keep control of their own AI. Anthropic and OpenAI tied at C+. Google got a D+, xAI a D-, Meta an F.
Guidelight was founded by Steven Adler and Page Hedley, both former OpenAI safety researchers. They scored six practices: logging, monitoring, gated actions, circuit breaking, outside review, containment.
The headline finding is the flat ceiling. On a 0 to 5 scale, nobody scored above 3 on anything. Anthropic scored 0 on having a containment plan.
Labs are best at watching and worst at stopping. Guidelight says control systems today could be switched off by a misbehaving model, or simply outpaced by one.