None
FR
AI labs shouldn’t be allowed to grade their own homework
['Andrew Freedman', 'Gillian Hadfield']
Fortune | FORTUNE
The OpenAI and Anthropic incidents last month — and more recently, Meta — showed why AI labs shouldn’t be allowed to grade their own homework.
Most reporting focused on the capabilities these incidents revealed but they also exposed a gap in how frontier AI is overseen.
Rather than asking frontier AI companies to effectively grade their own homework, it would establish licensed Independent Verification Organizations (IVOs) – technical experts outside the AI labs that would evaluate whether companies’ safety frameworks actually keep catastrophic risks within acceptable bounds.
Today, most frontier AI safety expertise resides inside the companies building the models.
The next time a frontier AI model behaves in an unexpected or dangerous way, the public shouldn’t have to hope the company involved decides to disclose it.