From the former Policy Frontiers Team Lead at OpenAI: She worked on testing AIs for dangerous capabilities, and warns we don’t really know how to do that. "The evaluations can't be comprehensive enough… the models can get smart enough to hide their capabilities."
Rosie Campbell highlights competitive and cultural pressures she experienced inside a frontier AI company, and discusses her hopes and misgivings about pacing the frontier.
"I assumed there would be a whole load of good arguments for why this isn't really something we need to worry about. There'd be answers to all of these concerns. And the more I read into it, the more I was kind of terrified at the lack of rebuttals to some of these arguments."
"Over time, I feel like the organization changed a lot. The incentives changed. There was a lot more pressure to move fast to build commercial products. And a lot of the people who had been very motivated by these big picture safety concerns started leaving."
"We're at a point where we don't understand what's going on with the systems. The rate of progress is kind of insane, and we need to make sure that our ability - to understand, control, deploy these systems safely - that keeps up with the capabilities of the systems themselves."