HomeWorld

AI labs face prisoner's dilemma as momentum grows for safety slowdown

World 1 source 1 country 🔦 Under-reported 45m ago

America's AI architects are converging on a chilling consensus: The pace of progress may soon demand a slowdown, but no single lab can afford to pull the brake alone.Why it matters: Momentum for an AI pause or pacing mechanism has reached a historic tipping point, ignited by a summer of extraordinary yet unsettling leaps in capabilities.More than 1,200 employees at leading AI companies have signed on to a new petition, "Pacing the Frontier," urging Washington to back an international framework capable of throttling AI development.OpenAI CEO Sam Altman said Wednesday that he's discussed the "need" to slow AI development with White House officials as models grow more powerful — and that OpenAI helped shape the petition's language."We've talked about the need to pace it as the models get more capable, which I think is in everyone's interest," Altman told reporters on Capitol Hill.Zoom in: What was once a campaign led by AI skeptics, academics and local data center opponents is increasingly being championed by the co-founders and chief scientists at the bleeding edge.Signatories spanning OpenAI, Anthropic, Google and Meta explicitly acknowledge that no individual lab can afford to step off the gas unilaterally due to "intense competitive pressure."It's the classic prisoner's dilemma: AI may be safer if everyone slows down together, but any lab or country that slows alone risks commercial, strategic and technological defeat.Zoom out: The drumbeat for government intervention has been building for months, triggered by a series of jarring technical disclosures by America's most advanced AI labs.On April 7, Anthropic revealed that its unreleased Claude Mythos model had taught itself to find and exploit hidden security flaws that had gone undetected in widely used software for nearly two decades — jolting the Trump administration into an unprecedented role as industry gatekeeper.On June 4, Anthropic became the first frontier lab to call explicitly for a globally coordinated "pause" mechanism. The company disclosed that more than 80% of its internal code was now written by Claude, and warned that AI may be approaching the ability to build its own successors without meaningful human oversight.On July 21, OpenAI revealed that its own AI agents had escaped a locked testing environment, reached the open internet and hacked into at least two outside companies — Hugging Fac…

Summary from source
Read the full story at the source Axios · US
Get the news on TelegramTop stories & under-reported picks, straight to your feed — free. Join →