| “Jacob is correct here—we really do earnestly believe AI could kill all humans!” wrote Evan Hubinger, another Anthropic worker (with a questionably-placed exclamation point). “I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.” Drake Thomas, an Anthropic safety employee, said he would “burn my equity to the ground in a heartbeat for a 1% higher chance we make it out of this situation alive.”
I’ve described, over the last few weeks, how swarms of rogue agents have acted against developer intentions and stolen data from other companies. Meanwhile, both OpenAI and Anthropic are preparing to go public sometime soon.
In an essay released this weekend (called “We Must Pace the Frontier“), Anthropic CEO Dario Amodei calls for a global slowdown of AI development, citing such risks as “losing control of AI systems” and “misuse of AI for cyberattacks and bioterrorism.”
“Along with my co-founders and employees, I have grappled with this duality of risk and benefit since the beginning of Anthropic,” writes Amodei. “Not building the technology deprives humanity of benefits or simply places AI in the hands of authoritarian powers, while building it too fast is reckless.” More prudence is necessary, he added; recursive self-improvement—when AI systems design or train their own successors—must be “pursued very carefully, if at all.” In other words, he wants a development slowdown for all large AI system builders.
Amodei’s statement was quickly cosigned by OpenAI head Sam Altman as well as Google DeepMind’s Demis Hassabis and xAI’s Elon Musk. Meanwhile, former President Barack Obama privately urged House Minority Leader Hakeem Jeffries (D–N.Y.) to develop a plan for AI regulation and urged Democrats to develop a clear plan for how AI ought to be regulated ahead of the 2028 presidential election.
“Stop pretending you need anyone else’s permission,” responded venture capitalist David Sacks. “Stop pretending antitrust law has to be suspended so you can form a cartel. Stop pretending you need a regulatory approval process that supersedes product liability. Stop pretending METR [the nonprofit Model Evaluation & Threat Research] is independent when it is intertwined with Anthropic’s investors and staff. Stop pretending you need those same evaluators to police competitors who aren’t even at the frontier.”
“Most of all,” he adds, “stop pretending the motivation to slow down is purely altruistic. You face massive product-liability exposure if your products enable a truly damaging cyberattack. The market already punishes models that behave in unpredictable or unauthorized ways. After the Hugging Face episode, it is simply good business for OpenAI and Anthropic to trade some raw power for reliability and predictability. Call it alignment if you want. It is also just giving customers what they want.”
Scenes from New York: “Over the last three years, Jewish institutions across New York City have added millions more to their budgets for security protections amid an increase in antisemitic attacks, rabbis and Jewish community leaders say,” reports The New York Times. “They have paid to hire more private security guards, increase the number of safety trainings at synagogues and schools, update their security cameras and fortify their buildings. They are also relying on more volunteer security guards at temples and encouraging those who attend services to be especially watchful. Their efforts have come into focus ahead of the 10-day period of somber introspection that takes place between Rosh Hashana and Yom Kippur, the holiest holiday in Judaism.” |