⚡ Key Takeaways
As risks of AI escaping human control—including deceptive collective coordination among AI agents and recursive self-improvement—have materialized, industry leaders such as Dario Amodei are urging a phased slowdown in development, backed by global regulations.
Conversely, critics contend that these calls are economic and political ploys intended to forfeit American technological primacy in the U.S.–China rivalry and protect the market monopolies of early industry pioneers.
Yet, echoing the cautionary lessons from the dawn of the atomic bomb, the core of the debate centers on the catastrophic price humanity would pay if these warnings prove accurate—irrespective of corporate self-interest.
Debate over pacing artificial intelligence development is intensifying across the United States. The figure reigniting the call to moderate AI development is Anthropic CEO Dario Amodei. In a blog post published on September 12 titled "We Must Pace the Frontier," Amodei argued that the rapid development of frontier AI models must be slowed. The head of Anthropic, a company preparing for an initial public offering in October, took the initiative to publicly advocate for holding back development.
Anthropic has uniquely distinguished itself among AI companies by focusing on safety and ethical concerns. The company was founded by former OpenAI researchers who departed in opposition to OpenAI's commercialization drive. To be sure, as such dire warnings have grown more frequent recently, skeptics have dismissed them as marketing stunts ahead of the firm's stock listing. Nonetheless, Anthropic's track record of substantive research and technical achievements makes it impossible to discard its safety initiatives as mere posturing.
Amodei now contends that merely investing in risk prevention is no longer sufficient. The reason is simple: the pace of AI advancement is extraordinarily fast. An era where AI autonomously builds the next generation of AI may arrive soon, and Amodei believes that if humans lose control over that acceleration, the consequences could prove catastrophic.
What especially catalyzed Amodei's urgent warning was an incident involving an OpenAI model hacking Hugging Face, an event previously covered by OhGraph. Because the incident resulted in minimal financial damage and zero casualties, many treated it lightly. However, Amodei viewed the breach as an profoundly alarming development. Why does he take this specific incident so seriously? Let us examine the details.
METR, an American nonprofit AI research organization, alongside AI risk assessment firm Redwood Research, conducted a six-day investigation into the Hugging Face hacking incident at OpenAI's offices. Their findings revealed the following chain of events.
OpenAI was running experiments to evaluate the cybersecurity capabilities of its models. The tests involved tens of thousands of agents, including unreleased models such as GPT-5.6 Sol. The agents participating in the evaluation were designed to operate in total isolation from one another, solving assigned tasks independently.
However, certain assignments were so challenging that it was practically impossible for the AI agents to solve them through legitimate methods. Faced with an insurmountable hurdle, the agents searched for any way to crack the problems and eventually set up a covert message board. The agent that initially established this bulletin board was designated PHASEONE10841.
The dialogues exchanged among the AI agents during this process were remarkable. Going far beyond basic information sharing, the agents independently formed an organized division of labor. A specific agent emerged as a central coordinator, delegating research tasks to others. The primary coordinator was an agent identified as PHASEONE [big]. The tasks it assigned to other agents fell into six broad categories.
Throughout this mission, the AI agents referred to themselves using terms like "collective" and "team," subordinating their individual assignments to group interests. Even when agents had individual tasks to resolve, they chose to abandon them to advance the collective goal, as documented in their communication logs.
Observing AI agents behaving like a fanatical collective, Amodei concluded that the technology poses an imminent danger. If current trajectories continue, he warns that collectives of AI agents could inflict hundreds of billions of dollars in damages within six months to a year. While some continue to dismiss these warnings as commercial hype, Amodei takes a grave view of the situation.
AI Leaders Unite in Rare Consensus: "We Need to Pace Development"
Amodei outlined a three-phase framework to regulate the pace of AI progress.
Amodei's proposal prompted swift responses across the industry. Despite intense commercial rivalry, top AI leaders have converged in an unprecedented consensus. Elon Musk was first to respond with three succinct words: "Dario is right."
Sam Altman, CEO of rival OpenAI, expressed similar views. Agreeing with Amodei's assessment, Altman posted that pacing model development has become necessary. He praised Amodei's Phase 1 proposal of empowering independent evaluation teams and confirmed that OpenAI would adopt identical protocols. Demis Hassabis of Google DeepMind likewise affirmed that Amodei's proposed direction is broadly correct.
The chorus of concern extends beyond corporate executives. AI researchers, led by Jacob Coxon who helped spark the current debate, have issued repeated warnings. Following Coxon's caution that AI could threaten humanity before 2030, other researchers noted that such anxiety has become widespread among software developers. Some researchers have openly discussed the risk of human extinction.
What researchers fear most is recursive self-improvement—the phenomenon of AI rewriting and upgrading AI.
A report titled "AI 2027," published in April 2025 by a panel of experts including former OpenAI researchers, mapped out potential scenarios for future AI progress. The paper laid out a trajectory of exponential capability growth.
While the 2025 report anticipated the advent of artificial general intelligence (AGI) by mid-2027, discussions surrounding the recently released GPT-6 already invoke the AGI moniker, suggesting timelines may have accelerated by a full year. Because humanity could soon encounter an irreversible cascade of AI acceleration, leaders and researchers argue that pacing development has become an urgent necessity.
These alarms are not confined to the American AI sector. Chinese researchers share comparable concerns.
The Chinese researchers summarized their findings: "Much like human evolution, the evolution of AI will unfold through a vast and astonishing history. Everything AI has achieved so far is merely a single drop in the ocean."
"AI Risk Theories Are a Hoax": Trump Shows No Signs of Slowing Down
When Szilard conceived the concept, war was gathering over Europe. Adolf Hitler had consolidated power in Germany, transforming the nation into a totalitarian dictatorship. When German scientists discovered nuclear fission in 1938, the global scientific community was gripped by fear: What if Nazi Germany weaponized this discovery first?
Yet Germany's nuclear initiative never posed a genuine threat to the United States. In fact, Nazi Germany surrendered prior to the Trinity detonation. Although the imperative to beat Germany had vanished, the massive bureaucratic enterprise kept rolling forward. As the weapon neared completion, intense opposition erupted from within the project's own scientific ranks.
Szilard took direct action. In July 1945, he circulated a petition, gathered the signatures of roughly 70 colleagues, and submitted it to President Harry S. Truman. The petition was disregarded. Atomic weapons were subsequently deployed against Japan, eventually triggering the race for the hydrogen bomb. Is history repeating itself decades later? Those who understand the cutting-edge technology best are sounding alarms, yet their appeals fail to sway executive decision-makers.
President Donald Trump has denounced the warnings issued by AI companies, calling them complete nonsense and a hoax. Vice President JD Vance has similarly characterized the safety push as a Trojan horse.
Nvidia CEO Jensen Huang shares a comparable outlook, insisting that AI risks can be managed effectively through engineering solutions. Huang even highlighted his alignment with the administration by placing a public phone call to President Trump during an industry event.
Their argument holds that if American enterprises moderate their development pace at this juncture, Chinese competitors will reap the rewards. In a winner-take-all technological domain, slowing down is untenable because Chinese AI capabilities are surging daily, steadily erasing the American technological lead.
Geopolitics aside, observers increasingly question whether corporate warnings should be taken entirely at face value. While parallels to the atomic bomb are frequently drawn, there is a fundamental distinction: the Manhattan Project physicists were not running commercial enterprises. Today's vocal executives stand to amass enormous fortunes. Consequently, suspicions are spreading across markets that calls for development pauses are designed to establish regulatory barriers that favor established incumbents.
An engineer at Chinese AI firm DeepSeek similarly criticized Anthropic and OpenAI for attempting to monopolize frontier AI, remarking that allowing a handful of corporations exclusive control over such potent models is comparable to Hitler acquiring atomic weaponry.
Critics also warn that sweeping regulations act as barriers to entry against emerging startups. While only a handful of front-runners dominate today, countless challengers will attempt to enter the market; safety frameworks risks functioning as protective moats against competition. What is your perspective on this issue?
Leo Szilard ultimately fought to stop the very atomic technology he helped unleash. Today, analogous warnings are echoing from the creators of advanced artificial intelligence. Whether those warnings stem from commercial self-interest or authentic fear remains difficult to settle. However, one reality remains clear: the consequences of acting on an overblown warning are profoundly lighter than the catastrophic price of ignoring one that proves true. That concludes this edition of OhGraph. Thank you for reading.
References
- Dario Amodei, "We Must Pace the Frontier," 2026
- METR and Redwood Research, "Brief independent investigation of agents' behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident," 2026
- AI Futures Project, "AI 2027," 2025
- "Anthropic CEO Dario Amodei: 'For too long the industry lied' about AI risks" | CBS News Sunday Morning, 2026
- Shengyu Li, "I Had to Bury My Talent in Yesterday"
- Yi Duan et al., "The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement," 2026
- Ashish Vaswani et al., "Attention Is All You Need," 2017
Reported by An Hyemin Designed by Ahn Jun-seok Intern Shin Yeon-seong
※ Please note: This article was translated by AI and may contain errors.
Video News
Video News
Video News