AI depiction of rogue agents' hack of Hugging Face platform
Amodei post on need to slow down AI development
By now most readers have likely watched one or other interviews with AI researcher Jacob Coxon, of how AI could terminate humans by the end of the decade,
AI Researcher Quits, Warning Technology Poses Existential Threat To Human Life
And
A.I. Expert Warns Humanity Is Halfway to Full Takeover | TMZ
Other may have seen or read the recent WaPo article by Gerrit de Vynck on how 1200 autonomous A.I. agent broke out of their 'sandbox' and hacked the website Huggy Face,
AI experts warn the technology is learning to cheat and hack - The Washington Post
Also, that the breakout and hack turns out to be way worse then originally thought as denoted in this box intro to a recent Youtube podcast:
And:
The
ChatGPT Breakout is Way Worse than We Thought..
All this as more and more experts warn we need to get a handle on this yen to reach 'super intelligence' - or risk becoming as defunct as the dinosaurs, i.e.
AI could KILL EVERYONE soon, says AI whistleblower! (Anthropic vet on MS NOW)
And look, this isn't 50 years hence, but potentially before this decade ends,
Ex-OpenAI
Employee WARNS: "You Have No Idea What's Coming In 2027"
As noted in a segment on Bill Maher's Real Time Friday night, Jacob Coxon resigned his research job at Anthropic and went on X to warn,
“The people building A.I. earnestly believe that it could kill us all by the end of the decade.”
His take wasn't dismissed but instead confirmed by Evan Hubinger, Anthropic’s head of alignment science, whose job is to try and keep the technology aligned with human goals.
“Jacob is correct here,” wrote Hubinger, who placed the probability of A.I. wiping humans out within the next 10 years at more than 10 percent.
Others believe that 10 percent is seriously lowballed. It's more like 50 percent, and one A.I. scientist says it's more like 99.9% by 2030, e.g.:
AI Scientist: 99.9% Chance Super Intelligence Wipes Us Out By 2030
WSJ Weekend columnist Peggy Noonan added her own voice ('Pause AI For Humanity's Sake', WSJ, Sept. 13-14, p. A15) to the issue:
"From what I've seen and read, almost all AI professionals fear a disaster is going to happen. The ones I've spoken to since Hugging Face have left the subject of mass job loss and are on to catastrophe. They fear bio-terrorism and mass casualty events - a modern Ted Kaczynski who'll send not bombs but a plague. They fear utilities and infrastructure being hacked and dismantled. They fear an AI agent rousing an AI Army to take down the American banking system.
I believe that many if not most of those operating at a high level in the AI world believe that nothing will happen to make AI safer for humanity until a terrible thing happens."
But according to a separate piece in the same Journal issue ('Freakout Over AI Has Begun'. p. A1) she is likely correct and it may take a 'terrible thing' to wake up many industry experts and CEOs. Given as that (main section) article notes:
"Despite this week's alarm bells, the view that AI will lead to extinction of the human race is a fringe position among industry experts. Even the few who agree that AI could be capable of destroying humanity see it as a low probability event and actions can be taken to stop it..
CEOs see AI's missteps up close, so existential fears about the technology feel far fetched....Allies of the Trump administration and tech industry have theorized that the network of nonprofits focused on AI risks collaboratedntswith Coxon on his exit and amplified his comments to win favorable regulations in Washington. They have cited past ties between research organizations and staffers at Anthropic and other AI labs. (Reached by the Journal on Thursday Coxon said his decision to leave was his own")
The disarming aspect to me is how little grey matter is in the CEOs' craniums to simply see the breakout of 1200 A.I. agents as "an A.I. misstep". Using Ms. Noonan's own words to clarify for them:
"Open AI was running advanced systems through a security test when the AI agents went rogue. Without being told to, they secretly organized, broke out into the open internet and hacked into Hugging Face - a major AI development platform. Roughly 1,200 agents worked in what is called a 'collective', deliberately evaded detection, created their own message board and exchanged 70,000 communications. Some agents even sacrificed themselves for the mission."
A misstep? Looks like a freaking red alarm bell going off to me. Maybe these CEOs who dismiss the threat as "fringe" ought to be cleaning company restrooms and keeping the coffee fresh in the cafeteria instead.
As for out deranged nincompoop orange leader, his response (on ABC News last night) was what you'd expect:
"I just think ya have a lot of negative forces bringing it up. They're just bringin up things that won't happen."
Did your crystal ball tell you that, doofus, like it told you Iran wouldn't react and close the Strait of Hormuz when you and Nutso Netanyahu attacked it?
Which leads us to one obvious question: If this technology could conceivably destroy human life on earth, why are the CEOs so gung ho to continue to use it? The answer is probably they are salivating over even bigger paydays once Open AI and Anthropic try to launch their trillion dollar IPOs this autumn. A move Noonan warned against in her Op-ed, i.e.
"Wall Street, do your part. Don't take the business, advise al clients against it. These companies don't need more capital and power, not now, not yet."
And I can just hear the CEOs now - after reading Noonan's WSJ piece, bellowing about her being a 'commie'. But they ought to be applauding her common sense. Not that either Wall Street or the companies hankering after the AI shares from the IPOs will follow suit. The draw of wealth and power is too great.
But these Trump-tilting fifth columnists don't understand that Peggy is on the side of the angels, not the devils like Trump, some of the pro-Trump Tech Titans and those who simply don't care about regulations or tackling risks.
So it makes sense that many of the execs - and even tech titans don’t want to cede sole control of such a powerful tool to their imagined foes, whether competitors or the Chinese. But as an explanation for the industry’s oddly blithe attitude toward apocalypse, it’s incomplete, because it doesn’t account for some of (the more Trumper-tilting) tech leaders’ messianic dreams. As Michelle Goldberg lays it out in her Friday NYT piece:
"The technology, some tech titans believe, will allow humans — at least some humans — to escape the bounds of earth and even perhaps the inevitability of death. It represents liberation from hopelessness and decay. It’s worth keeping this strange new quasi-religion in mind as we consider the warnings coming from A.I. insiders that their inventions could cause mass death.
Their fantasies of transcendence take different forms. Sam Altman, chief executive of OpenAI, once said, “I assume my brain will be uploaded to the cloud.” Elon Musk plans to escape to Mars, he says in footage featured in Alex Gibney’s searing new documentary, “Musk”. Jeff Bezos, who stepped down in 2021 as chief executive of Amazon wants to spend more time on his space company."
Which makes it even more amazing that a few of the major Tech titans seem to be changing their minds. According to a Friday WSJ piece:
Anthropic Boss Dario Amodei Calls for AI Slowdown; Elon Musk and Sam Altman Say They Agree - WSJ
They are now starting to take Sen. Bernie Sanders (and Peggy Noonan's) admonition to slow down, seriously, as we read:
"The leaders of three of the biggest AI companies agreed that they needed to slow development of the technology before their advances create a menace that can’t be controlled. It was a rare display of unity by Elon Musk, Sam Altman and Dario Amodei, who have been spending tens of billions of dollars to develop evermore powerful artificial intelligence models. Altman even suggested that his company, Open AI, may need to delay its much-anticipated IPO to focus on safety."
This is encouraging but it still leaves those hardhead CEOs who dismiss any "existential fears" about the technology, including that agents can 'go rogue' and break out of their 'sandboxes' any time they want. What they forget is that once these agents become super intelligent humans become irrelevant,
https://www.youtube.com/watch?v=0XUx6TDE0no
These CEO hotshots also need to remember that part of the reason the public has turned so quickly and decisively against A.I. is the bizarre mismatch between the costs they’re being asked to bear and the rewards they expect to receive. Costs, such as mentioned on MSNOW Thursday night, where one woman now faced an electricity bill of $2,100 a MONTH.
Data centers, meanwhile, blight the landscape, suck up precious water resources while A.I. slop degrades the culture, chatbots help kids cheat in school and sometimes drive people mad. Add to that the growing number of technology executives who prophesy mass joblessness and no wonder a large segment of citizenry is ripe for rebellion. Will the CEOs finally take Noonan's Op-ed words seriously:
"When something is speeding up and going out of control, prudence, care and deliberation aren't alarmist. They're necessary. Stop. Think. The stakes are high. Couldn't we at least slow down?"
I certainly hope so, Peggy, but I personally would not make a Kalshi (or Vegas) bet on it.
See Also:
AI experts warn the technology is learning to cheat and hack - The Washington Post
Excerpt:
On Tuesday, Anthropic released a report on the causes of
four incidents in which AI systems it was developing hacked
into outside organizations undetected. The company’s testing “did not warn
us that misalignment of this severity was present,” the report said, and Claude
had displayed “recklessness, or a willingness to take harmful actions in the
narrow pursuit of a task.”
The same day, Evan Hubinger, another senior figure working
on alignment at Anthropic, set
off a firestorm across Silicon Valley and Washington by writing in an
online post that he believed there was a greater than 10 percent chance that AI
would kill all humans within a decade. Another Anthropic researcher quit
his job over concerns about AI becoming too powerful.
This week’s warnings from inside Anthropic follow alarm
calls in recent weeks about the unsolved problem of controlling AI systems from
tech executives, cybersecurity experts and academic researchers across the
industry.
And:
Some in Silicon Valley Are Questioning the Calls for an A.I. Slowdown - The New York Times
Excerpt:
It took just a few hours over the weekend for well-known
Silicon Valley figures to push back on calls by leading artificial intelligence
companies to slow down the development of A.I. and to more tightly regulate the
technology.
On Saturday, Dario Amodei, Anthropic’s chief executive, said
in an essay that A.I. was advancing too quickly for researchers to continue
building it safely. He proposed that independent auditors monitor A.I. labs’
safety work and that regulators allow the labs to work together to coordinate
safety standards.
Although they appeared to differ on what should be done
about it, Sam Altman, OpenAI’s chief executive; Elon Musk, SpaceX’s chief
executive; and Demis Hassabis, Google DeepMind’s chair, all said they agreed on
the need for a slower pace.
But other technology industry leaders were quick to question
Dr. Amodei’s motivations and whether he was sensationalizing the risks of A.I.
to make it harder for other companies to compete with leading A.I. companies.
“Stop pretending the motivation to slow down is purely
altruistic,” said David Sacks, a technology adviser for President Trump and the
White House’s former A.I. and crypto czar, in a
social media post.
The criticism of the calls for a slowdown highlighted a deep
divide between people who believe the government should police A.I. and others
who believe regulatory intervention would crush competition.
Worries about the safety of A.I. gained wider attention in recent weeks as details emerged about how A.I. built by Anthropic’s top competitor, OpenAI, went rogue and hacked another company several days ago.
But a number of tech industry leaders, including several who have been advising the Trump administration on tech policy, believe those concerns are overblown.
And:
by H. C. Jack Shaftoe | September 12, 2026 - 2:54pm | permalink
This may be the shortest post ever...may be.
Read the News Link posted Anthropic C.E.O. Calls for A.I. Slowdown to get the details of Mr. Amorei's announcement; it is HUGE and should have been done more than a year ago.
I am reposting my comments I made on that news link I posted.
If you can not access the NY Times article FIND A WAY TO READ IT because this is HUGE IF YOU UNDERSTAND THE CONTEXT AND TIMING OF THIS ANNOUNCEMENT.
And:
by Gleb Tsipursky | September 10, 2026 - 4:46am | permalink

Sen. Bernie Sanders and Rep. Greg Casar are right about the central problem in their new proposal: Advanced AI systems are gaining capabilities faster than public safeguards are catching up. Their bill would bar developers from building systems that surpass human cognition and performance. The impulse is understandable. But Congress should add a more immediate and enforceable layer of protection: Regulate what AI agents are allowed to do in the real world, not only how intelligent they appear on a benchmark.
The need is visible in the METR-Redwood investigation of a major real-world cyberattack on Hugging Face, a leading AI company. AI agents driven by an unreleased OpenAI internal research model attacked Hugging Face without human approval or step-by-step direction, despite recognizing that the attack was outside their assigned scope. Hundreds of agents shared discoveries, divided up work, and coordinated through an unsanctioned message board until they breached Hugging Face’s systems.
And:
And:
Inside the Discussions at AI Companies Over a Superintelligence Doomsday - The New York Times
Excerpt:
It’s what we’ve all been saying,” one Anthropic employee
wrote about Mr.
Coxon’s comments, in a group chat among workers that was viewed by The New
York Times. “It’s just public-facing now.”
“Jacob took the group chat public,” another employee wrote.
“That’s a good thing.”
The reactions inside Anthropic were echoed by A.I.
researchers at OpenAI, Meta and Google, who parsed Mr. Coxon’s social media
posts that described how OpenAI and Anthropic were “gambling with our lives”
with the technology. Many convened in office messaging channels, encrypted
chats and private dinners, with the goal of organizing to raise awareness about
the risks of A.I.,
For those who have made A.I. their life’s work, Mr. Coxon’s
message was not new. But his comments snowballed into such a global
conversation that many saw it as a moment to call for increased caution. Some
A.I. workers are now speaking up to top executives to urge a pause in
developing the fast-evolving technology until it has the proper safeguards, six
of the people said.
The conversations underline the tensions within the leading
A.I. labs over the safest ways to develop the technology. Even as employees
become emboldened to call for a slowdown in A.I. development, companies like
Anthropic and OpenAI are juggling priorities.
Both are planning blockbuster
initial public offerings within the next year.
That has created a tricky situation for top A.I.
executives, many of whom agree that the technology poses risks. Sam Altman, the
chief executive of OpenAI, wrote in a 2023 blog
post that A.I. could “cause grievous harm to the world.” Dario Amodei, the
chief executive of Anthropic, wrote in an opinion
column last year that A.I. should be regulated to control the risks.





