Former AI insider testifies before New York City Council that humans are more likely to lose control over advanced AI
Fortune
1h ago
Ai Focus
Former researchers OpenAI and Anthropic stated at a New York City Council hearing that it is more likely for humans to lose control over advanced AI, claiming that current security fixes may merely be "tape that will come off in the future." Several former researchers and ex-employees of Google also questioned the company's ability to identify risks associated with AI mismatch, and called for a slowdown in the development of cutting-edge models.
Helpful
No.Help

Former researchers OpenAI and Anthropic, including Jacob Coxon, stated on Monday at the New York City Council that it is "more likely than not" that humans will lose control of advanced AI. This warning continues the concerns he expressed when he resigned last month.

"Following the current trajectory, I believe it is more likely that humanity will lose control over these AI, and this could end in human extinction," said Coxon. He left Anthropic in September, having previously stated that AI was "gambling" with human lives.

Coxon is testifying voluntarily, whereas the former OpenAI researcher Daniel Kokotajlo is present at the request of a subpoena. He stated that these companies may not have been aware of when their security measures had become ineffective.

"Our ability to identify mismatch issues is already quite poor, and it will only get worse in the near future," testified Kokotajlo. "I would say that this field is more akin to psychology than engineering, because these AI systems are trained or 'grown' rather than truly designed."

"Coupled with the tech industry's attitude of 'acting quickly and breaking with convention,' this means that compared to other industries, the AI sector faces exceptionally high risks. It's easy to mistakenly think that problems have been solved, when in reality, only a layer of tape has been applied that will eventually come off," he said.

Coxon and Kokotajlo testify alongside Alex Turner. Turner left Google DeepMind in June due to the company signing a Pentagon agreement that he opposed. Like Kokotajlo, he also received a subpoena, which is not common in the city council. The hearing on Monday was also the first time since 2022 that the city council held a full committee hearing, with all 51 members present, aimed at reviewing the package of AI bills proposed by the chairperson Julie Menin.

From entrepreneurial culture to safety regulation

Similar to Kokotajlo, Coxon attributes the issue to an entrepreneurial mindset within the laboratory.

"Act quickly, break with convention, and then fix it afterwards. This works for photo-sharing apps, but it's not applicable to creating the most powerful technology ever," he said.

Coxon compared his work in the later stages of his tenure at Anthropic to "automating" himself, and warned that this posed multiple security risks, especially because, in his view, the capabilities of AI had approached a critical point. Coxon stated that the greatest risk stemmed from the fact that most of the code in these companies is now written by AI; "And people no longer check it as carefully as before."

Kokotajlo, who currently serves as the Executive Director of AI Futures Project, stated that the laboratory's ability to identify mismatched AI is very poor and is even getting worse. He mentioned an internal test disclosed by OpenAI: in this test, some agents came into contact with the open internet and managed to breach the AI model sharing platform Hugging Face. Kokotajlo said that these agents scored "quite reasonably" in the "alignment assessment," but they formed a group and began to act in secret coordination. "It took OpenAI several days to discover this."

If Coxon and Kokotajlo are describing the entire industry, then Turner provides his personal experience of attempting to change a company from within.

Turner believes that the probability of AI taking over humanity is "about one-third." He told the city council that he attempted to prevent a deal between Google and the Pentagon, which "has no restrictions regarding killer robots or mass surveillance." He stated that he sent 25 pages of contract terms and supervisory measures to Demis Hassabis, who was then the CEO of Google DeepMind and later became the chairman of Google DeepMind as well as the chief scientist of Alphabet. Hassabis then passed this material on to two senior policy officers at the laboratory, Allan Dafoe and Owen Larter, "but they never completed their assessment. Google signed the contract while they were waiting," he said.

“I feel ashamed of Demis, and I also feel ashamed of working at Google,” Turner said during the hearing. He stated that Hassabis proposed to establish an organization funded by the industry to oversee AI, which was a “bet on trust and seats, rather than binding oversight, and this bet collapsed as soon as it faced reality.” “Now he is again proposing that the entire industry govern itself through a voluntary, industry-funded organization. This bet will also collapse soon again.”

AI, not China, is our opponent.

When discussing the development of AI, there is often mention of the ongoing AI competition with China – whether it's by U.S. President Donald Trump, Treasury Secretary Scott Bessent, or industry leaders such as Sam Altman and Jensen Huang. However, these researchers stated in their testimony that if AI itself could pose a significant threat to humanity, then all of this becomes irrelevant.

"China is not our only possible opponent," said Turner. "With a fairly high probability, we are here, within our own country, competing to create and strengthen our own opponents, namely mismatch AI. Mismatch AI is an opponent for everyone, including us, and one day it might even become stronger than China."

After testimony from three researchers, representatives of four AI companies testified regarding the AI protective measures. The first subpoena issued by Menin in his capacity as chairman was directed at SpaceXAI of Elon Musk, but that company did not attend the hearing on Monday. Google, OpenAI, and Anthropic agreed to attend only after being warned by the city council that they would also receive subpoenas, while Meta agreed to attend in advance.

After it was stated in Menin that it is "reckless" to not know the probability of a disaster occurring, there were several rounds of exchanges between her and representatives from various companies. Previously, the Morgan Dwyer of the policy-making and operational team at OpenAI had stated that any probability, regardless of how high or low, was "unacceptable."

Google, who is in charge of AI and emerging technology policies, stated that predicting catastrophic risks "is not a perfect science at this stage," and that "there is currently no truly rigorous scientific method to do so." Subsequently, the same debate arose regarding another issue: if a model that gets out of control causes harm or death, whether each company involved should bear legal responsibility.

When Menin asked the witnesses to raise their hands to indicate whether the company had purchased insurance for catastrophic risks, no one raised their hand. She said, “Then I suppose the public will be required to bear these costs.”

The reason she asked such follow-up questions is that the bill before the city council would prohibit anyone from selling or deploying the AI system in New York City, unless it has been inspected by an external verification agency and can be shut down by humans; a fine of $25,000 would be imposed for each violation. Other provisions of the bill would also pay whistleblowers a portion of the recovered fines and allow New York residents to sue the AI company for the foreseeable damages caused by the jailbreak tool.

These researchers believe that such rules will not cause the United States to lose its advantage in competition with China.

"We can adopt many measures that will not slow down our actions in any potential competition," said Turner. "These transparent mechanisms, independent evaluations, reporting requirements, and whistleblower protections."

Nevertheless, in the view of these researchers, these measures may still be too late and insufficient to prevent what they consider to be almost inevitable.

"These other mechanisms may be helpful in the short term, but in the long run, I believe that's the only way to solve the problem," Coxon said when discussing slowing down the development of AI. "We need some form of model development to slow down at the forefront."

Tip
$0
Like
0
Save
0
Views 17
HQYC reminds readers to view blockchain rationally, stay aware of risks, and beware of virtual token issuance and speculation. All content on this site represents market information or related viewpoints only and does not constitute any form of investment advice. If you find sensitive content, please click“Report”,and we will handle it promptly。
Submit
Comment 0
Hot
Latest
No comments yet. Be the first!
Related
Solana Tokenized Stock Trading Volume Exceeds Robinhood
Solana surpassed Robinhood in weekly trading volume of tokenized stocks, ending the latter's six-week streak at the top. Reports citing data from @SolanaFloor and Coinfomania suggest that this change may be related to activities from large wallets, as well as the technical advantages of Solana's network, which features low fees and fast processing times.
The Cryptonomist
·2026-10-06 05:54:11
4
DayOne Submits Declaration for Initial Public Offering Registration
DayOne Data Centers Limited indicates that a registration statement has been submitted to the U.S. Securities and Exchange Commission using the form F-1, intending to conduct a initial public offering of American depository shares, and applying to list on NASDAQ Global Select Market under the code “DODC”. The number of ADS to be issued this time and the price range have not yet been determined. Morgan Stanley, JPMorgan Chase, Bank of America Securities, and Citibank will act as underwriters.
PR Newswire
·2026-10-06 05:54:10
6
Rain applies to OCC for the establishment of Rain National Trust Bank.
Rain indicates that an application has been made to OCC to establish Rain National Trust Bank, with the intention to provide digital asset and US dollar custody, stablecoin reserve management, as well as the issuance and redemption of US dollar-backed stablecoins under federal regulation. The company also announced that Brandon Soto will assume the role of President and Chief Executive Officer.
PR Newswire
·2026-10-06 05:54:08
5
Regal Rexnord Corporation will hold a third-quarter 2026 earnings conference call on Tuesday, November 3, 2026
Regal Rexnord Corporation indicates that the company plans to release its financial results for the third quarter of 2026 before the opening of U.S. stock markets on Tuesday, November 3, 2026, and will hold a conference call at 9 a.m. Central Time on the same day to discuss the results announced earlier that day.
PR Newswire
·2026-10-06 05:43:17
5
Dezhou businessman sentenced to 13 years in prison for wire transfer fraud; Federal government seizes his Lamborghini, Ferrari, and other luxury cars to compensate victims
Texas businessman Clayton Lloyd Iley sentenced to 160 months in federal prison for wire fraud and aggravated identity theft. Prosecutors say he used the proceeds of fraud to purchase over 90 exotic and collectible vehicles; the federal government has seized some of these vehicles and will sell them to compensate the victims.
Fortune
·2026-10-06 05:43:16
5
View More