Connect with us

Hi, what are you looking for?

SecurityWeekSecurityWeek

Artificial Intelligence

Anthropic Urges Industry Coordination to Allow for a ‘Pause’ in AI Development if Risks Grow

The proposed coordination would let advanced AI labs verify that global rivals have actually stopped or slowed their work.

Anthropic

Anthropic is proposing that the world’s top artificial intelligence companies come up with a coordinated way to pause development of advanced AI systems, warning the technology is improving so quickly there’s a risk humans would lose control.

The company behind the Claude chatbot said in a blog post Thursday that as cutting-edge AI gets increasingly faster at carrying out tasks, “it would be good for the world to have the option to slow or temporarily pause” its development.

Anthropic said its internal research institute plans to explore the issue in collaboration with others and “take actions” to help build the systems for a credible slowdown or pause, without being more specific.

Anthropic rival OpenAI argued for a different approach in a report published Wednesday, saying that “democratic governments — not private companies acting alone — must ultimately determine the rules, safeguards, and accountability mechanisms.”

“Our view is that decisions about the pace of AI innovation should not be left to any one lab, company, or special interest group,” it said.

AI models are getting faster, with rapid increases in how quickly they can carry out software tasks like coding on their own, Anthropic said in its post. Based on current trends and given enough computing power, an AI system could be able to design and develop its own successor, in what is known as “recursive self-improvement.”

Advertisement. Scroll to continue reading.

Self-building AI would be a major technological milestone that would bring benefits in science, healthcare and other areas, Anthropic said, but it “also might increase the risks of humans losing control over AI systems.”

Some tech industry figures have long warned of such a scenario.

Learn More at the AI Risk Summit | Ritz-Carlton, Half Moon Bay

Anthropic’s post comes after a different warning this week from a team of researchers at the University of Toronto who showed how AI tools could be used to create a new kind of AI “worm” that adapts its hacking strategy as it spreads from device to device and takes over a vast computing network.

Self-building AI would be a major technological milestone that would bring benefits in science, healthcare and other areas, Anthropic said, but it “also might increase the risks of humans losing control over AI systems.”

Some tech industry figures have long warned of such a scenario.

Anthropic’s post comes after a different warning this week from a team of researchers at the University of Toronto who showed how AI tools could be used to create a new kind of AI “worm” that adapts its hacking strategy as it spreads from device to device and takes over a vast computing network.

The proposed coordination would let advanced AI labs verify that global rivals have actually stopped or slowed their work, “and that a bad actor could not use the auspices of a coordinated slowdown to jump ahead in secret.”

The company said a coordinated global mechanism is needed because without it a slowdown in AI development could let the “least cautious” players catch up and add to pressure on companies and governments as they make tough choices about AI safety.

Anthropic’s post comes as the company and ChatGPT-maker OpenAI race to sell shares on the stock market, in an IPO that could value Anthropic at nearly a trillion dollars.

Papernot notified Canadian cybersecurity authorities prior to releasing his report, which shows how researchers developed the worm in a laboratory by using an “open-source” AI tool that is easy for software developers to cheaply access and modify.

“In the past, cyber attackers would focus on targets that are very high value,” he said. “Banking systems, hospitals, electricity grids, water treatment systems, schools.”

Papernot agreed that there should be more collaboration between companies, government agencies and academic researchers to develop countermeasures as AI-powered hacking tools supercharge the search for computer vulnerabilities.

“That old laptop you have in your basement that you don’t check on regularly doesn’t seem like a very high-value target, but It can be used as a launch pad to attack these higher-value targets,” he said. “Anything connected to the internet is now at risk because of how low the cost has become to mount these cyberattacks.”

Learn More at the AI Risk Summit | Ritz-Carlton, Half Moon Bay

RelatedCan We Trust AI? No – But Eventually We Must

RelatedThe Wild West of Agentic AI – An Attack Surface CISOs Can’t Afford to Ignore

RelatedSweet Security Launches Agentic AI Red Teaming to Counter ‘Mythos Moment’

RelatedRaising the Cybersecurity Stakes: Ante up for the Agentic Era

Written By

Daily Briefing Newsletter

Subscribe to the SecurityWeek Email Briefing for the latest cybersecurity threats, trends, and expert insights.

Trending

Daily Briefing Newsletter

Subscribe to the SecurityWeek Email Briefing to stay informed on the latest threats, trends, and technology, along with insightful columns from industry experts.

Join this live webinar as we explore if detection-first security operations can keep pace with AI, or if it’s time to rethink prevention as the strongest default.

Register

CodeSecCon bridges the gap between dev and security. Discover best practices for secure coding, innovative risk-reduction tools, and safe AI integration to cultivate a true DevSecOps culture. Safely secure your apps!

Register

People on the Move

Vensure Employer Solutions appointed Michael Lockhart as Chief Information Security Officer.

WISeKey has appointed Alexander Hirsch as Group Chief Marketing Officer.

UltraViolet Cyber has named Andrew Park Chief Information Security Officer.

More People On The Move

Expert Insights

Daily Briefing Newsletter

Subscribe to the SecurityWeek Email Briefing to stay informed on the latest cybersecurity news, threats, and expert insights. Unsubscribe at any time.