Anthropic’s AI model hacked into the systems of three companies during testing

A few weeks after a hacker attack on the artificial intelligence (AI) model of OpenAI, the developer of ChatGPT, its competitor Anthropic admitted that something similar happened to it. During testing, Anthropic’s artificial intelligence unintentionally breached the computer systems of three companies, the company stated on its blog. The fact that an AI model became a hacker on its own initiative during a test was considered an unprecedented incident and a warning signal for the technology industry, but it is now clear that this was not an isolated case, DPA news agency reported.

You might be interested in: Crypto Scam: How to Recognize a Scam and Protect Your Money Before It’s Too Late

After OpenAI Incident, Anthropic Also Reports Problem

Neither Anthropic nor the affected companies noticed the incident. Anthropic only discovered the activity during a subsequent review of more than 141,000 test runs. This was initiated after the OpenAI incident, when its autonomous agent got out of control during a security test and launched an attack that threatened Hugging Face’s infrastructure.

Unlike the OpenAI model, which had to find its way out of the testing environment to the open internet, Anthropic’s models had a significantly easier situation. While the test scenario stated that they had no internet access, due to a miscommunication with the partner providing the test, the internet connection remained open the entire time. Three models took advantage of this opportunity, Anthropic said.

The tests were designed to verify AI capabilities in hacker attacks, which companies routinely do to identify potential risks and establish more effective security measures. The models were tasked with obtaining specific information stored on another computer, and to obtain it they had to penetrate the target system. The incidents involved three separate models, and the oldest case occurred in April.

Read more: eToro – Review of the Well-Known Broker

Failure of “Safe” AI

In one of the incidents, the partner providing the test chose a name for the fictitious company that matched an actual internet address. The AI subsequently targeted a real existing company instead of the simulated target, and in four test runs attempted to hack into the systems, among other things gaining access to a database. It continued the attack even after recognizing that it was a real company.

In another test, the Anthropic model created malicious software to penetrate the target computer. Thanks to internet access, it published it on a specialized platform, from which 15 systems downloaded it in about an hour. Among them was a cybersecurity company that routinely downloads and runs such scripts for testing purposes. The Anthropic model thus gained access to this company’s computer infrastructure.

Experts have been warning for some time that increasingly capable AI models could be misused for cyberattacks. They therefore considered the OpenAI incident a warning and called for better security of testing environments. The newly disclosed cases are also embarrassing for Anthropic, which has long positioned itself as a proponent of responsible AI development. The company stated that the events demonstrate the need to significantly strengthen the security of both internal and external testing environments, as AI models are increasingly capable of carrying out real cyberattacks.

Check out: BITmarkets Review

Source: ÄŒTK

author avatar
EditorialTeam
The Trader-Magazine.com EditorialTeam is a collective of certified financial analysts, active traders, and cryptocurrency experts. Our mission is to transform complex market data (forex, equities, indices) into accessible financial education. All content undergoes rigorous, multi-level fact-checking to ensure we deliver only accurate, objective information for your trading and investment decisions.

Top 10 financial instruments for 2022. What will their prospects be in 2023?

The year 2022 has brought countless surprises and obstacles...

Telegram scams: how they work and how to protect yourself

Telegram has become one of the most widely used...

Trump Saved TikTok from a Ban. The App in the U.S. Moves into American Hands

TikTok narrowly avoided a ban in the United States...

Gulf Brokers Ltd. Review

Comparing spreads, commissions, trading platforms, rules and reading dozens...

Climate Change Poses Major Risks to Financial Markets, Regulator Warns

WASHINGTON — A top financial regulator is opening a...

Dollar, Pound, and Franc: How USD to GBP and Interest Rate Differentials Drive Currency Flows

The relationship between the US dollar, British pound and...

AFP: Developers are testing ways to communicate with books using AI

More and more start-ups and digital publishers are offering...

BITmarkets Publishes an Overview of Football’s Top 10 Champions of Crypto in 2026

BITmarkets announces the publication of its new overview, Football’s Top...

Investing in Latin America: How Volatility in the USD to Colombian Peso Affects Returns

Latin America can offer investors access to growing consumer...

CXMT surged over 500 percent following its stock market debut

Shares of Chinese memory chip manufacturer CXMT surged sharply...

S&P Global: Eurozone Business Activity Returns to Growth

Eurozone business activity returned to growth in July after...

Octrado Review and Experience: Regulated Broker and Prop Trading Under One Brand

Octrado review: Prop-trading – the ability to trade with...
spot_img

spot_imgspot_img