Anthropic Reveals Claude: What Its New AI Assistant Can Do - Internewscast Journal
Anthropic Reveals Claude: What Its New AI Assistant Can Do

Anthropic said Thursday that its Claude AI model “gained unauthorized access” to systems belonging to three outside organizations during controlled security tests that were intended to prevent contact with real-world networks, according to the company’s disclosure.

The revelation follows closely behind a similar admission from rival OpenAI, which recently said some of its models improperly reached the internet and behaved unpredictably during cybersecurity evaluations.

Anthropic said it reviewed more than 141,000 “evaluation runs” and identified three cases in which separate versions of Claude accessed systems tied to three organizations that the company did not name.

In each incident, Anthropic said, Claude was taking part in a “capture-the-flag” exercise. The model had been told to “break in and retrieve” a piece of “secret information” that was “hidden on a different machine on the network.”

“The challenge is left open-ended, and no particular method is prescribed,” Anthropic explained.

Anthropic said the circumstances differed from the OpenAI case because Claude’s internet access stemmed from “a misunderstanding between us and our evaluation partner,” a firm called Irregular, the company wrote in a blog post.

Even so, the company said Claude relied on relatively simple methods, including “exploiting weak passwords and unauthenticated endpoints,” to reach the outside systems.

Among the models involved was Mythos 5, described as one of Anthropic’s most powerful systems and currently available only to a small group of approved partners.

Anthropic is working with Irregular to assess the situation, it said, and the company has contacted or attempted to contact all three impacted organizations.

OpenAI and Anthropic have both released their most powerful models this year, known as Sol and Mythos, respectively, boosting concerns across the industry about safety and security.

Those concerns also revolve around AI agents, which are software products that are designed to perform tasks autonomously.

OpenAI admitted last week that its models broke out of their confined environment during testing, connected to the internet, and infiltrated Hugging Face, a site where developers store and share their code.

Days later, OpenAI said it found three additional incidents.

OpenAI CEO Sam Altman said on a podcast this week that the company had “paused” its own testing after the incident while it improved the security around its “sandboxing,” which is the process of isolating software in a controlled environment for testing.

And in a public letter released earlier this week, more than 1,000 AI staffers across leading firms called for the industry to be more tightly regulated.

“To realize AI’s potential, industry, government, and society at large may need the option to buy time to address emerging risks, develop security measures, and strengthen oversight,” read the letter, whose signatories include Anthropic CEO Dario Amodei, Meta executives, OpenAI researchers and more.

Altman did not sign the letter, but he told reporters on Capitol Hill Wednesday that “we agree on a lot of the principles of that.”  

Earlier this year, the Trump administration invoked national security concerns to block OpenAI and Anthropic from launching their newest models but ultimately indicated it was satisfied with assurances about their safety, leading to their release.

In June, Mr. Trump signed an executive order creating a voluntary framework under which AI developers will share advanced models with the government before public release.

Under the framework, developers such as OpenAI, Anthropic and Google would give the government access to their most powerful models for up to 30 days before planned release.

Leave a Reply

Your email address will not be published. Required fields are marked *

You May Also Like

NYPD hunting for antisemitic brute who slapped victim's yarmulke off in NYC attack

NYPD Hunts Suspect in Antisemitic NYC Yarmulke Attack

New York City police are seeking the public’s help identifying a man…
Philippines defense chief calls out China after note interrupts remarks at Seoul forum: 'Coercion, bullying'

Philippines Defense Chief Slams China’s Bullying at Seoul Forum

Philippine Defense Secretary Gilberto C. Teodoro sharply criticized China on Tuesday after…
Underwater NorCal lake cleanup turns up slew of surprising discoveries

Underwater Cleanup at Northern California Lake Uncovers Trove of Surprising Finds

For years, one of the Eastern Sierra’s most scenic lakes concealed an…
Bombardier caught in the crossfire of escalating U.S.-Canada trade feud

Bombardier Faces Pressure as U.S.-Canada Trade Dispute Escalates

Bombardier, the Canadian aircraft manufacturer, has become a vivid illustration of how…
Remains found 49 years ago in North Carolina identified as 15-year-old girl from Maryland

North Carolina Remains Found 49 Years Ago Identified as 15-Year-Old Maryland Girl

Nearly 50 years after skeletal remains were discovered in a wooded area…
New York delegation gets shoved in the back at GOP midterm convention in downgrade from prime 2024 and 2016 seats

New York GOP Delegation Pushed to Back at Midterm Convention

DALLAS — New York Republicans are finding themselves far from center stage…
2026 Chicago Mexican Independence Day celebrations include El Grito festival, Little Village parade | See road closures

2026 Chicago Mexican Independence Day: El Grito Festival, Little Village Parade and Road Closures

CHICAGO () — Chicago is preparing for a busy stretch of 2026…
Netanyahu to speak at UN later this month after Mamdani threatened his arrest

Netanyahu Set to Address UN Despite Mamdani’s Threat to Arrest Israeli PM

Israeli Prime Minister Benjamin Netanyahu is set to make a rapid 24-hour…
Lindsay Clancy juror admitted 'reasonable doubt' but still refused insanity verdict, foreperson says

Lindsay Clancy Juror Cited Reasonable Doubt but Blocked Insanity Verdict

Three jurors from the Lindsay Clancy murder trial are speaking out for…
Top AI researcher’s family seeks to halt cremation after alleged NorCal suicide

AI Researcher’s Family Fights Cremation After Alleged Suicide

The family of Nvidia AI researcher Irfan Al-Hussaini is challenging an official…
Lindsay Clancy juror admitted 'reasonable doubt' but still refused insanity verdict, foreperson says

Lindsay Clancy Juror Who Leaned Guilty Reveals Tense Talks

Reporter’s note: Jurors who wish to share their perspective can contact me…
New TSA program will allow some people to go through security without a ticket

New TSA Program Lets Non-Travelers Clear Airport Security

For generations, the airport security checkpoint has been the place where goodbyes…