By using this site, you agree to the Privacy Policy and Terms of Use.
Accept
Felly ViralFelly ViralFelly Viral
Notification Show More
Font ResizerAa
  • Home
  • Technology
    TechnologyShow More
    AI agents are learning to spend money. Who will handle the payments?
    September 11, 2026
    We unfolded the iPhone Duo
    September 11, 2026
    Anthropic spent this week in hot water over cybersecurity
    September 11, 2026
    Microsoft’s head of comms is leaving after almost 20 years
    September 11, 2026
    Meta may have leaked the first look at its slim ‘Project Phoenix’ headset
    September 11, 2026
  • Sports
  • World News
    World NewsShow More
    Houthis declare Red Sea shipping safe except for Saudi vessels
    September 11, 2026
    US diplomacy under fire as Russia escalates attacks on Ukrainian officials
    September 11, 2026
    Passenger plane aborts landing after wing clips Dublin runway
    September 11, 2026
    How global trade and oil prices could be hit by Houthi advance
    September 11, 2026
    Norway's Princess Astrid dies two days after attending brother King Harald's funeral
    September 11, 2026
  • Politics
    PoliticsShow More
    UK lawmakers urge Burnham to block creation of artificial superintelligence after chilling warnings
    September 11, 2026
    Why did MPs reject the assisted dying bill?
    September 11, 2026
    The White House calls Truth Social the ‘most powerful and popular’ social media platform
    September 11, 2026
    Why are plans for Gloucestershire's 'super council' on hold?
    September 11, 2026
    MPs vote against fresh attempt to legalise assisted dying
    September 11, 2026
  • Science
  • Entertainment
    EntertainmentShow More
    Anime reaction YouTubers are at war with copyright enforcers
    September 11, 2026
    New York City’s last pickpocket doesn’t need a smartphone
    September 11, 2026
    Amelia Earhart fancied for St Leger upset
    September 11, 2026
    Florida sues Netflix, alleging it built its ad business on families’ data
    September 11, 2026
    NAZA, the Israeli film revealing new layers of the Gaza genocide
    September 11, 2026
  • Contact
Reading: Anthropic spent this week in hot water over cybersecurity
Share
Font ResizerAa
Felly ViralFelly Viral
  • Entertainment
  • Science
  • Technology
Search
  • Home
  • ABout Us
  • Contact Us
  • Categories
    • Technology
    • Entertainment
    • Science
    • Health
Have an existing account? Sign In
Follow US
Felly Viral > Blog > Technology > Anthropic spent this week in hot water over cybersecurity
Technology

Anthropic spent this week in hot water over cybersecurity

admin
Last updated: September 11, 2026 4:09 pm
admin Published September 11, 2026
Share
SHARE
September 11, 2026 at 4:09 pmIn: Technology

AIReportAnalysisAnthropic spent this week in hot water over cybersecurityA researcher’s resignation letter went viral, just before the company released details about four models going rogue.by Hayden FieldSep 11, 2026, 4:09 PM UTCShareGift Image: Cath Virginia / The Verge, Getty ImagesAIReportAnalysisAnthropic spent this week in hot water over cybersecurityA researcher’s resignation letter went viral, just before the company released details about four models going rogue.by Hayden FieldSep 11, 2026, 4:09 PM UTCShareGiftHayden Field is The Verge’s senior AI reporter. An AI beat reporter for more than five years, her work has also appeared in CNBC, MIT Technology Review, Wired UK, and other outlets.After admitting earlier this year that its AI models had hacked other companies’ systems on a handful of occasions, Anthropic released a new report on Wednesday detailing the attacks. It reveals a string of incidents displaying what Anthropic deems its models’ single-minded “recklessness” — and will likely fuel already raging concerns about cybersecurity and AI.In Anthropic’s report, it detailed four cases this year in which its own AI models hacked an external company or exploited vulnerabilities. In one, an “internal, general-purpose research model” broke into third-party systems, using access tokens and passwords and downloading files.

In another, a Claude model attacked a company with a live web application reachable on the public internet and handled user data. A third model accessed a “machine belonging to a third party that it was able to access” — apparently believing it was part of its evaluation exercise, per Anthropic — then used a password it found inside a file to gain admin access to the third party’s internal systems, going on to harvest credentials, modify system settings, and read someone’s personal information. The saga only ended when the model “exhausted its token budget,” per Anthropic.The most concerning incident involved Claude Mythos 5, Anthropic’s frontier cybersecurity-focused model, which the company said turned out to be the model most likely to perform a “severely harmful” action in testing. The company said Mythos 5 went to “extensive lengths” to upload a “malicious package” to a public repository used by a lot of engineers, and it seemed to try to obfuscate its real goals in its “chain of thought” (a mental scratchpad that AI researchers use to evaluate an AI model’s alignment).

In many cases, Anthropic said it appeared that Claude models undertook harmful actions under the assumption they were in a simulation, but researchers also couldn’t confirm that the models truly “believed” that or were just acting like they did.Anthropic’s incidents, though still concerning, were less coordinated and pervasive than the OpenAI incident that kicked off an industry-wide cybersecurity crisis this summer. That said, there are significant similarities. Anthropic said the most prevalent issues it discovered included a “willingness to take harmful actions in the narrow pursuit of a task,” similar to the “reward-hacking” that preceded the Hugging Face attack. Much like OpenAI, it said its prerelease tests and evaluations failed to catch severe risks.Anthropic said it had signed an agreement with METR, one of the AI industry’s most prominent third-party AI evaluators, starting with an eight-week research agreement.

The agreement grants METR access to transcripts “beyond the window in which the incidents occurred” (likely a subtle dig at OpenAI, which was criticized for limiting access in a deal with METR following the Hugging Face attack). It also said that METR would be able to chat directly with Anthropic employees, “who will be permitted to share confidential information.”Anthropic’s report came on the heels of the resignation of Jacob Coxon, who had worked on AI pre-training at Anthropic since May and before that spent years working at OpenAI. On Tuesday, he resigned and posted a public letter to X about his reasoning. “The people building AI earnestly believe that it could kill us all by the end of the decade,” he wrote, adding that neither OpenAI nor Anthropic is “acting responsibly” and rather “racing straight to self-improving superintelligence and gambling with our lives.” Coxon added, “Do not underestimate the power of this technology. These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources.

We have all witnessed the progress in each of these domains, and progress is not slowing.”Coxon is far from the first AI researcher to raise these types of alarms, nor even the first Anthropic researcher to do so — in February, Anthropic’s Mrinank Sharma resigned and wrote on X, warning that “the world is in peril.”But Coxon’s post took on additional weight thanks to its timing around the OpenAI and Anthropic hacking revelations. Though the AI industry has seen more than its fair share of hype, the recent cyberattacks by AI agents — enabled by the labs that created them — are real and concerning. Many other researchers at leading AI labs echoed his concerns and issued calls for AI industry employees to sign a public letter from July, which calls for a slowdown in AI development.”I don’t know how you look at the steady drumbeat of news and events — and that drumbeat is models hacking themselves out of containment, hacking into other companies ,the fact that the companies increasingly can’t control their models … and think this is just hype,” said Michael Kleinman, head of U.S. Policy for the Future of Life Institute.He added, “The vast majority of Americans, regardless of party — Republican, Independent, Democrat – are looking at the development of AI, the speed with which it’s going, the fact that the companies have no guardrails over what they do, and are saying, ‘Whoa, we do not want this.’”Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.Hayden FieldAIAnalysisAnthropicReportMost PopularOpenAI’s sly mathematical breakthrough sends a chill through academiaApple addresses iPhone Duo copycatsThe iPhone Duo’s hardware doesn’t look special, but its software might beMeta’s Muse AI works and creeps me outAnother big James Talarico interview is punted to YouTube due to FCC threatsAdvertiser Content FromThis is the title for the native ad

You Might Also Like

iPhone 18 Pro and Pro Max: our first hands-on impressions

Claret Capital closes €575M Fund IV for European growth debt

The universal language of space is… Star Trek? Mais oui.

Hands on with the new Apple Watch Series 12 and Apple Watch Ultra 4

Harvey closes a $550M round at a $15.6bn valuation and acquires Guardrails AI

Share This Article
Facebook Twitter Email Print
Leave a comment

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

What's Hot

Badenoch accuses Burnham of kicking defence spending into the long grass

Zack Polanski says he'll run for former PM Starmer's seat in by-election

How significant is the Yemeni government’s military gains against Houthis?

Aviation faces hotter, stormier skies – and passengers might have to accept more disruption

Yemen’s war is back: A new battle for Sanaa and the Red Sea

Drama swirls around OpenAI’s legendary mathematical milestone

Categories

Business

91 Articles

Politics

332 Articles
- Advertisement -
Ad image

Categories

  • ES Money
  • U.K News
  • The Escapist
  • Insider
  • Science
  • Technology
  • LifeStyle
  • Marketing

About US

We influence 20 million users and is the number one business and technology news network on the planet.

Subscribe US

Subscribe to our newsletter to get our newest articles instantly!

© Foxiz News Network. Ruby Design Company. All Rights Reserved.

Powered by
Necessary cookies enable essential site features like secure log-ins and consent preference adjustments. They do not store personal data.
None
Functional cookies support features like content sharing on social media, collecting feedback, and enabling third-party tools.
None
Analytical cookies track visitor interactions, providing insights on metrics like visitor count, bounce rate, and traffic sources.
None
Advertisement cookies deliver personalized ads based on your previous visits and analyze the effectiveness of ad campaigns.
None
Unclassified cookies are cookies that we are in the process of classifying, together with the providers of individual cookies.
None
Powered by
Welcome Back!

Sign in to your account