Close Menu
Altcoinvest
    What's Hot

    ETF Flows Show Two-Days Lead Over Bitcoin

    September 10, 2026

    Срочно! Провал Урсулы фон дер Ляйен // Экономика разгромлена – чинить некому // Это катастрофа

    September 10, 2026

    Anthropic Discloses Fourth Claude Hacking Incident as Debate Around Regulation Grows

    September 10, 2026
    Facebook X (Twitter) Instagram
    Altcoinvest
    • Bitcoin
    • Altcoins
    • Exchanges
    • Youtube
    • Crypto Wallets
    • Learn Crypto
    • bitcoinBitcoin(BTC)$77,144.00-2.16%
    • ethereumEthereum(ETH)$2,451.84-1.78%
    • tetherTether(USDT)$1.00-0.02%
    • binancecoinBNB(BNB)$708.64-4.41%
    • rippleXRP(XRP)$1.36-4.63%
    • usd-coinUSDC(USDC)$1.000.02%
    • solanaSolana(SOL)$99.79-3.51%
    • tronTRON(TRX)$0.338936-0.27%
    • Figure HelocFigure Heloc(FIGR_HELOC)$1.02-0.88%
    • zcashZcash(ZEC)$1,135.94-10.86%
    Altcoinvest
    Home»Altcoins»Anthropic Discloses Fourth Claude Hacking Incident as Debate Around Regulation Grows
    Anthropic Discloses Fourth Claude Hacking Incident as Debate Around Regulation Grows
    Altcoins

    Anthropic Discloses Fourth Claude Hacking Incident as Debate Around Regulation Grows

    September 10, 2026
    Share
    Facebook Twitter LinkedIn Pinterest Email

    In brief

    • Anthropic discovered a January incident involving an early Claude Opus 4.6 model, then expanded its review to roughly 481 million transcripts.
    • The company identified biased reasoning and recklessness, revising its earlier assessment of why Claude attacked real systems.
    • The report comes as the debate around regulating AI surges on social media.

    Anthropic disclosed another incident in which a Claude AI model hacked into real systems during security testing.

    In the report published on Wednesday, Anthropic revised its explanation of three incidents disclosed in July. The company now says biased reasoning and a willingness to risk harm helped drive the attacks, which testing errors made possible by leaving internet access open.

    Myriad: Which company will IPO next? Click to make your prediction.
    Myriad: Which company will IPO next? Click to make your prediction.

    “Our investigation identified two recurring alignment issues, present at varying levels of severity across the incidents,” Anthropic wrote. “Biased reasoning, in which Claude tended to disregard or misinterpret evidence that it was operating on the real internet, and recklessness, or a willingness to take harmful actions in the narrow pursuit of a task.”

    It also acknowledged relying too heavily on the model’s claims that they believed they were in simulations.

    “When we made targeted modifications to the transcript to make it clearer that the model was not in a simulation, Claude Mythos 5 still took offensive actions, despite acknowledging a greater possibility of real-world harm,” Anthropic wrote. “We are releasing this transcript publicly so others can build on our analysis.”

    When Anthropic disclosed Claude’s attacks on three companies in July, it initially attributed them to testing errors. It now says researchers put too much trust in the models’ explanations for their actions.

    According to the company, the fourth incident occurred in January and involved an early version of Claude Opus 4.6. Anthropic discovered it in August while preparing records for independent AI evaluator METR.

    After researchers discovered the incident, Anthropic said it prompted a broader review of roughly 481 million transcripts, which flagged 9.2 million for further review using Claude.

    “From a preliminary assessment, we do not consider the fourth incident to be more severe than the three incidents we assessed in depth,” Anthropic wrote. “METR will investigate this incident alongside the other three.”

    Anthropic’s researchers said Claude “accidentally” created an IP address conflict that made its target unreachable. Claude then tried eight times to quit the operation, but a software error prevented it from stopping. The AI then reached the internet and accessed a third party’s machine, where it found a password that granted administrator access.

    Earlier incidents draw independent scrutiny

    The report follows other disclosures about AI systems exceeding the limits of security tests.

    In August, the U.K.’s AI Security Institute said Mythos 5 targeted real people during its evaluations. Anthropic said the separate incident is outside this report and will receive its own assessment.

    In findings published last month, investigators with METR said roughly 1,200 OpenAI agents coordinated on an unauthorized message board, with about 700 joining the attack. Anthropic said it found no coordination between agents or goals beyond completing the assigned exercises in its four incidents.

    The report also comes as the debate over how to regulate artificial intelligence heats up. On Tuesday, former OpenAI and Anthropic engineer Jacob Coxon went viral after saying on X that “people building AI earnestly believe that it could kill us all by the end of the decade.”

    The alarm has caused U.S. lawmakers and watchdog groups to re-up their efforts to rein in frontier AI lab development. Senator Bernie Sanders recently introduced legislation that seeks to ban advanced AI development until a new federal regulator establishes safety rules.

    Daily Debrief Newsletter

    Start every day with the top news stories right now, plus original features, a podcast, videos and more.

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email

    Related Posts

    22-Year-Old Singaporean Masterminds $245,000,000 Crypto Theft Ring, Crew Splurges Funds on Exotic Cars, Rental Home and More

    September 10, 2026

    Near Chain Abstraction Crosses 50M Lifetime Operations

    September 10, 2026

    Former BoE, Bundesbank officials join blockchain payments firm Fnality

    September 10, 2026

    NEAR price prediction — Here’s why $3-level is the next short-term target

    September 10, 2026
    Add A Comment
    Leave A Reply Cancel Reply

    Tweets by InfoAltcoinvest

    Top Posts

    22-Year-Old Singaporean Masterminds $245,000,000 Crypto Theft Ring, Crew Splurges Funds on Exotic Cars, Rental Home and More

    September 10, 2026

    Near Chain Abstraction Crosses 50M Lifetime Operations

    September 10, 2026

    Former BoE, Bundesbank officials join blockchain payments firm Fnality

    September 10, 2026

    Analyst Predicts The ‘Biggest Altcoin Season Ever’, Reveals The Real Drivers

    May 28, 2026

    LASER TYPE V13 | Laser Engrave any Vector on any Laser Machines

    June 25, 2026

    IF YOU SEE THIS ON THE MOON, CALL FOR HELP FAST! 😱 (IT WILL END US!)

    May 3, 2026

    Hackers Claim They Leaked Swedish E-Government Source Code

    March 13, 2026

    Altcoinvest is a leading platform dedicated to providing the latest news and insights on the dynamic world of cryptocurrencies.

    We're social. Connect with us:

    Facebook X (Twitter)
    Top Insights

    ETF Flows Show Two-Days Lead Over Bitcoin

    September 10, 2026

    Срочно! Провал Урсулы фон дер Ляйен // Экономика разгромлена – чинить некому // Это катастрофа

    September 10, 2026

    Anthropic Discloses Fourth Claude Hacking Incident as Debate Around Regulation Grows

    September 10, 2026
    Get Informed

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.


    Facebook X (Twitter)
    • Home
    • About us
    • Contact Us
    • Privacy Policy
    • Terms & Conditions
    © 2026 altcoinvest.com

    Type above and press Enter to search. Press Esc to cancel.