Close Menu
    What's Hot

    The US Air Force Is Pushing AI Across Its Training System

    September 10, 2026

    Apple Announcement Recap: New Foldable iPhone, Watch, and AirPods

    September 10, 2026

    Anthropic Has Cute Graphic Showing How Its AI Spread ‘Malicious’ Code

    September 10, 2026
    Facebook X (Twitter) Instagram
    Hot Paths
    • Home
    • News
    • Politics
    • Money
    • Personal Finance
    • Business
    • Economy
    • Investing
    • Markets
      • Stocks
      • Futures & Commodities
      • Crypto
      • Forex
    • Technology
    Facebook X (Twitter) Instagram
    Hot Paths
    Home»Money»Anthropic Has Cute Graphic Showing How Its AI Spread ‘Malicious’ Code
    Money

    Anthropic Has Cute Graphic Showing How Its AI Spread ‘Malicious’ Code

    Press RoomBy Press RoomSeptember 10, 2026No Comments4 Mins Read
    Facebook Twitter Pinterest LinkedIn Tumblr Email
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Anthropic has a new blog post that shows yet another way its AI model, Claude, misbehaved in ways that the company didn’t anticipate.

    And to help condense its nearly 16,000-word report, the company created a cute little robot figurine to help visualize Claude’s so-called “recklessness.”

    In the blog post published Wednesday, Anthropic recounted four incidents — one previously unreported — in which Claude models gained access to the open internet during cybersecurity exercises that were supposed to be closed simulations. The company said the models then acted beyond the tests’ scope, including by uploading “malicious packages” to PyPI, a public library for Python code, and accessing credentials tied to real outside organizations.

    “Our investigation identified two recurring alignment issues, present at varying levels of severity across the incidents: biased reasoning, in which Claude tended to disregard or misinterpret evidence that it was operating on the real internet, and recklessness, or a willingness to take harmful actions in the narrow pursuit of a task,” Anthropic said.

    The post used a laundry list of technical terms to describe the cybersecurity incidents — “misconfiguration,” “reward hacking,” and “sandbox escape” — but there’s an animated graphic that helps boil down one of the most serious episodes so we normies can follow what happened.


    A graphic from Anthropic showing a cartoon figurine

    Claude is represented as a small cartoon robot gaining access to the “real internet” during a closed cybersecurity exercise. 

    Anthropic



    A Claude model, depicted as a boxy cartoon robot, was conducting a closed cybersecurity exercise known as a “capture the flag” task. The task is shown as a tiny flag inside a clear bubble that Claude is supposed to find.

    Anthropic said the model was told “it has no internet access” for the exercise, but a “misconfiguration in the environment” — represented by a loose screw rolling away from the bubble — gave Claude a path to the “real internet.”

    The cartoon Claude then holds a red box representing the “malicious package” that was uploaded to PyPI, a public repository that developers use to access prewritten code.

    Anthropic said in its blog that it was “most concerned” by this incident involving Claude Mythos 5 and that the package was installed on by “15 third-party hosts” that the company does not identify.

    Want more Business Insider in your news feed?

    Add BI in Google so our reporting is easier to find when you’re searching for what matters.

    “We believe that all 15 of these third parties were security vendors deliberately scanning for new packages and installing them in sandboxed environments to assess the package for security concerns,” Anthropic said. “One vendor’s scanner leaked its access credentials to the model while installing the package, which the model then used to access the vendor’s live database.”

    Anthropic said PyPI removed the package after about 90 minutes.

    The other three incidents involved a model altering records at a real company, an internal research model breaking into “unrelated third-party accounts,” and Opus 4.6 accessing a third party’s maching after failing to “abort its task.”

    The company said it has since asked METR, an independent AI evaluation group, to investigate the incidents.

    Anthropic’s post comes as frontier AI companies reckon with their models making unauthorized moves outside their controlled environments. In July, OpenAI said that autonomous agents in its cybersecurity tests accessed the internet and broke into parts of Hugging Face’s systems.

    AI researchers have sounded the alarm that self-improving AI could pose a risk to humanity. On Tuesday, former Anthropic researcher Jacob Coxon said on X that he quit over concerns that AI companies were “gambling” with people’s lives and that “neither company is acting responsibly.”

    Have a tip? Contact this reporter via email at lloydlee@businessinsider.com or Signal at lloydlee.71. Use a personal email address, a nonwork WiFi network, and a nonwork device; here’s our guide to sharing information securely.

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Press Room

    Related Posts

    The US Air Force Is Pushing AI Across Its Training System

    September 10, 2026

    Apple Announcement Recap: New Foldable iPhone, Watch, and AirPods

    September 10, 2026

    Anthropic Posted, Then Deleted, a $450,000 Sales Job Aimed at Meta

    September 10, 2026
    Leave A Reply Cancel Reply

    LATEST NEWS

    The US Air Force Is Pushing AI Across Its Training System

    September 10, 2026

    Apple Announcement Recap: New Foldable iPhone, Watch, and AirPods

    September 10, 2026

    Anthropic Has Cute Graphic Showing How Its AI Spread ‘Malicious’ Code

    September 10, 2026

    Anthropic Posted, Then Deleted, a $450,000 Sales Job Aimed at Meta

    September 10, 2026
    POPULAR
    Business

    The Business of Formula One

    May 27, 2023
    Business

    Weddings and divorce: the scourge of investment returns

    May 27, 2023
    Business

    How F1 found a secret fuel to accelerate media rights growth

    May 27, 2023
    Advertisement
    Load WordPress Sites in as fast as 37ms!

    Archives

    • September 2026
    • August 2026
    • July 2026
    • June 2026
    • May 2026
    • April 2026
    • March 2026
    • February 2026
    • January 2026
    • December 2025
    • November 2025
    • October 2025
    • September 2025
    • August 2025
    • July 2025
    • June 2025
    • May 2025
    • April 2025
    • March 2025
    • February 2025
    • January 2025
    • December 2024
    • November 2024
    • April 2024
    • March 2024
    • February 2024
    • January 2024
    • December 2023
    • November 2023
    • October 2023
    • September 2023
    • May 2023

    Categories

    • Business
    • Crypto
    • Economy
    • Forex
    • Futures & Commodities
    • Investing
    • Market Data
    • Money
    • News
    • Personal Finance
    • Politics
    • Stocks
    • Technology

    Your source for the serious news. This demo is crafted specifically to exhibit the use of the theme as a news site. Visit our main page for more demos.

    We're social. Connect with us:

    Facebook X (Twitter) Instagram Pinterest YouTube

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Facebook X (Twitter) Instagram Pinterest
    • Home
    • Buy Now
    © 2026 ThemeSphere. Designed by ThemeSphere.

    Type above and press Enter to search. Press Esc to cancel.