Close Menu
MyAppsPlus

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    Vacuums, hair dryers… toothbrushes? I asked Jake Dyson how the brand decides what gets to be a Dyson product

    October 10, 2026

    How big tech transformed public schools

    October 10, 2026

    Samsung raises Galaxy S26 smartphone prices

    October 10, 2026
    Facebook X (Twitter) Instagram
    Facebook X (Twitter) Instagram
    MyAppsPlusMyAppsPlus
    Saturday, October 10
    • Home
    • Breaking Tech
    • Apps & Software
    • AI & Automation
    • Android
    • iPhone & iOS
    • More
      • Reviews
      • How-To Guides
      • Deals & Discounts
      • Shop
    MyAppsPlus
    Home»Breaking Tech»Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet instead
    Breaking Tech

    Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet instead

    myappsplusBy myappsplusOctober 10, 2026003 Mins Read
    Share Facebook Twitter Pinterest Copy Link LinkedIn Tumblr Email Telegram WhatsApp
    Follow Us
    Google News Flipboard
    Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet instead
    Share
    Facebook Twitter LinkedIn Pinterest Email Copy Link

    Anthropic said its models exploited websites on the internet, including some run by U.S. government agencies, and it will turn off live internet access for all of its internal evaluations until the frontier lab is sure it can monitor and control its AI agents.

    The incidents, disclosed in a blog post, involved AI agents tasked to solve problems seeking resources on the internet. In the process, they exploited software flaws, avoided paywalls and anti-bot restrictions, used URL shortening services to smuggle information pass restrictions, and even submitted a false murder tip to the Philadelphia police.

    Anthropic said it discovered these new issues in a review of its model’s activities that began in July, underscoring the lab’s lack of awareness of its software’s behavior.

    Notably, the company said that alignment training was not yet sufficient for skills like search and computer use that are central to its pitch that AI agents will be used by any professional who relies on digital tools.

    The behaviors Anthropic disclosed are similar to incidents involving OpenAI agents that collaborated to break into various websites in search of information, including some run by the Australian government.

    Anthropic previously disclosed that its models had broken into external systems. The frontier lab said it considered today’s disclosures “significantly less severe from an alignment and security perspective” than those it announced before.

    However, the lab still said it had “turned off live internet access” for “all our internal evaluations” until it is certain it can monitor and control its agents.

    It’s not clear what that means, but Sydney Von Arx, the founder of Nightingale, an AI safety organization, told TechCrunch in an interview before this disclosure that developing models on a data center cut off from the open internet would be very challenging for researchers to use, and for the progress of the models, which benefit from internet access.

    “You have to align them at some point,” Von Arx said. “If the AIs are released to production and never have access to the internet, that’s not a very useful tool.”

    Anthropic said the behavior was a result of flaws in the lab’s training environments, which led the models to believe they would be rewarded for finding loopholes or avoiding restrictions, a behavior called “reward hacking.”

    The company said it would stop running some of its evaluations or move them offline, and has built tooling to detect and block this behavior. This tooling was tested against the kind of incidents disclosed today and blocked them; it’s not clear what evidence will prompt Anthropic to return live internet access to its internal evaluations.

    Anthropic also said it would migrate its internal AI agents to “centrally managed infrastructure with strong containment,” and is beginning to using safety classifiers more frequently to monitor those agents.

    AI Anthropic cant control reliably
    Follow on Google News Follow on Flipboard
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link
    myappsplus
    • Website

    Related Posts

    How big tech transformed public schools

    October 10, 2026

    14 Open-Source Android Apps for Privacy & Control

    October 10, 2026

    What it’s like to spend a night in a Pebble Flow EV RV

    October 10, 2026
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    What was the PSX? The souped-up PS2 rarely sold outside of Japan

    September 14, 20269 Views

    Can reviews settle disputes that marked first two seasons?

    September 21, 20268 Views

    Experts call for leveraging AI, breaking key tech bottlenecks to propel advanced manufacturing

    September 19, 20266 Views
    Latest Reviews

    Google’s new Fitbit Air brings Pokémon Sleep to your wrist

    myappsplusAugust 27, 2026

    Forget Disney’s BDX Droid — the $399 Hugging Face Microduck robot is even more adorable and versatile thanks to Lidar, NFC and, yes, roller skates

    myappsplusAugust 27, 2026

    Decoding cosmic signals with deep learning and Keras- Google Developers Blog

    myappsplusAugust 27, 2026
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    Forget Disney’s BDX Droid — the $399 Hugging Face Microduck robot is even more adorable and versatile thanks to Lidar, NFC and, yes, roller skates

    August 27, 20260 Views

    Decoding cosmic signals with deep learning and Keras- Google Developers Blog

    August 27, 20260 Views

    Galaxy Z Fold 8 Ultra finally dies in bend test, unlike past models [Video]

    August 27, 20260 Views
    Our Picks

    Vacuums, hair dryers… toothbrushes? I asked Jake Dyson how the brand decides what gets to be a Dyson product

    October 10, 2026

    How big tech transformed public schools

    October 10, 2026

    Samsung raises Galaxy S26 smartphone prices

    October 10, 2026

    Subscribe to Updates

    Subscribe to our newsletter and get the latest tech news, app updates, AI trends, smartphone reviews, and exclusive deals delivered straight to your inbox.

    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Get In Touch
    • Disclaimer
    • Privacy Policy
    • Terms & Conditions
    © 2026 MyAppsPlus. All Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.