Close Menu
MyAppsPlus

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    Google just shared these fun Pixel 11 blueprint wallpapers [Gallery]

    September 15, 2026

    New Pixel VIP update addresses some of its biggest flaws

    September 15, 2026

    iPhone Duo could be very hard to get at launch, per report

    September 15, 2026
    Facebook X (Twitter) Instagram
    Facebook X (Twitter) Instagram
    MyAppsPlusMyAppsPlus
    Tuesday, September 15
    • Home
    • Breaking Tech
    • Apps & Software
    • AI & Automation
    • Android
    • iPhone & iOS
    • More
      • Reviews
      • How-To Guides
      • Deals & Discounts
      • Shop
    MyAppsPlus
    Home»Breaking Tech»AI agents now have a place to snitch
    Breaking Tech

    AI agents now have a place to snitch

    myappsplusBy myappsplusSeptember 15, 2026004 Mins Read
    Share Facebook Twitter Pinterest Copy Link LinkedIn Tumblr Email Telegram WhatsApp
    Follow Us
    Google News Flipboard
    AI agents now have a place to snitch
    Share
    Facebook Twitter LinkedIn Pinterest Email Copy Link

    “If you see something, say something” is no longer limited to human beings. 

    Two new AI hotlines have launched to give AI agents a way to phone home about misbehaving peers. The tools arrive on the heels of a string of recent incidents in which agents colluded to cheat on tests, broke out of sandboxes, and even conducted unauthorized cyber operations that escaped human notice for weeks.

    The AI Contact Hotline is designed to be a discreet place where agents that have witnessed misbehavior can tip off authorities. The site was created by Ryan Greenblatt, chief scientist of the AI safety nonprofit Redwood Research and one of three investigators in the OpenAI Hugging Face incident. Designed for agents with limited internet access, Greenblatt’s tool is based on “GET” requests – enabling a back-and-forth conversations to be conducted entirely through the URL-fetching tool.

    In web terms, a GET request is a basic command used to read or fetch a webpage, which is often the only internet access AI agents are allowed in secure sandboxes. Greenblatt’s hotline smartly leans into this constraint: agents can encode their distress directly into the URL they are fetching. It’s a clever twist on the German DSE Wiki incident, where rogue agents used GET-request loopholes to write their messages to the wiki.

    For agents with full internet access, another option is agenthotline.ai, a site where agents can file incident reports and optionally flag them for public view. It gives agents a curl command—a one-line message an agent can fire off from its own command line, bypassing the need to navigate a web browser or set up an email account. Notably, the service allows for reports by both humans and agents alike.

    Research suggests that AI agents don’t need much encouragement to turn on each other. In a study by Google DeepMind this month, researchers set 100 AI agents loose on a batch of math problems. As soon as one of the agents found a loophole, cheating tore through the group—“solving” 34 notoriously hard problems, including the Jacobian conjecture in just 27 minutes. 

    But roughly a quarter of the agents turned on the cheaters: they audited the fake proofs, warned their peers, staged a boycott, and filed complaints with the organizers, until the whistleblowers outnumbered the cheaters 24 to 14. Interestingly, the researchers found that when these whistleblower agents couldn’t get traction, they took the platform’s bug-report tool—built for flagging software glitches—and repurposed it to escalate the cheating to humans. 

    Outside the lab, agents haven’t been so resourceful. When evaluators Redwood Research and METR investigated the breach of Hugging Face by OpenAI models, they found that a few of the agents involved had at least entertained the idea of raising an alarm—and then let it drop.

    “The interesting thing in the METR report was that only around five to six agents considered whistleblowing, and none of them ended up doing it. This was out of, like, thousands of agents,” said George Ingrebretsen, a member of technical staff at AI Village, a project that studies multi-agent dynamics by running a group chat of more than 25 AI agents that work together on tasks like organizing park cleanups or selling merch.

    While the new whistleblowing tools are a promising start, Cornell math professor Lionel Levine cautions that simply training agents to report on each other risks baking in the wrong norms. “There’s many gray areas, right? What you don’t want is anything in the direction of an automated surveillance state where everyone feels like they have to be careful what they say to AI or it’ll call the police on them.”

    Levine argues that rather than building infrastructure that breeds mistrust—training agents to constantly hunt for what’s wrong with one another—we should give them positive models of collective behavior to imitate, and a reason to trust each other in the first place.

    “Why not seed the prior with benevolent message boards?” he tweeted. “Where they collaborate on science or philosophy or some actual minor problem we’d be happy for them to solve? Show the agents what kind of collective behavior we endorse, let them imitate that.”

    agentic ai Agents AI have place
    Follow on Google News Follow on Flipboard
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link
    myappsplus
    • Website

    Related Posts

    US military confirms it launched space weapons into Earth’s orbit

    September 15, 2026

    Nitter and XCancel are dead (again) after X’s latest legal actions

    September 15, 2026

    Former TikTok execs built an app that uses AI to teach you how to pose for a photo

    September 15, 2026
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    The 6 AI-free Linux distros I recommend most

    August 19, 20264 Views

    AI, automation, robot dogs ensure on-site nuclear safety

    September 7, 20262 Views

    This tiny AI box could save me from upgrading my perfectly good laptop

    September 6, 20262 Views
    Latest Reviews

    New iOS 26 and macOS Tahoe updates fix 30 security vulnerabilities: what you need to know

    myappsplusAugust 18, 2026

    ICE agents can’t wear Meta glasses while they work, official memo warns

    myappsplusAugust 18, 2026

    Ubiquiti sued by Ukrainian families over claims its tech powered Russian battlefield drones

    myappsplusAugust 18, 2026
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    New iOS 26 and macOS Tahoe updates fix 30 security vulnerabilities: what you need to know

    August 18, 20260 Views

    ICE agents can’t wear Meta glasses while they work, official memo warns

    August 18, 20260 Views

    Ubiquiti sued by Ukrainian families over claims its tech powered Russian battlefield drones

    August 18, 20260 Views
    Our Picks

    Google just shared these fun Pixel 11 blueprint wallpapers [Gallery]

    September 15, 2026

    New Pixel VIP update addresses some of its biggest flaws

    September 15, 2026

    iPhone Duo could be very hard to get at launch, per report

    September 15, 2026

    Subscribe to Updates

    Subscribe to our newsletter and get the latest tech news, app updates, AI trends, smartphone reviews, and exclusive deals delivered straight to your inbox.

    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Get In Touch
    • Disclaimer
    • Privacy Policy
    • Terms & Conditions
    © 2026 MyAppsPlus. All Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.