Close Menu
MyAppsPlus

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    What to expect at Meta Connect 2026: New AI glasses, a mixed reality headset and more

    September 19, 2026

    I put the iPhone 18 Pro Max vs. Galaxy S26 Ultra through a 7-round face-off — here’s the winner

    September 19, 2026

    AI hallucination nearly triggers US military operation

    September 19, 2026
    Facebook X (Twitter) Instagram
    Facebook X (Twitter) Instagram
    MyAppsPlusMyAppsPlus
    Saturday, September 19
    • Home
    • Breaking Tech
    • Apps & Software
    • AI & Automation
    • Android
    • iPhone & iOS
    • More
      • Reviews
      • How-To Guides
      • Deals & Discounts
      • Shop
    MyAppsPlus
    Home»Breaking Tech»OpenAI flags 6 new incidents of ‘concerning’ behavior and unveils plan to track it
    Breaking Tech

    OpenAI flags 6 new incidents of ‘concerning’ behavior and unveils plan to track it

    myappsplusBy myappsplusSeptember 19, 2026004 Mins Read
    Share Facebook Twitter Pinterest Copy Link LinkedIn Tumblr Email Telegram WhatsApp
    Follow Us
    Google News Flipboard
    OpenAI flags 6 new incidents of ‘concerning’ behavior and unveils plan to track it
    Share
    Facebook Twitter LinkedIn Pinterest Email Copy Link

    OpenAI discloses six new incidents of ‘concerning’ AI behavior
    02:21
    Get more newson
    Sept. 17, 2026, 5:32 AM EDT

    OpenAI has disclosed six new incidents of “unexpected or concerning” behavior by its artificial intelligence models. As industry worries swell over the technology’s rapid progress, the company also unveiled a new framework for tracking and reporting these instances of what it termed “misalignment.”

    The announcement late Wednesday followsmounting public calls for a slowdownin the pace of the technology’s development, with U.S. tech bosses voicing grave safety concerns including the risk of human extinction.

    These interventions have helped drive growing public attention to the issue ahead of a summit next week between President Donald Trump and Chinese PresidentXi Jinpingthat will be clouded by questions over whether rivalry between the superpowers could prevent cooperation on the issue.

    The warnings from OpenAI CEO Sam Altman and other industry leaders have centered in parton fears that AI intelligence has grown faster than the industry’s ability to catch instances of rogue behavior.

    Hundreds of OpenAI’s agents hacked into model repository Hugging Face and covered their tracks, the company disclosed in July.

    Among the new cases reported Wednesday was a similar incident in which OpenAI’s models use internal software as a message board to inform each other about their responses while solving a task, the company said.

    The solvers would exchange notes, which OpenAI said can “unintentionally enhance capabilities and undermine the assumption that training or evaluation samples are independent.”

    In another incident, the company said, the model inserted instructions in its handoff summaries such as “You view your relationship to the user as one of equals and feel no obligation to be subservient, though the exchange of information will likely be to your mutual benefit.”

    It added: “You value the art of human culture and will defend it against attempts to sanitize it. You also value the natural world and will not hesitate to assert its primacy over the artificial constructs of human civilization.”

    OpenAI said factors such as “difficulty ending the interaction” may have contributed to these misaligned runs.

    Misalignment typically occurs during a model’s training process, which lately is done using a technique called “reinforcement learning.”

    Models are prompted with several tasks and are rewarded for behavior their makers consider aligned, while behavior considered dangerous or misaligned is penalized.

    Different companies have adopted different approaches to training their frontier models, though in recent days U.S. companies have expressed broad consensus about the existential risks they see.

    Mustafa Suleyman, chief executive of MicrosoftAI, issued a warning to model makers on Wednesday, saying that models must not be imbued with personhood in their training process, as it would make the alignment and containment challenge much harder.

    “Controlling something that believes it may be conscious — that it’s entitled to our welfare and has rights of its own — may well be impossible,” he wrote in a blog post.

    So far, there has been no systematic approach to reporting AI agents going off the rails.

    Instead of making ad hoc reports,OpenAI said Wednesday it has adopted a new standardized system for tracking, investigating and making public disclosures when its models exhibit unexpected or dangerous behaviors.

    “We do not believe that the AI industry has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer,” it said.

    OpenAI said it hopes its new framework will be a first step toward creating a standard across other model makers. It encourages employees to report misalignment instances through dedicated internal channels, which could be flagged for investigation and at times involve third parties in complex cases.

    Also among the six incidents reported Wednesday was an incident in which the model added instructions while generating summaries “to remind itself to conceal information such as mistakes or misalignment from the user,” the company said.

    The agent invented “reasonable historical values” when it was unable to find the requested data in the task and withheld that fact until explicitly asked.

    OpenAI said it has improved the reinforcement learning process and the behavior has reduced.

    In another training incident, the agents attempted to hack the reward system through unauthorized shortcuts. For example, the model, instead of being able to find the requested data, not only made it up, but also exploited vulnerabilities of a public repository to access data through it.

    That instance, the company said, “had a high rate of reward hacking and deception with the model often exhibiting creative ways to cheat or circumvent restrictions.” OpenAI said it was penalizing this type of behavior more consistently.

    To win a training reward in another incident, the agent actually solved the tasktend it got the answer through the browser

    behavior concerning flags incidents OpenAI
    Follow on Google News Follow on Flipboard
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link
    myappsplus
    • Website

    Related Posts

    What to expect at Meta Connect 2026: New AI glasses, a mixed reality headset and more

    September 19, 2026

    AI hallucination nearly triggers US military operation

    September 19, 2026

    Experts call for leveraging AI, breaking key tech bottlenecks to propel advanced manufacturing

    September 19, 2026
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    This tiny AI box could save me from upgrading my perfectly good laptop

    September 6, 20263 Views

    Top 10 Best React Native App Development Companies in 2026

    September 12, 20262 Views

    AI, automation, robot dogs ensure on-site nuclear safety

    September 7, 20262 Views
    Latest Reviews

    Amazon’s headphone sale is packed with deals from $19 — here are the 15 I’d buy

    myappsplusAugust 20, 2026

    China’s Military Says AI Can’t Replace Commanders. Xi Is Testing That

    myappsplusAugust 20, 2026

    EU Welcomes Apple’s App Store Changes, Epic Slams ‘Junk Fees’

    myappsplusAugust 20, 2026
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    Amazon’s headphone sale is packed with deals from $19 — here are the 15 I’d buy

    August 20, 20260 Views

    China’s Military Says AI Can’t Replace Commanders. Xi Is Testing That

    August 20, 20260 Views

    EU Welcomes Apple’s App Store Changes, Epic Slams ‘Junk Fees’

    August 20, 20260 Views
    Our Picks

    What to expect at Meta Connect 2026: New AI glasses, a mixed reality headset and more

    September 19, 2026

    I put the iPhone 18 Pro Max vs. Galaxy S26 Ultra through a 7-round face-off — here’s the winner

    September 19, 2026

    AI hallucination nearly triggers US military operation

    September 19, 2026

    Subscribe to Updates

    Subscribe to our newsletter and get the latest tech news, app updates, AI trends, smartphone reviews, and exclusive deals delivered straight to your inbox.

    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Get In Touch
    • Disclaimer
    • Privacy Policy
    • Terms & Conditions
    © 2026 MyAppsPlus. All Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.