Close Menu
GeekBlog

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    Rockstar Showed 26 Minutes of GTA 6. Fans Watched and Decided the Leaker Had Been Lying.

    August 28, 2026

    NASA Has Closed the Book on 2024 YR4. Rocks That Size Pass Us a Few Times a Year.

    August 28, 2026

    Apple Maps Now Has Ads. The Off Switch You Are Looking For Does Not Exist.

    August 28, 2026
    Facebook X (Twitter) Instagram Threads
    GeekBlog
    • Home
    • Mobile
    • Tech News
    • Blog
    • How-To Guides
    • AI & Software
    Facebook
    GeekBlog
    Home»Tech News»OpenAI’s research on AI models deliberately lying is wild 
    Tech News

    OpenAI’s research on AI models deliberately lying is wild 

    Michael ComaousBy Michael ComaousSeptember 18, 20254 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link
    OpenAI’s research on AI models deliberately lying is wild 
    Share
    Facebook Twitter LinkedIn Pinterest Email Copy Link

    Every now and then, researchers at the biggest tech companies drop a bombshell. There was the time Google said its latest quantum chip indicated multiple universes exist. Or when Anthropic gave its AI agent Claudius a snack vending machine to run and it went amok, calling security on people, and insisting it was human.  

    This week, it was OpenAI’s turn to raise our collective eyebrows.

    OpenAI released on Monday some research that explained how it’s stopping AI models from “scheming.” It’s a practice in which an “AI behaves one way on the surface while hiding its true goals,” OpenAI defined in its tweet about the research.   

    In the paper, conducted with Apollo Research, researchers went a bit further, likening AI scheming to a human stock broker breaking the law to make as much money as possible. The researchers, however, argued that most AI “scheming” wasn’t that harmful. “The most common failures involve simple forms of deception — for instance, pretending to have completed a task without actually doing so,” they wrote. 

    The paper was mostly published to show that “deliberative alignment⁠” — the anti-scheming technique they were testing — worked well. 

    But it also explained that AI developers haven’t figured out a way to train their models not to scheme. That’s because such training could actually teach the model how to scheme even better to avoid being detected. 

    “A major failure mode of attempting to ‘train out’ scheming is simply teaching the model to scheme more carefully and covertly,” the researchers wrote. 

    Techcrunch event

    San Francisco
    |
    October 27-29, 2025

    Recommended for you:

    Two of the Kremlin’s most active hack groups are collaborating, ESET says
    Tech News·Sep 19, 2025

    Two of the Kremlin’s most active hack groups are collaborating, ESET says

    Perhaps the most astonishing part is that, if a model understands that it’s being tested, it can pretend it’s not scheming just to pass the test, even if it is still scheming. “Models often become more aware that they are being evaluated. This situational awareness can itself reduce scheming, independent of genuine alignment,” the researchers wrote. 

    It’s not news that AI models will lie. By now most of us have experienced AI hallucinations, or the model confidently giving an answer to a prompt that simply isn’t true. But hallucinations are basically presenting guesswork with confidence, as OpenAI research released earlier this month documented. 

    Scheming is something else. It’s deliberate.  

    Even this revelation — that a model will deliberately mislead humans — isn’t new. Apollo Research first published a paper in December documenting how five models schemed when they were given instructions to achieve a goal “at all costs.”  

    What is? Good news that the researchers saw significant reductions in scheming by using “deliberative alignment⁠.” That technique involves teaching the model an “anti-scheming specification” and then making the model go review it before acting. It’s a little like making little kids repeat the rules before allowing them to play. 

    OpenAI researchers insist that the lying they’ve caught with their own models, or even with ChatGPT, isn’t that serious. As OpenAI’s co-founder Wojciech Zaremba told TechCrunch’s Maxwell Zeff when calling for better safety-testing: “This work has been done in the simulated environments, and we think it represents future use cases. However, today, we haven’t seen this kind of consequential scheming in our production traffic. Nonetheless, it is well known that there are forms of deception in ChatGPT. You might ask it to implement some website, and it might tell you, ‘Yes, I did a great job.” And that’s just the lie. There are some petty forms of deception that we still need to address.”

    The fact that AI models from multiple players intentionally deceive humans is, perhaps, understandable. They were built by humans, to mimic humans and (synthetic data aside) for the most part trained on data produced by humans. 

    Recommended for you:

    Donald Trump Is Saying There’s a TikTok Deal. China Isn’t
    Tech News·Sep 19, 2025

    Donald Trump Is Saying There’s a TikTok Deal. China Isn’t

    It’s also bonkers. 

    While we’ve all experienced the frustration of poorly performing technology (thinking of you, home printers of yesteryear), when was the last time your not-AI software deliberately lied to you? Has your inbox ever fabricated emails on its own? Has your CMS logged new prospects that didn’t exist to pad its numbers? Has your fintech app made up its own bank transactions? 

    It’s worth pondering this as the corporate world barrels towards an AI future where companies believe agents can be treated like independent employees. The researchers of this paper have the same warning.

    “As AIs are assigned more complex tasks with real-world consequences and begin pursuing more ambiguous, long-term goals, we expect that the potential for harmful scheming will grow — so our safeguards and our ability to rigorously test must grow correspondingly,” they wrote. 

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Telegram Email Copy Link
    Previous ArticleAmazon October Prime Day 2025: how to find the best deals
    Next Article Ready to download iOS 26? See if your iPhone is eligible for the free update first
    Michael Comaous
    • Website

    Michael Comaous is a dedicated professional with a passion for technology, innovation, and creative problem-solving. Over the years, he has built experience across multiple industries, combining strategic thinking with hands-on expertise to deliver meaningful results. Michael is known for his curiosity, attention to detail, and ability to explain complex topics in a clear and approachable way. Whether he’s working on new projects, writing, or collaborating with others, he brings energy and a forward-thinking mindset to everything he does.

    Related Posts

    7 Mins Read

    Rockstar Showed 26 Minutes of GTA 6. Fans Watched and Decided the Leaker Had Been Lying.

    8 Mins Read

    NASA Has Closed the Book on 2024 YR4. Rocks That Size Pass Us a Few Times a Year.

    8 Mins Read

    Apple Maps Now Has Ads. The Off Switch You Are Looking For Does Not Exist.

    8 Mins Read

    Trump Declared a Power Grid Emergency. The Parts It Restricts Already Have a 128 Week Wait.

    9 Mins Read

    Amazon Coined “Artificial Artificial Intelligence” in 2005. Real AI Just Made It Obsolete.

    8 Mins Read

    A Lab Turned a Plastic Bottle Into a Cookie. Nobody Has Been Allowed to Eat One Yet.

    Top Posts

    AliExpress Was Playing Silent Sound Through Your Speakers to Work Out Who You Are

    August 24, 20262 Views

    How to Fix PS5 Controller Stick Drift (2026): 7 Working Methods

    July 10, 20262 Views

    Best AI Video Generators in 2026: Tested and Compared

    July 10, 20262 Views
    Stay In Touch
    • Facebook

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    Best Stores for Buying MP3 and Digital Music You Can Keep Forever (2026)

    August 2, 2025932 Views

    Discord will require a face scan or ID for full access next month

    February 9, 2026770 Views

    Trade in your old phone and get up to $1,100 off a new iPhone 17 at AT&T – here’s how

    September 10, 2025383 Views
    Our Picks

    Rockstar Showed 26 Minutes of GTA 6. Fans Watched and Decided the Leaker Had Been Lying.

    August 28, 2026

    NASA Has Closed the Book on 2024 YR4. Rocks That Size Pass Us a Few Times a Year.

    August 28, 2026

    Apple Maps Now Has Ads. The Off Switch You Are Looking For Does Not Exist.

    August 28, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    HEICJPG.online - Convert HEIC to JPG online
    Facebook
    • About Us
    • Contact us
    • Privacy Policy
    • Disclaimer
    • Terms and Conditions
    © 2026 GeekBlog

    Type above and press Enter to search. Press Esc to cancel.