Close Menu
GeekBlog

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    A Toddler Needed a $20,000 Wheelchair. A High School Robotics Team Built Him One Instead.

    August 5, 2026

    AI’s Memory Boom Is Quietly Crushing the Budget Smartphone

    August 5, 2026

    Scientists Didn’t Say Earth Is Becoming Uninhabitable. Here’s What the “Hothouse Earth” Study Actually Says

    August 5, 2026
    Facebook X (Twitter) Instagram Threads
    GeekBlog
    • Home
    • Mobile
    • Tech News
    • Blog
    • How-To Guides
    • AI & Software
    Facebook
    GeekBlog
    Home»Tech News»Are bad incentives to blame for AI hallucinations?
    Tech News

    Are bad incentives to blame for AI hallucinations?

    Michael ComaousBy Michael ComaousSeptember 7, 20253 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link
    ChatGPT logo
    Share
    Facebook Twitter LinkedIn Pinterest Email Copy Link

    Recommended for you:

    AMD’s AI-powered FSR 4 upscaling is now available in most FSR 3.1 games
    Tech News·Sep 8, 2025

    AMD’s AI-powered FSR 4 upscaling is now available in most FSR 3.1 games

    A new research paper from OpenAI asks why large language models like GPT-5 and chatbots like ChatGPT still hallucinate, and whether anything can be done to reduce those hallucinations.

    In a blog post summarizing the paper, OpenAI defines hallucinations as “plausible but false statements generated by language models,” and it acknowledges that despite improvements, hallucinations “remain a fundamental challenge for all large language models” — one that will never be completely eliminated.

    To illustrate the point, researchers say that when they asked “a widely used chatbot” about the title of Adam Tauman Kalai’s Ph.D. dissertation, they got three different answers, all of them wrong. (Kalai is one of the paper’s authors.) They then asked about his birthday and received three different dates. Once again, all of them were wrong.

    How can a chatbot be so wrong — and sound so confident in its wrongness? The researchers suggest that hallucinations arise, in part, because of a pretraining process that focuses on getting models to correctly predict the next word, without true or false labels attached to the training statements: “The model sees only positive examples of fluent language and must approximate the overall distribution.”

    “Spelling and parentheses follow consistent patterns, so errors there disappear with scale,” they write. “But arbitrary low-frequency facts, like a pet’s birthday, cannot be predicted from patterns alone and hence lead to hallucinations.”

    The paper’s proposed solution, however, focuses less on the initial pretraining process and more on how large language models are evaluated. It argues that the current evaluation models don’t cause hallucinations themselves, but they “set the wrong incentives.”

    The researchers compare these evaluations to the kind of multiple choice tests random guessing makes sense, because “you might get lucky and be right,” while leaving the answer blank “guarantees a zero.” 

    Techcrunch event

    San Francisco
    |
    October 27-29, 2025

    “In the same way, when models are graded only on accuracy, the percentage of questions they get exactly right, they are encouraged to guess rather than say ‘I don’t know,’” they say.

    Recommended for you:

    Intel’s chief executive of products departs among other leadership changes
    Tech News·Sep 9, 2025

    Intel’s chief executive of products departs among other leadership changes

    The proposed solution, then, is similar to tests (like the SAT) that include “negative [scoring] for wrong answers or partial credit for leaving questions blank to discourage blind guessing.” Similarly, OpenAI says model evaluations need to “penalize confident errors more than you penalize uncertainty, and give partial credit for appropriate expressions of uncertainty.”

    And the researchers argue that it’s not enough to introduce “a few new uncertainty-aware tests on the side.” Instead, “the widely used, accuracy-based evals need to be updated so that their scoring discourages guessing.”

    “If the main scoreboards keep rewarding lucky guesses, models will keep learning to guess,” the researchers say.

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Telegram Email Copy Link
    Previous ArticleGoogle finally details Gemini usage limits
    Next Article This pettable Poké Ball is a Tamagotchi-style toy with over 150 Pokémon inside and I need it now
    Michael Comaous
    • Website

    Michael Comaous is a dedicated professional with a passion for technology, innovation, and creative problem-solving. Over the years, he has built experience across multiple industries, combining strategic thinking with hands-on expertise to deliver meaningful results. Michael is known for his curiosity, attention to detail, and ability to explain complex topics in a clear and approachable way. Whether he’s working on new projects, writing, or collaborating with others, he brings energy and a forward-thinking mindset to everything he does.

    Related Posts

    6 Mins Read

    AI’s Memory Boom Is Quietly Crushing the Budget Smartphone

    6 Mins Read

    Meta Wants an AI Agent Managing Your Life. Wall Street Isn’t So Sure

    6 Mins Read

    MakuluLinux’s New AI-OS Wants to Run Your Whole Desktop, Not Just Answer Questions

    7 Mins Read

    AI Tokens Got 98% Cheaper. Corporate AI Bills Are Exploding Anyway

    7 Mins Read

    Hackers Are Hijacking Hotel Wi-Fi to Steal Microsoft 365 Logins Without a Single Phishing Email

    5 Mins Read

    Microsoft’s Biggest Patch Tuesday Ever Just Showed Us Where Cybersecurity Is Heading

    Top Posts

    Best Stores for Buying MP3 and Digital Music You Can Keep Forever (2026)

    August 2, 202530 Views

    Every iPhone Camera Ranked in 2026 (Best to Worst)

    July 6, 202613 Views

    How to Fix PS5 Controller Stick Drift (2026): 7 Working Methods

    July 10, 20268 Views
    Stay In Touch
    • Facebook

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    Best Stores for Buying MP3 and Digital Music You Can Keep Forever (2026)

    August 2, 2025930 Views

    Discord will require a face scan or ID for full access next month

    February 9, 2026770 Views

    Trade in your old phone and get up to $1,100 off a new iPhone 17 at AT&T – here’s how

    September 10, 2025383 Views
    Our Picks

    A Toddler Needed a $20,000 Wheelchair. A High School Robotics Team Built Him One Instead.

    August 5, 2026

    AI’s Memory Boom Is Quietly Crushing the Budget Smartphone

    August 5, 2026

    Scientists Didn’t Say Earth Is Becoming Uninhabitable. Here’s What the “Hothouse Earth” Study Actually Says

    August 5, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Facebook
    • About Us
    • Contact us
    • Privacy Policy
    • Disclaimer
    • Terms and Conditions
    © 2026 GeekBlog

    Type above and press Enter to search. Press Esc to cancel.