Close Menu

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    And Daphne Koller’s Thoughts On The Future

    June 20, 2026

    Nobel Winner John Jumper to Leave Google DeepMind for Anthropic

    June 20, 2026

    DeepSeek Climbs US Corporate Expense Reports as Companies Chase Cheaper AI

    June 20, 2026
    Facebook X (Twitter) Instagram
    ailogicnews.aiailogicnews.ai
    • Home
    ailogicnews.aiailogicnews.ai
    Home»OpenAI»OpenAI’s latest AI models are smarter, but they make things up more often. Here’s what we know
    OpenAI

    OpenAI’s latest AI models are smarter, but they make things up more often. Here’s what we know

    AI Logic NewsBy AI Logic NewsApril 19, 2025No Comments2 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    Share
    Facebook Twitter LinkedIn Pinterest Email

    OpenAI launched its new 3 and o4 mini reasoning models on Wednesday with many new features. Some enthusiastic OpenAI employees even went on to state that o3 had is nearing Artificial General Intelligence (AGI) – a technical term which has no fixed definition but is usually meant to believe a stage when AI achieves near or equivalent level of intelligence as humans. However, as it turns a new document by OpenAI itself proves that its new AI models are prone to not just hallucination (making stuff up), but even more hallucinations than its previous reasoning and non-reasoning models. 

    OpenAI had first rolled out its reasoning model last year which claims to mimic human level thinking in order to solve for more complex queries. However, with its latest and most powerful reasoning model yet, OpenAI says that it can make ‘accurate’ and ‘inaccurate’ claims. 

    In its technical report for o3 and o4 mini (first reported by TechCrunch), OpenAI says “o3 tends to make more claims overall, leading to more accurate claims as well as more inaccurate/hallucinated claims. More research is needed to understand the cause of this result.”

    As per OpenAI’s own PersonaQA benchmark, the company found that o3 hallucinated while responding to around 33% of the questions compared to the 16% hallucination rate for o1 and 14.8% for o3 mini.

    Source link

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous ArticleMinistry of SMEs allocates 5 trilli
    Next Article From gifted student to AI hero, DeepSeek’s founder Liang Wenfeng inspires a nation
    AI Logic News

    Related Posts

    OpenAI

    OpenAI, Anthropic and the New Battle for A.I. Trust

    June 20, 2026
    OpenAI

    Inklings #021 📧

    June 19, 2026
    OpenAI

    OpenAI & Anthropic Accelerate Experienced Enterprise Sales Hires. AI-RTZ #1121

    June 19, 2026
    Demo
    Top Posts

    DeepSeek V4 And Tencent’s New Hunyuan Model To Launch In April

    March 17, 202647 Views

    OpenAI’s Simo Said to Warn Staff Ag

    March 17, 202640 Views

    Hunter Alpha Sparks DeepSeek V4 Speculation

    March 18, 202632 Views
    Latest Reviews
    ailogicnews.ai
    © 2026 Lee Enterprises

    Type above and press Enter to search. Press Esc to cancel.