Close Menu

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    The best game about an unhinged goose is just $7 on Steam right now

    HP OmniBook 5 14 review: Over 25 hours of battery power

    I don’t need AI in Windows. I need an operating system that works

    Facebook X (Twitter) Instagram
    • Artificial Intelligence
    • Business Technology
    • Cryptocurrency
    • Gadgets
    • Gaming
    • Health
    • Software and Apps
    • Technology
    Facebook X (Twitter) Instagram Pinterest Vimeo
    Tech AI Verse
    • Home
    • Artificial Intelligence

      Blue-collar jobs are gaining popularity as AI threatens office work

      August 17, 2025

      Man who asked ChatGPT about cutting out salt from his diet was hospitalized with hallucinations

      August 15, 2025

      What happens when chatbots shape your reality? Concerns are growing online

      August 14, 2025

      Scientists want to prevent AI from going rogue by teaching it to be bad first

      August 8, 2025

      AI models may be accidentally (and secretly) learning each other’s bad behaviors

      July 30, 2025
    • Business

      Why Certified VMware Pros Are Driving the Future of IT

      August 24, 2025

      Murky Panda hackers exploit cloud trust to hack downstream customers

      August 23, 2025

      The rise of sovereign clouds: no data portability, no party

      August 20, 2025

      Israel is reportedly storing millions of Palestinian phone calls on Microsoft servers

      August 6, 2025

      AI site Perplexity uses “stealth tactics” to flout no-crawl edicts, Cloudflare says

      August 5, 2025
    • Crypto

      Circle Partners With Finastra on $5 Trillion USDC Settlement

      August 28, 2025

      US and China Are Laundering Europeans’ Personal Data — Is Blockchain the Fix?

      August 28, 2025

      Does Coinbase’s New Hiring Policy Contradict US Federal Law?

      August 28, 2025

      Nvidia Earnings Report Shows Record Revenues Despite Zero Sales in China

      August 28, 2025

      One Sleuth Sounds The Alarm: Crypto Scam Prevention Isn’t Working

      August 28, 2025
    • Technology

      The best game about an unhinged goose is just $7 on Steam right now

      August 28, 2025

      HP OmniBook 5 14 review: Over 25 hours of battery power

      August 28, 2025

      I don’t need AI in Windows. I need an operating system that works

      August 28, 2025

      A new cloud storage doesn’t charge monthly fees, and their 1TB plan just went on sale

      August 28, 2025

      Windows 11 Pro is normally $199, but right now, it’s only $13

      August 28, 2025
    • Others
      • Gadgets
      • Gaming
      • Health
      • Software and Apps
    Check BMI
    Tech AI Verse
    You are at:Home»Technology»AI search engines fail accuracy test, study finds 60% error rate
    Technology

    AI search engines fail accuracy test, study finds 60% error rate

    TechAiVerseBy TechAiVerseMarch 12, 2025No Comments4 Mins Read5 Views
    Facebook Twitter Pinterest Telegram LinkedIn Tumblr Email Reddit
    AI search engines fail accuracy test, study finds 60% error rate
    Share
    Facebook Twitter LinkedIn Pinterest WhatsApp Email

    BMI Calculator – Check your Body Mass Index for free!

    AI search engines fail accuracy test, study finds 60% error rate

    Serving tech enthusiasts for over 25 years.

    TechSpot means tech analysis and advice you can trust.

    In context: It is a foregone conclusion that AI models can lack accuracy. Hallucinations and doubling down on wrong information have been an ongoing struggle for developers. Usage varies so much in individual use cases that it’s hard to nail down quantifiable percentages related to AI accuracy. A research team claims it now has those numbers.

    The Tow Center for Digital Journalism recently studied eight AI search engines, including ChatGPT Search, Perplexity, Perplexity Pro, Gemini, DeepSeek Search, Grok-2 Search, Grok-3 Search, and Copilot. They tested each for accuracy and recorded how frequently the tools refused to answer.

    The researchers randomly chose 200 news articles from 20 news publishers (10 each). They ensured each story returned within the top three results in a Google search when using a quoted excerpt from the article. Then, they performed the same query within each AI search tool and graded accuracy based on whether the search correctly cited A) the article, B) the news organization, and C) the URL.

    The researchers then labeled each search based on degrees of accuracy from “completely correct” to “completely incorrect.” As you can see from the diagram below, other than both versions of Perplexity, the AIs did not perform well. Collectively, AI search engines are inaccurate 60 percent of the time. Furthermore, these wrong results were reinforced by the AI’s “confidence” in them.

    Click to enlarge.

    The study is fascinating because it quantifiably confirms what we have known for a few years – that LLMs are “the slickest con artists of all time.” They report with complete authority that what they say is true even when it is not, sometimes to the point of argument or making up other false assertions when confronted.

    In a 2023 anecdotal article, Ted Gioia (The Honest Broker) pointed out dozens of ChatGPT responses, showing that the bot confidently “lies” when responding to numerous queries. While some examples were adversarial queries, many were just general questions.

    “If I believed half of what I heard about ChatGPT, I could let it take over The Honest Broker while I sit on the beach drinking margaritas and searching for my lost shaker of salt,” Gioia flippantly noted.

    Even when admitting it was wrong, ChatGPT would follow up that admission with more fabricated information. The LLM is seemingly programmed to answer every user input at all costs. The researcher’s data confirms this hypothesis, noting that ChatGPT Search was the only AI tool that answered all 200 article queries. However, it only achieved a 28-percent completely accurate rating and was completely inaccurate 57 percent of the time.

    ChatGPT isn’t even the worst of the bunch. Both versions of X’s Grok AI performed poorly, with Grok-3 Search being 94 percent inaccurate. Microsoft’s Copilot was not that much better when you consider that it declined to answer 104 queries out of 200. Of the remaining 96, only 16 were “completely correct,” 14 were “partially correct,” and 66 were “completely incorrect,” making it roughly 70 percent inaccurate.

    Arguably, the craziest thing about all this is that the companies making these tools are not transparent about this lack of accuracy while charging the public $20 to $200 per month to access their latest AI models. Moreover, Perplexity Pro ($20/month) and Grok-3 Search ($40/month) answered slightly more queries correctly than their free versions (Perplexity and Grok-2 Search) but had significantly higher error rates (above). Talk about a con.

    However, not everyone agrees. TechRadar’s Lance Ulanoff said he might never use Google again after trying ChatGPT Search. He describes the tool as fast, aware, and accurate, with a clean, ad-free interface.

    Feel free to read all the details in the Tow Center’s paper published in the Columbia Journalism Review, and let us know what you think.

    Do you trust AI search engines to return accurate results?

    BMI Calculator – Check your Body Mass Index for free!

    Share. Facebook Twitter Pinterest LinkedIn Reddit WhatsApp Telegram Email
    Previous ArticleApple Mac Studio gets reviewed: incredibly fast, incredibly expensive
    Next Article Seagate hard drive scammers figure out how to compromise FARM metrics
    TechAiVerse
    • Website

    Jonathan is a tech enthusiast and the mind behind Tech AI Verse. With a passion for artificial intelligence, consumer tech, and emerging innovations, he deliver clear, insightful content to keep readers informed. From cutting-edge gadgets to AI advancements and cryptocurrency trends, Jonathan breaks down complex topics to make technology accessible to all.

    Related Posts

    The best game about an unhinged goose is just $7 on Steam right now

    August 28, 2025

    HP OmniBook 5 14 review: Over 25 hours of battery power

    August 28, 2025

    I don’t need AI in Windows. I need an operating system that works

    August 28, 2025
    Leave A Reply Cancel Reply

    Top Posts

    Ping, You’ve Got Whale: AI detection system alerts ships of whales in their path

    April 22, 2025166 Views

    6.7 Cummins Lifter Failure: What Years Are Affected (And Possible Fixes)

    April 14, 202548 Views

    New Akira ransomware decryptor cracks encryptions keys using GPUs

    March 16, 202530 Views

    Is Libby Compatible With Kobo E-Readers?

    March 31, 202527 Views
    Don't Miss
    Technology August 28, 2025

    The best game about an unhinged goose is just $7 on Steam right now

    The best game about an unhinged goose is just $7 on Steam right now Image:…

    HP OmniBook 5 14 review: Over 25 hours of battery power

    I don’t need AI in Windows. I need an operating system that works

    A new cloud storage doesn’t charge monthly fees, and their 1TB plan just went on sale

    Stay In Touch
    • Facebook
    • Twitter
    • Pinterest
    • Instagram
    • YouTube
    • Vimeo

    Subscribe to Updates

    Get the latest creative news from SmartMag about art & design.

    About Us
    About Us

    Welcome to Tech AI Verse, your go-to destination for everything technology! We bring you the latest news, trends, and insights from the ever-evolving world of tech. Our coverage spans across global technology industry updates, artificial intelligence advancements, machine learning ethics, and automation innovations. Stay connected with us as we explore the limitless possibilities of technology!

    Facebook X (Twitter) Pinterest YouTube WhatsApp
    Our Picks

    The best game about an unhinged goose is just $7 on Steam right now

    August 28, 20252 Views

    HP OmniBook 5 14 review: Over 25 hours of battery power

    August 28, 20252 Views

    I don’t need AI in Windows. I need an operating system that works

    August 28, 20252 Views
    Most Popular

    Xiaomi 15 Ultra Officially Launched in China, Malaysia launch to follow after global event

    March 12, 20250 Views

    Apple thinks people won’t use MagSafe on iPhone 16e

    March 12, 20250 Views

    French Apex Legends voice cast refuses contracts over “unacceptable” AI clause

    March 12, 20250 Views
    © 2025 TechAiVerse. Designed by Divya Tech.
    • Home
    • About Us
    • Contact Us
    • Privacy Policy
    • Terms & Conditions

    Type above and press Enter to search. Press Esc to cancel.