Close Menu

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    New Nvidia N1X and Nvidia N1V laptops revealed

    Microsoft executive addresses whether all Xbox games will launch day one on PS5 like Fable

    Citizen launches new blue chronograph with Caliber B620 Eco-Drive movement

    Facebook X (Twitter) Instagram
    • Artificial Intelligence
    • Business Technology
    • Cryptocurrency
    • Gadgets
    • Gaming
    • Health
    • Software and Apps
    • Technology
    Facebook X (Twitter) Instagram Pinterest Vimeo
    Tech AI Verse
    • Home
    • Artificial Intelligence

      ChatGPT can embrace authoritarian ideas after just one prompt, researchers say

      January 24, 2026

      Ashley St. Clair, the mother of one of Elon Musk’s children, sues xAI over Grok sexual images

      January 17, 2026

      Anthropic joins OpenAI’s push into health care with new Claude tools

      January 12, 2026

      The mother of one of Elon Musk’s children says his AI bot won’t stop creating sexualized images of her

      January 7, 2026

      A new pope, political shake-ups and celebs in space: The 2025-in-review news quiz

      December 31, 2025
    • Business

      New VoidLink malware framework targets Linux cloud servers

      January 14, 2026

      Nvidia Rubin’s rack-scale encryption signals a turning point for enterprise AI security

      January 13, 2026

      How KPMG is redefining the future of SAP consulting on a global scale

      January 10, 2026

      Top 10 cloud computing stories of 2025

      December 22, 2025

      Saudia Arabia’s STC commits to five-year network upgrade programme with Ericsson

      December 18, 2025
    • Crypto

      Solana’s Privacy Coin Jumps 60% After New Cross-Chain Swap Reveal

      January 25, 2026

      Did Axie Infinity (AXS) Whales Just Buy Into a Pullback Risk After a 41% Rally?

      January 25, 2026

      Kraken’s Breakout Acquisition Signals Institutional Bet on Crypto Prop Trading’s Explosive Growth

      January 25, 2026

      Smart Money Exit Solana’s Seeker Token after 200% Rally

      January 25, 2026

      Zcash Bear Trap Active After 15% Rebound: What’s Next for ZEC Price?

      January 25, 2026
    • Technology

      New Nvidia N1X and Nvidia N1V laptops revealed

      January 25, 2026

      Microsoft executive addresses whether all Xbox games will launch day one on PS5 like Fable

      January 25, 2026

      Citizen launches new blue chronograph with Caliber B620 Eco-Drive movement

      January 25, 2026

      Super Mario Bros. Wonder Switch 2 upgrade price renews criticism about costly Nintendo ecosystem

      January 25, 2026

      Retroid Pocket 6: reviews praise its performance and value, but knock some of its design choices

      January 25, 2026
    • Others
      • Gadgets
      • Gaming
      • Health
      • Software and Apps
    Check BMI
    Tech AI Verse
    You are at:Home»Technology»New study shows AI isn’t ready for office work
    Technology

    New study shows AI isn’t ready for office work

    TechAiVerseBy TechAiVerseJanuary 25, 2026No Comments5 Mins Read1 Views
    Facebook Twitter Pinterest Telegram LinkedIn Tumblr Email Reddit
    New study shows AI isn’t ready for office work
    Share
    Facebook Twitter LinkedIn Pinterest WhatsApp Email

    New study shows AI isn’t ready for office work

    Your job is safe for now as AI still struggles with real office tasks


    Levart_Photographer / Unsplash

    It has been nearly two years since Microsoft CEO Satya Nadella predicted that generative AI would take over knowledge work, but if you look around a typical law firm or investment bank today, the human workforce is still very much in charge. Despite all the hype about “reasoning” and “planning,” a new study from training-data company Mercor explains exactly why the robot revolution is stalled: AI just can’t handle the messiness of real work.

    A reality check for the “replacement” theory

    Mercor released a new benchmark called APEX-Agents, and it is brutal. Unlike the usual tests that ask AI to write a poem or solve a math problem, this one uses actual queries from lawyers, consultants, and bankers. It asks the models to do complete, multi-step tasks that require jumping between different types of information.

    Adobe Stock Image

    The results? Even the absolute best models on the market—we are talking about Gemini 3 Flash and GPT-5.2—couldn’t crack a 25% accuracy rate. Gemini led the pack at 24%, with GPT-5.2 right behind it at 23%. Most others were stuck in the teens.

    Why AI is failing the “office test”

    Mercor CEO Brendan Foody points out that the issue isn’t raw intelligence; it’s context. In the real world, answers aren’t served up on a silver platter. A lawyer has to check a Slack thread, read a PDF policy, look at a spreadsheet, and then synthesize all that to answer a question about GDPR compliance.

    Microsoft

    Humans do this context-switching naturally. AI, it turns out, is terrible at it. When you force these models to hunt for information across “scattered” sources, they either get confused, give the wrong answer, or just give up entirely.

    The “Unreliable Intern”

    For anyone worried about their job security, this is a bit of a relief. The study suggests that right now, AI functions less like a seasoned professional and more like an unreliable intern who gets things right about a quarter of the time.

    That said, the progress is terrifyingly fast. Foody noted that just a year ago, these models were scoring between 5% and 10%. Now they are hitting 24%. So, while they aren’t ready to take the wheel yet, they are learning to drive much faster than we expected. For now, though, the “knowledge work” revolution is on hold until the bots learn how to multitask p

    Moinak Pal is has been working in the technology sector covering both consumer centric tech and automotive technology for the…

    Google Research suggests AI models like DeepSeek exhibit collective intelligence patterns

    AI models are now holding meetings in their own heads

    It turns out that when the smartest AI models “think,” they might actually be hosting a heated internal debate. A fascinating new study co-authored by researchers at Google has thrown a wrench into how we traditionally understand artificial intelligence. It suggests that advanced reasoning models – specifically DeepSeek-R1 and Alibaba’s QwQ-32B – aren’t just crunching numbers in a straight, logical line. Instead, they appear to be behaving surprisingly like a group of humans trying to solve a puzzle together.

    The paper, published on arXiv with the evocative title Reasoning Models Generate Societies of Thought, posits that these models don’t merely compute; they implicitly simulate a “multi-agent” interaction. Imagine a boardroom full of experts tossing ideas around, challenging each other’s assumptions, and looking at a problem from different angles before finally agreeing on the best answer. That is essentially what is happening inside the code. The researchers found that these models exhibit “perspective diversity,” meaning they generate conflicting viewpoints and work to resolve them internally, much like a team of colleagues debating a strategy to find the best path forward.


    Read more

    Microsoft tells you to uninstall the latest Windows 11 update

    Microsoft says uninstall the January 2026 security update after POP email bugs and system issues surface.

    Microsoft has issued an unusual public advisory telling users to uninstall the Windows 11 January 2026 security update (KB5074109) after widespread reports that it is causing serious system and application issues. The update, which began rolling out automatically on January 13 and advances affected systems to OS Build 26200.7623 or similar releases, has been linked to problems including Outlook Classic freezing, black screens, and app crashes.

    Outlook is not working.
    KB5074109 This is the cause.
    Microsoft, do something about it.#Microsoft

    — 喘息登山者 (@hapico0109) January 20, 2026


    Read more

    You could see faster AMD Ryzen AI Max chips soon

    New leaks suggest Ryzen AI Max 400 “Gorgon Halo” could land with slicker performance.

    AMD appears to be working on a refreshed version of its Ryzen AI MAX 400 family, codenamed “Gorgon Halo”. According to recent leaks by VideoCardz, this next-gen refresh targets faster performance for Ryzen-powered machines, especially those focused on AI workloads and integrated graphics.

    The rumored Gorgon Halo series would essentially be a clock-bumped iteration of the current Strix Halo-branded processors, with the same core counts but higher boost speeds on both the CPU and Radeon iGPU sides. Additionally, it’ll also add support for faster LPDDR5X-8533 memory to further improve responsiveness and performance under AI-heavy workloads.


    Read more

    Share. Facebook Twitter Pinterest LinkedIn Reddit WhatsApp Telegram Email
    Previous ArticleThis is the tech that makes Volvo’s latest EV a major step forward
    Next Article Your future BMW electric M3 will still sound like a real M car
    TechAiVerse
    • Website

    Jonathan is a tech enthusiast and the mind behind Tech AI Verse. With a passion for artificial intelligence, consumer tech, and emerging innovations, he deliver clear, insightful content to keep readers informed. From cutting-edge gadgets to AI advancements and cryptocurrency trends, Jonathan breaks down complex topics to make technology accessible to all.

    Related Posts

    New Nvidia N1X and Nvidia N1V laptops revealed

    January 25, 2026

    Microsoft executive addresses whether all Xbox games will launch day one on PS5 like Fable

    January 25, 2026

    Citizen launches new blue chronograph with Caliber B620 Eco-Drive movement

    January 25, 2026
    Leave A Reply Cancel Reply

    Top Posts

    Ping, You’ve Got Whale: AI detection system alerts ships of whales in their path

    April 22, 2025639 Views

    Lumo vs. Duck AI: Which AI is Better for Your Privacy?

    July 31, 2025240 Views

    6.7 Cummins Lifter Failure: What Years Are Affected (And Possible Fixes)

    April 14, 2025141 Views

    6 Best MagSafe Phone Grips (2025), Tested and Reviewed

    April 6, 2025111 Views
    Don't Miss
    Technology January 25, 2026

    New Nvidia N1X and Nvidia N1V laptops revealed

    New Nvidia N1X and Nvidia N1V laptops revealed – NotebookCheck.net News The Legion 7 could…

    Microsoft executive addresses whether all Xbox games will launch day one on PS5 like Fable

    Citizen launches new blue chronograph with Caliber B620 Eco-Drive movement

    Super Mario Bros. Wonder Switch 2 upgrade price renews criticism about costly Nintendo ecosystem

    Stay In Touch
    • Facebook
    • Twitter
    • Pinterest
    • Instagram
    • YouTube
    • Vimeo

    Subscribe to Updates

    Get the latest creative news from SmartMag about art & design.

    About Us
    About Us

    Welcome to Tech AI Verse, your go-to destination for everything technology! We bring you the latest news, trends, and insights from the ever-evolving world of tech. Our coverage spans across global technology industry updates, artificial intelligence advancements, machine learning ethics, and automation innovations. Stay connected with us as we explore the limitless possibilities of technology!

    Facebook X (Twitter) Pinterest YouTube WhatsApp
    Our Picks

    New Nvidia N1X and Nvidia N1V laptops revealed

    January 25, 20262 Views

    Microsoft executive addresses whether all Xbox games will launch day one on PS5 like Fable

    January 25, 20262 Views

    Citizen launches new blue chronograph with Caliber B620 Eco-Drive movement

    January 25, 20262 Views
    Most Popular

    A Team of Female Founders Is Launching Cloud Security Tech That Could Overhaul AI Protection

    March 12, 20250 Views

    7 Best Kids Bikes (2025): Mountain, Balance, Pedal, Coaster

    March 13, 20250 Views

    VTOMAN FlashSpeed 1500: Plenty Of Power For All Your Gear

    March 13, 20250 Views
    © 2026 TechAiVerse. Designed by Divya Tech.
    • Home
    • About Us
    • Contact Us
    • Privacy Policy
    • Terms & Conditions

    Type above and press Enter to search. Press Esc to cancel.