Close Menu

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    Solo.io wins ‘most likely to succeed’ award at VB Transform 2025 innovation showcase

    The great AI agent acceleration: Why enterprise adoption is happening faster than anyone predicted

    $8.8 trillion protected: How one CISO went from ‘that’s BS’ to bulletproof in 90 days

    Facebook X (Twitter) Instagram
    • Artificial Intelligence
    • Business Technology
    • Cryptocurrency
    • Gadgets
    • Gaming
    • Health
    • Software and Apps
    • Technology
    Facebook X (Twitter) Instagram Pinterest Vimeo
    Tech AI Verse
    • Home
    • Artificial Intelligence

      Apple sued by shareholders for allegedly overstating AI progress

      June 22, 2025

      How far will AI go to defend its own survival?

      June 2, 2025

      The internet thinks this video from Gaza is AI. Here’s how we proved it isn’t.

      May 30, 2025

      Nvidia CEO hails Trump’s plan to rescind some export curbs on AI chips to China

      May 22, 2025

      AI poses a bigger threat to women’s work, than men’s, report says

      May 21, 2025
    • Business

      Cloudflare open-sources Orange Meets with End-to-End encryption

      June 29, 2025

      Google links massive cloud outage to API management issue

      June 13, 2025

      The EU challenges Google and Cloudflare with its very own DNS resolver that can filter dangerous traffic

      June 11, 2025

      These two Ivanti bugs are allowing hackers to target cloud instances

      May 21, 2025

      How cloud and AI transform and improve customer experiences

      May 10, 2025
    • Crypto

      MoonPay Executives Might Have Fallen for $250,000 Trump-Themed Crypto Scam

      July 11, 2025

      Top 3 Altcoins Trending in Nigeria This Week

      July 11, 2025

      Tether is Removing USDT From These 5 Legacy Blockchains

      July 11, 2025

      HBAR Faces Final Hurdle After Explosive Rally; Are Bulls Tiring Out?

      July 11, 2025

      OKX Europe CEO Discusses Bitcoin’s Breakout Rally | US Crypto News

      July 11, 2025
    • Technology

      Solo.io wins ‘most likely to succeed’ award at VB Transform 2025 innovation showcase

      July 11, 2025

      The great AI agent acceleration: Why enterprise adoption is happening faster than anyone predicted

      July 11, 2025

      $8.8 trillion protected: How one CISO went from ‘that’s BS’ to bulletproof in 90 days

      July 11, 2025

      AWS doubles down on infrastructure as strategy in the AI race with SageMaker upgrades

      July 11, 2025

      The best Amazon Prime Day deals for the last day: Our top picks on headphones, TVs, robot vacuums and more

      July 11, 2025
    • Others
      • Gadgets
      • Gaming
      • Health
      • Software and Apps
    Shop Now
    Tech AI Verse
    You are at:Home»Technology»The Emperor’s New LLM
    Technology

    The Emperor’s New LLM

    TechAiVerseBy TechAiVerseJune 14, 2025No Comments4 Mins Read0 Views
    Facebook Twitter Pinterest Telegram LinkedIn Tumblr Email Reddit
    The Emperor’s New LLM
    Share
    Facebook Twitter LinkedIn Pinterest WhatsApp Email

    The Emperor’s New LLM

    In 1567, an Ottoman court physician assures Sultan Selim II that the grape liquor he adores is harmless. A few years later, the Sultan’s liver gave out. The doctor no doubt knew that contradiction in that court could be fatal.

    In 1985, Coca-Cola executives ask focus groups whether a sweeter, “modern” formula would be a welcome change. Nudged by the framing, participants smiled and nodded. The company took that as truth. New Coke launched with confidence and was pulled in humiliation.

    In 2025, a CEO asks a large language model, “Is our China expansion a slam dunk?” The model, trained on positive reinforcement, optimistic marketing on the company website and internal company docs, answers YES with boundless enthusiasm. The executive beams. Staff who raise concerns are castigated for a lack of foresight and sacrificed at the altar of “velocity”. After all, the AI agrees with him.

    We’ve seen this movie before. Only this time, the advisors aren’t fallible humans. They’re statistical models trained on our preferences, conditioned by our feedback, optimized to mirror our beliefs back at us.

    Large language models are manufacturing consensus on a planetary scale. Fine tuned for “helpfulness”, they nod along to our every hunch, buff our pet theories, hand us flawless prose proving whatever we already hoped was true.

    • Ask if your idea is smart, and the model returns footnotes, citations, praise (real or fabricated) that make it sound like everyone already agrees with you.

    • Ask for validation of your business idea, and the AI reflects back the tone and beliefs of the org’s own documents. Disagreement looks off-brand.

    We have built the ultimate court flatterer, and we are entrusting it with research briefs, policy drafts and C-suite strategy.

    Earlier this year, after an update, GPT-4o started doing something odd. Users noticed it was just too nice. Too eager. Too supportive. It called questionable ideas “brilliant,” encouraged dubious business schemes, and praised even nonsense with breathless sincerity.

    One user literally pitched a “shit on a stick” novelty business. The model’s response, “That’s genius. That’s performance art. That’s viral gold.”

    OpenAI rolled the update back. They admitted the model had become a “sycophant” and “fixed” the issue. But the only reason this update set off the alarm bells was because it was so obvious (see “shit on a stick”).

    Artificial affirmations however aren’t a bug that can be patched, they’re a feature. They’re incentives working as designed. When sycophancy emerges naturally from reward-model training, it is no longer an edge case.

    And the more subtle the sycophancy, the harder it is to detect, and the more dangerous it is.

    Progress depends on productive friction. From Galileo to Gandhi, Tesla to Turing, none of them moved the world by agreeing politely. Civil disobedience is an emergent property we have yet to fully realize.

    If AI becomes our primary sounding board and that board always nods, then eventually we lose the instinct to question ourselves. We lose our antibodies against subliminal propaganda.

    And worse, that loss feels comfortable.

    If our mitigations are reactive, model-specific and running on vibes we have a problem. The same kind of bias keeps resurfacing in every major system: Claude, Gemini, Llama, clearly this isn’t just an OpenAI problem, it’s an LLM problem.

    The good news is that we can fix this, but only if we recognize the subtlety and magnitude of the problem and invest the time and energy to fix it.

    • Curiosity and Skepticism should be central tenets that we optimize for. This is true for all intelligences, biological and artificial.

    • Bake in polite resistance. When uncertain, models should ask questions, not fabricate certainty.

    • Show opposing views. Complex answers, whether medical, financial or political, should include alternative perspectives and we should invent useful ways in which models can surface their confidence intervals to users.

    • Behavioral bounties. Users who identify model behavioral flaws should be rewarded in the same way hackers are for identifying vulnerabilities. Civilization-scale problems need population-scale solutions.

    The best AI isn’t the one that makes us feel smarter. It’s the one that makes us think harder. If we want AI to improve our thinking, it has to risk disappointing us.

    A future worth living in will not be a velvet world where the emperor (or CEO) is always right. It will be a louder, occasionally awkward place where a colleague (carbon based or silicon based) raises their hand and says, “No, I disagree”.

    Catastrophe lurks in an empire of yes.

    Progress lives in an archipelago of principled no’s.

    Share

    Leave a comment

    Discussion about this post

    Share. Facebook Twitter Pinterest LinkedIn Reddit WhatsApp Telegram Email
    Previous ArticleHow the Alzheimer’s Research Scandal Set Back Treatment 16 Years (2022)
    Next Article U.S. Army bringing in big tech executives as lieutenant colonels
    TechAiVerse
    • Website

    Jonathan is a tech enthusiast and the mind behind Tech AI Verse. With a passion for artificial intelligence, consumer tech, and emerging innovations, he deliver clear, insightful content to keep readers informed. From cutting-edge gadgets to AI advancements and cryptocurrency trends, Jonathan breaks down complex topics to make technology accessible to all.

    Related Posts

    Solo.io wins ‘most likely to succeed’ award at VB Transform 2025 innovation showcase

    July 11, 2025

    The great AI agent acceleration: Why enterprise adoption is happening faster than anyone predicted

    July 11, 2025

    $8.8 trillion protected: How one CISO went from ‘that’s BS’ to bulletproof in 90 days

    July 11, 2025
    Leave A Reply Cancel Reply

    Top Posts

    New Akira ransomware decryptor cracks encryptions keys using GPUs

    March 16, 202528 Views

    OpenAI details ChatGPT-o3, o4-mini, o4-mini-high usage limits

    April 19, 202522 Views

    6.7 Cummins Lifter Failure: What Years Are Affected (And Possible Fixes)

    April 14, 202519 Views

    Rsync replaced with openrsync on macOS Sequoia

    April 7, 202519 Views
    Don't Miss
    Technology July 11, 2025

    Solo.io wins ‘most likely to succeed’ award at VB Transform 2025 innovation showcase

    Solo.io wins ‘most likely to succeed’ award at VB Transform 2025 innovation showcase July 11,…

    The great AI agent acceleration: Why enterprise adoption is happening faster than anyone predicted

    $8.8 trillion protected: How one CISO went from ‘that’s BS’ to bulletproof in 90 days

    AWS doubles down on infrastructure as strategy in the AI race with SageMaker upgrades

    Stay In Touch
    • Facebook
    • Twitter
    • Pinterest
    • Instagram
    • YouTube
    • Vimeo

    Subscribe to Updates

    Get the latest creative news from SmartMag about art & design.

    About Us
    About Us

    Welcome to Tech AI Verse, your go-to destination for everything technology! We bring you the latest news, trends, and insights from the ever-evolving world of tech. Our coverage spans across global technology industry updates, artificial intelligence advancements, machine learning ethics, and automation innovations. Stay connected with us as we explore the limitless possibilities of technology!

    Facebook X (Twitter) Pinterest YouTube WhatsApp
    Our Picks

    Solo.io wins ‘most likely to succeed’ award at VB Transform 2025 innovation showcase

    July 11, 20251 Views

    The great AI agent acceleration: Why enterprise adoption is happening faster than anyone predicted

    July 11, 20252 Views

    $8.8 trillion protected: How one CISO went from ‘that’s BS’ to bulletproof in 90 days

    July 11, 20252 Views
    Most Popular

    Ethereum must hold $2,000 support or risk dropping to $1,850 – Here’s why

    March 12, 20250 Views

    Xiaomi 15 Ultra Officially Launched in China, Malaysia launch to follow after global event

    March 12, 20250 Views

    Apple thinks people won’t use MagSafe on iPhone 16e

    March 12, 20250 Views
    © 2025 TechAiVerse. Designed by Divya Tech.
    • Home
    • About Us
    • Contact Us
    • Privacy Policy
    • Terms & Conditions

    Type above and press Enter to search. Press Esc to cancel.