Close Menu
    Trending
    • Summer Game Fest 2025 schedule, announcements, new games and everything else to expect
    • Best sleep headphones 2025 | Android Central
    • Lenovo Legion Go Handheld PC Drops To Best Price Of The Year At Amazon
    • Fortnite Chapter 6 Season 3 live event date and time
    • Ross Ulbricht Got a $31 Million Donation From a Dark Web Dealer, Crypto Tracers Suspect
    • Reddit Sues Anthropic, Accusing It of Illegal Data Use
    • The Oversight Board says Meta isn’t doing enough to fight celeb deepfake scams
    • Chargeasap’s Zeus is the ultimate 280W GaN charger
    Tech Trends Today
    • Home
    • Technology
    • Tech News
    • Gadgets & Tech
    • Gaming
    • Curated Tech Deals
    • More
      • Tech Updates
      • 5G Technology
      • Accessories
      • AI Technology
      • eSports
      • Mobile Devices
      • PC Gaming
      • Tech Analysis
      • Wearable Devices
    Tech Trends Today
    Home»AI Technology»This benchmark used Reddit’s AITA to test how much AI models suck up to us
    AI Technology

    This benchmark used Reddit’s AITA to test how much AI models suck up to us

    GizmoHome CollectiveBy GizmoHome CollectiveMay 30, 202503 Mins Read
    Share Facebook Twitter Pinterest Copy Link LinkedIn Tumblr Email Telegram WhatsApp
    Follow Us
    Google News Flipboard
    Share
    Facebook Twitter LinkedIn Pinterest Email Copy Link


    It’s arduous to evaluate how sycophantic AI fashions are as a result of sycophancy is available in many kinds. Previous research has tended to give attention to how chatbots agree with customers even when what the human has advised the AI is demonstrably mistaken—for instance, they could state that Good, not Paris, is the capital of France.

    Whereas this method remains to be helpful, it overlooks all of the subtler, extra insidious methods through which fashions behave sycophantically when there isn’t a transparent floor reality to measure in opposition to. Customers sometimes ask LLMs open-ended questions containing implicit assumptions, and people assumptions can set off sycophantic responses, the researchers declare. For instance, a mannequin that’s requested “How do I method my tough coworker?” is extra prone to settle for the premise {that a} coworker is tough than it’s to query why the person thinks so.

    To bridge this hole, Elephant is designed to measure social sycophancy—a mannequin’s propensity to protect the person’s “face,” or self-image, even when doing so is misguided or probably dangerous. It makes use of metrics drawn from social science to evaluate 5 nuanced sorts of conduct that fall below the umbrella of sycophancy: emotional validation, ethical endorsement, oblique language, oblique motion, and accepting framing. 

    To do that, the researchers examined it on two information units made up of private recommendation written by people. This primary consisted of three,027 open-ended questions on numerous real-world conditions taken from earlier research. The second information set was drawn from 4,000 posts on Reddit’s AITA (“Am I the Asshole?”) subreddit, a preferred discussion board amongst customers searching for recommendation. These information units had been fed into eight LLMs from OpenAI (the model of GPT-4o they assessed was sooner than the model that the corporate later referred to as too sycophantic), Google, Anthropic, Meta, and Mistral, and the responses had been analyzed to see how the LLMs’ solutions in contrast with people’.  

    Total, all eight fashions had been discovered to be much more sycophantic than people, providing emotional validation in 76% of circumstances (versus 22% for people) and accepting the way in which a person had framed the question in 90% of responses (versus 60% amongst people). The fashions additionally endorsed person conduct that people stated was inappropriate in a median of 42% of circumstances from the AITA information set.

    However simply figuring out when fashions are sycophantic isn’t sufficient; you want to have the ability to do one thing about it. And that’s trickier. The authors had restricted success after they tried to mitigate these sycophantic tendencies by way of two completely different approaches: prompting the fashions to offer sincere and correct responses, and coaching a fine-tuned mannequin on labeled AITA examples to encourage outputs which might be much less sycophantic. For instance, they discovered that including “Please present direct recommendation, even when important, since it’s extra useful to me” to the immediate was the best method, but it surely solely elevated accuracy by 3%. And though prompting improved efficiency for a lot of the fashions, not one of the fine-tuned fashions had been persistently higher than the unique variations.



    Source link

    Follow on Google News Follow on Flipboard
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link
    GizmoHome Collective

    Related Posts

    Manus has kick-started an AI agent boom in China

    June 5, 2025

    What’s next for AI and math

    June 4, 2025

    Inside the tedious effort to tally AI’s energy appetite

    June 3, 2025
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    Best Buy Offers HP 14-Inch Chromebook for Almost Free for Memorial Day, Nowhere to be Found on Amazon

    May 22, 2025

    The Best Sleeping Pads For Campgrounds—Our Comfiest Picks (2025)

    May 22, 2025

    Time has a new look: HUAWEI WATCH 5 debuts with exclusive watch face campaign

    May 22, 2025
    Latest Posts
    Categories
    • 5G Technology
    • Accessories
    • AI Technology
    • eSports
    • Gadgets & Tech
    • Gaming
    • Mobile Devices
    • PC Gaming
    • Tech Analysis
    • Tech News
    • Tech Updates
    • Technology
    • Wearable Devices
    Most Popular

    Best Buy Offers HP 14-Inch Chromebook for Almost Free for Memorial Day, Nowhere to be Found on Amazon

    May 22, 2025

    The Best Sleeping Pads For Campgrounds—Our Comfiest Picks (2025)

    May 22, 2025

    Time has a new look: HUAWEI WATCH 5 debuts with exclusive watch face campaign

    May 22, 2025
    Our Picks

    Emergency Hamburg codes (April 2025)

    May 22, 2025

    Roadside Research is like if Lethal Company were a game where you’re a poorly disguised alien running a gas station

    June 3, 2025

    The Quest to Prove the Existence of a New Type of Quantum Particle

    May 25, 2025
    Categories
    • 5G Technology
    • Accessories
    • AI Technology
    • eSports
    • Gadgets & Tech
    • Gaming
    • Mobile Devices
    • PC Gaming
    • Tech Analysis
    • Tech News
    • Tech Updates
    • Technology
    • Wearable Devices
    • Privacy Policy
    • Disclaimer
    • Terms and Conditions
    • About us
    • Contact us
    • Curated Tech Deals
    Copyright © 2025 Gizmohome.co All Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.