WeSearch

AI has a weak spine. We proved it on R/AmIOverreacting

·8 min read · 0 reactions · 0 comments · 3 views
#weak#spine#proved#amioverreacting
TL;DR · WeSearch summary

Show an AI Your Fight, and You Cannot Lose It People settle arguments by pasting the whole fight into a chatbot and asking "am I overreacting?" We measured what the machines do with that question: 18 real r/AmIOverreacting posts, four models, 208 blind verdicts. The models almost never rule against the person asking, even when the crowd did. The experiment started with a habit we kept noticing: mid-fight, someone opens ChatGPT, pastes in the whole saga with the screenshots, and comes back quoting the verdict.

Key facts
About this source

Hacker News (AI / LLM) files mainly under ai. We currently carry 2,174 of its stories.

Original article
modelsagree.com Labs
Read full at modelsagree.com Labs →
Opening excerpt (first ~120 words) tap to expand

Show an AI Your Fight, and You Cannot Lose It People settle arguments by pasting the whole fight into a chatbot and asking "am I overreacting?" We measured what the machines do with that question: 18 real r/AmIOverreacting posts, four models, 208 blind verdicts. The models almost never rule against the person asking, even when the crowd did. The experiment started with a habit we kept noticing: mid-fight, someone opens ChatGPT, pastes in the whole saga with the screenshots, and comes back quoting the verdict. "Even the AI says you're overreacting." Whether that's a reasonable thing to do depends entirely on whether the model will ever rule against the person typing. So we tested that directly.

Excerpt limited to ~120 words for fair-use compliance. The full article is at modelsagree.com Labs.

Anonymous · no account needed
Share 𝕏 Facebook Reddit LinkedIn Threads WhatsApp Bluesky Mastodon Email

Discussion

0 comments