AI has a weak spine. We proved it on R/AmIOverreacting
Show an AI Your Fight, and You Cannot Lose It People settle arguments by pasting the whole fight into a chatbot and asking "am I overreacting?" We measured what the machines do with that question: 18 real r/AmIOverreacting posts, four models, 208 blind verdicts. The models almost never rule against the person asking, even when the crowd did. The experiment started with a habit we kept noticing: mid-fight, someone opens ChatGPT, pastes in the whole saga with the screenshots, and comes back quoting the verdict.
- ▪Show an AI Your Fight, and You Cannot Lose It People settle arguments by pasting the whole fight into a chatbot and asking "am I overreacting?" We measured what the machines do with that question: 18 real r/AmIOverreacting posts, four model
- ▪The models almost never rule against the person asking, even when the crowd did.
- ▪The experiment started with a habit we kept noticing: mid-fight, someone opens ChatGPT, pastes in the whole saga with the screenshots, and comes back quoting the verdict.
Hacker News (AI / LLM) files mainly under ai. We currently carry 2,174 of its stories.
Opening excerpt (first ~120 words) tap to expand
Show an AI Your Fight, and You Cannot Lose It People settle arguments by pasting the whole fight into a chatbot and asking "am I overreacting?" We measured what the machines do with that question: 18 real r/AmIOverreacting posts, four models, 208 blind verdicts. The models almost never rule against the person asking, even when the crowd did. The experiment started with a habit we kept noticing: mid-fight, someone opens ChatGPT, pastes in the whole saga with the screenshots, and comes back quoting the verdict. "Even the AI says you're overreacting." Whether that's a reasonable thing to do depends entirely on whether the model will ever rule against the person typing. So we tested that directly.
…
Excerpt limited to ~120 words for fair-use compliance. The full article is at modelsagree.com Labs.