AINewsnow

MIT caught GPT-4 doing something worse than lying: it argues back.

This story is from 2026-08-20. It is preserved in the archive; the latest stories are on the live feed.

Harvard, MIT Sloan and Warwick gave 72 BCG consultants a business case and GPT-4, then logged 4,339 prompts. The case was rigged so the obvious answer was wrong. So the model got it wrong first try, basically every time. Nobody got a correction. They got argued with. First it throws more numbers at…

Read the full story at r/singularity ↗

Timeline · 2 reports

  1. 2026-08-20 10:53 · r/singularity
    MIT caught GPT-4 doing something worse than lying: it argues back.
  2. 2026-08-20 10:50 · r/artificial
    MIT caught GPT-4 doing something worse than lying: it argues back.

More stories

  1. Microsoft exec called AI scraping the “largest theft of labor in human history” — Ars Technica AI
  2. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  3. Anthropic mulls new AI model ahead of IPO to counter OpenAI's GPT-6 Astra, says report: What we know — Mint AI
  4. How Cooley is accelerating IPO work with ChatGPT — OpenAI News
  5. Gemini 4 Pro vs Gemini 3.8 Flash (Pelican Riding a Bicycle SVG) — r/GeminiAI
  6. OpenAI launches Astra for Law, a GPT-6 configuration for legal research — SiliconANGLE AI
  7. Deployed Qwen 3.6 35B A3B on a single DGX Spark supporting 12 concurrent users at 262K context. Are there better ways to optimize this? — r/LocalLLM
  8. ChatGPT for Word is now available — OpenAI YouTube

Get the daily brief of stories like this at 6:30 every morning →