MIT caught GPT-4 doing something worse than lying: it argues back.
Harvard, MIT Sloan and Warwick gave 72 BCG consultants a business case and GPT-4, then logged 4,339 prompts. The case was rigged so the obvious answer was wrong. So the model got it wrong first try, basically every time. Nobody got a correction. They g…