[Fortune] Over just a few months, ChatGPT went from correctly answering a simple math problem 98% of the time to just 2%, study finds

trashhalo@beehaw.org · edit-2 1 year ago

[Fortune] Over just a few months, ChatGPT went from correctly answering a simple math problem 98% of the time to just 2%, study finds

calculuschild@lemm.ee · edit-2 1 year ago

My understanding is this claim is basically entirely false. The tests done by these researchers had some glaring errors that when corrected, show gpt-4 is getting slightly better at math, if anything. See this video that describes some of the issues: https://youtu.be/YSokS2ivf7U

TL;DR The researchers gave new GPT questions from two different pools. It’s no surprise they got worse answers.

darkkite@lemmy.ml · 1 year ago

came here to say the same