@futurebird It's really weird for me to see them use a LLM for anything to do with math because it literally — by definition — can't do math. It can try to approximate "statistically speaking, this probably follows that question" and be trained to raise the statistics such that it's usually right, but it still can't actually perform the math because that's fundamentally just not how it works... Or, more fundamentally, they can't do logic for the same reason — their methodology of logic is "this statistically seems to follow that a lot in the training data, thus I output it" not "ah hah, if this is true, then that must be true."
Thus they can only be used to pretend at math... They might get it right but if they do it's by luck, not by actually applying logic or mathematics.