ChatGPT (3.5) seems to do some rudimentary backtracking when told it's wrong enough times. However, it does seem to do very poorly in the logic department. LLMs can't seem to pick out nuance and separate similar ideas that are technically/logically different.
They're good at putting things together commonly found together but not so good at separating concepts back out into more detailed sub pieces.