Why would the fact that it failed to follow one instruction increase the likelihood that it failed to follow others within the same response?
So when you tell it that it made a mistake, or is stupid, then those things are now prompting it to be more of the same.
And only slightly more obliquely: if part of the context includes the LLM making mistakes, expect similar activations.
Best results come if you throw away such prompts and start again. That is, iterate outside the function, not inside it.
Source, all the way down to the ability to "pay attention to" part.