But as far as we know does AI internally assemeble subtasks into graphs and then evaluate them and pick the best one?
Is there any evidence in the memory traces of the executing AI that there are tasks and sub-tasks and ordering and evaluating of them, then taking a decision to choose and EXECUTE the best plan?
Where is the evidence that AI-programs do "planning"?
I really don't see the controversy here. My prompts, including ones meant for actual hard productivity (programming, image OCR and analysis, Q&A and summarisation of news articles), behave very differently when I introduce elements that work on the assumption that the model is partly anthropomorphic. We can't pretend that the behaviour replication isn't there, where is demonstrably is there.