Chess reminds me more of programming given the set of defined rules in each. However, I'm biased as I work in radiology and program more as a hobby. So far I've seen way more tools to help me code than to accurately detect radiologic findings.
Do you have any links to research or work being done on computer vision that leads you to this conclusion? Would love to check it out!
The most recent of which you mentioned, Transformers, is used by both LLMs and image synthesis/understanding. The parent posits that while computer vision lags behind NLP, this may not continue. While your comment points out that image synthesis and understanding has improved over time, I'm not sure I follow the argument that it may soon leapfrog or even catch up with LLMs (i.e. text understanding and synthesis.)