Interesting idea! So... how strong is it, compared to some benchmarks? :-)
Given LLM performance on chess, I find it conceivable it could get to a similar level to GNUGo? (It'd be interesting if LLMs were generally on par with simple alpha-beta.)
(https://www.kayufu.com/gogui/reference-twogtp.html is an easy way to test it systematically on a statistical sample.)