No, those are benchmark, evaluation questions. The fine tune dataset was a custom, synthetically generated dataset of ~20k PostgreSQL Text to SQL pairs covering different SQL categories and question types.
I mention a little more about it here https://x.com/calebfahlgren/status/1754247740291207198?s=20