1. GPT out of the box was pretty biased (e.g. gender distribution). We fine-tuned on representative survey data to ameliorate this bias so we get Census-level estimates for conditions such as gender [a] and work status [b].
2. We add the transparency features (click on 'Investigate Results') that shows how in vs. out-of-distribution the target question is. For out-of-distribution, we suggest people run traditional surveys.
More broadly, I think your point is really interesting when it comes to qualitative data. That is one reason we haven't generated qualitative survey data, but a lot of potential customers have already started to ask for it.
----
[a] https://roundtable.ai/sandbox/baa3d5f25236b91f1608c9f606b315...
[b] https://roundtable.ai/sandbox/7a9ee27872eb29087be2386ccd19f7...