I have an interesting dissonance with this. On one hand, I understand how huge parameter sets can and do model specific personas well. I've also read some of these cited papers and _know_ intellectually that predicted results can be close to actual survey data. The other part of me is screaming at my laptop that language modeling is about aggregate statistics, revealed preference counts for a lot, and how could a language model actually substitute for market research?
I imagine the biggest hurtle you're going to face are research teams that:
- A. Want to see actual proof behind data
- B. Disbelieve a LLM could generate statistically significant insights about real people that would make individual decisions
- C. Need to justify their own existence / organizational clout with boots on the ground facilitating surveys
A and C might be surmountable, but I'm not sure of a good way of tackling B.