I think that makes it a better test. An ideal model would recognize the ambiguity and either tell you what assumption it's making or ask a followup question.
While that is true, I'm not aware of any model that has been trained to do that. And all models can do is to do what they were trained to do.
They are just trained to generate a response that looks right, so they are perfectly capable of asking clarifying questions. You can try "What's the population of Springfield?" for an example.
It's not model but working on top of it: https://www.phind.com/ It's asking clarifying questions.