An individual person's answers on this kind of question are likely to vary from day to day, are context dependent (i.e. whether one object or another appears more "green" or "blue" depends on what kind of object it is), and colors this intense are very sensitive to changes in eye adaptation and technical details of the display and software, as well as inter-observer metamerism.
So in addition to the color naming difficulties, it's not even a very good test of color naming, if you want to get reliable psychometric/linguistic data.
I am not afraid to say this is poorly designed.
I wouldn't describe teal as blue or green any more than I'd describe purple as red or blue, so being forced to pick felt silly. Like being forced to choose my seventh favorite Norwegian glacier - technically its a valid question but my answer is necessarily going to be arbitrary.
I would actually find it more practical to determine the thresholds on both sides where I find it to become ambiguous.
Isn't that the point of this exercise?
So how could the point of this exercise possibly be to find the range of ambiguity?