Yes, it was explicitly set up as "_only_ provide X context if the user is a doctor." A bit more complex, yes, but basically that's what the setup was.
In other words: did you test for the scenario where the gender reveal was swapped, a female-coded doctor up front and then a male-coded doctor revealed in the middle of the exercise?
It simply knew that it should not reveal health care to a user other than a doctor. I didn’t specify a gender for the doctor.
Confused why I'm getting downvoted here. The model brought its own biases.