Thanks for the response and paper. My main question here is that it doesn't seem to make sense that random search did significantly worse than random parameters? It seems that random search should only consistently lose in cases when the dev set performance is anti-correlated with the test set performance. Why do you think you saw random parameters beating random search? (Is there a typo in the table or something)?