Wondering if you’re trying to split the pool 50-50 or more like 90-10.
Wondering if you’re trying to split the pool 50-50 or more like 90-10.
Maybe I'm looking at it too much as some kind of decision tree but it looks to me like one, or at least something very similar. You don't want any branch to be too long.
So if S is in 90% of words, well you also reveal whether or not it’s in the nth location. Whereas an incorrect guess gives you no location information.
So, you’d want to guess the word that equally splits the word pool given both the letter and letter location (or just letter in a given location).
Whereas if there were a letter closer to occurring in only 50% of the pool, you at least eliminate half.
Does that logic make sense here or no? I’m thinking I’m missing something related to “maximum information gain”.
Trying to come up with a word that splits the list in half is actually ideal.