I can see why it was the case from engineering standpoint: adding a name from spoken phrase to a list is trivial, but looking up an item (that might not even be there) from a list based on a spoken phrase is prone to transcription errors. The user might refer to it with different phrasing than when it was added too.
This is feasible only very recently with LLMs.