> This is similar to how, not too long ago, LLM's had extreme difficulty counting the number of letters in some words.
The specific issue of Google is that they are using an underpowered model, not fit to task, and much prone to hallucination than either OpenAI or Anthropic free tier offerings.
Google should at least match the frontier labs at the free tier (with some limit; after that, degrade quality), ffs