- The assumption that both parties know about the same as an LLM does. An LLM know orders of magnitude more.
- The assumption that the output of the LLM is not refined over a few cycles.
The point is that you might give 300 bits of semantic information to an LLM, it fills it to a 1000 with perhaps 400 wrong bits. You correct it half a dozen times. It's now 950. You do the final touch ups. It's now at 1000. And it still took you 20% of the time to do it.