That’s a good point. Imagine if an LLM could only read, speak, and hear at the same speed as a human. How long would training a model take?
We can make them read digital media really quickly, but we can’t really accelerate its interactions with the physical world.