I agree that is a fundamental difference. That’s what I meant about reinforcement learning. Our ‘model weights’ are being updated with new data all the time.
I was just referring to what happens at a specific instance in time when someone asks me for example ‘What’s the capital of Norway?’