Supervised training uses this precise loop. The difference between inference and observation is called the loss function.
Models are now starting to show emergent behavior that they were not explicitly trained for. This is leading to the current debate.