> Imagine a calculator program that computes billions of two number multiplications accurately by looking up prior examples
This is a poor analogy because:
* Multiplying numbers has a single objective answer. Whether a code is good (sometimes even just whether it's correct) can be quite subjective.
* LLMs certainly do some level of composition between the data sources they were trained on i.e. they are more than just lookup tables.
* We have calculators that actually do multiply large numbers accurately. We don't have anything that automatically writes code that is definitely correct and "good".
To address only the last point: imagine you had a device that would quickly factor large numbers used in modern criticality, but occasionally got it wrong. You could waste a lot time debating whether it's a "calculator", but it's still certainly useful to have one.