Really enjoyed the historical framing here. Starting from bag-of-words and working up to Jev makes the “just a classifier” argument much more nuanced. The tradeoff between generality, speed, and cost is probably the most interesting part to me — especially where Jev sits between task-specific classifiers and full LLMs.