None of those matters (except multimodal). If you are running a business, the only thing that matters is
a) How does it perform on my set of evals
b) What is the cost/latency of serving it to my consumers.
It shouldn't matter to me how many parameters, corpus it is trained on, whether it's LLM or Transformer or something else