Is it just me or is Nvidia trolling hard by calling a model with 30b parameters "nano"? With a bit of context, it doesn't even fit on a RTX 5090.
Other LLMs with the "nano" moniker are around 1b parameters or less.
Other LLMs with the "nano" moniker are around 1b parameters or less.
https://github.com/jameschrisa/Ollama_Tuning_Guide/blob/main...