Memory reclamation really useful if you are using Ollama or similar in WSL2 - after awhile, Ollama will unload the model and WSL2 will give the memory back to the OS
1. If using `ollama run`: `ollama run llama3 --keepalive -1`
2. If running ollama serve directly, use `OLLAMA_KEEP_ALIVE=-1` ollama serve
3. If using the api, there's a `keep_alive` parameter you can set to -1