I don't think Windows does this, but Ollama does
Most people who know it does this turns it off because it kicks in too early so if you have 24GB it'll offload to RAM and tank your inference speed when you hit around 22GB use.
https://nvidia.custhelp.com/app/answers/detail/a_id/5490/~/s...
https://nvidia.custhelp.com/app/answers/detail/a_id/5490/~/s...