They have also published smaller 33B model called Laguna XS 2.1, its Q4 gguf is 20GB.
https://huggingface.co/poolside/Laguna-XS-2.1-GGUF/tree/main
https://huggingface.co/poolside/Laguna-XS-2.1-GGUF/tree/main
Maybe the chat template is broken, maybe not. But I can't get it to think/reason, or at least sometimes it seems to decide not to even when asked. And it is wildly inconsistent when it does it, sometimes thinking out loud.
However, without thinking enabled it is shockingly good at researching using web tool calls. It's also very fast.
I am assuming something is not yet right with the chat template, even using a recent build of llama.cpp, and I guess I will revisit this model again at some point in the future.