Full threadnickthesick·Is it possible to try this out on something like llama.cpp? Does the different architecture make a difference there?View on HN