Full threadbrucethemoose2·Well, this post is the second result for even running llama on Versal, so...I think MLC-LLM (though TVM) can maybe run inference?View on HN