Pool spare GPU capacity to run LLMs at larger scalegithub.com11 points·i386··3 commentsOpen articleSaveView on HN