ParentFull threadsujayk_33·It's faster inference because of the Hardware (LPUs), here the question is about architectures (AR or Diffusions)View on HN