Beating The DALL-E Mini: Too much traffic, please try again
wandb.ai
wandb.ai
I once asked in the LAION discord what that was and they said basically flask/fastapi server + load the model + gunicorn + ec2 instances.
I'm curious if any HN readers have any further insight or thoughts on ML model deployment for projects like DALL-E Mini.