I see you and other commenters don't quite understand my point. If you're wrapping model into a docker container, you don't need amalgamated single file version. It makes it harder to upgrade llamafile/model weights separately afterwards, it needs you to store separate llamafile binary for each container, etc, etc. Why not just build proper layered image with separate layer for llama.cpp and separate layer or volume for model?
Cargo cult is not in using Docker, but in using Docker to wrap something already wrapped into a comparable layer of abstraction.
Besides,
> not polluting my OS (filesystem and processes) with bits and bobs I downloaded off the internet
is purely self-deception. It's not like Docker images are not stored in some folder deep in the filesystem. If anything, it's harder to clean up after Docker than just doing rm -rf on a directory with llamafiles.