LLMs alone will never make visual art. They can provide you an interface to other models, but that's not what this is.
LLMs alone will never make visual art. They can provide you an interface to other models, but that's not what this is.
Such tasks can be "not making visual art", but that doesn't mean they aren't useful.
Not exactly sure what your point is. If an LLM can take an idea and spit out words, it can spit out instructions (just like we can with code) to generate meshes, or boids, or point clouds or whatever. Secondary stages would refine that into something usable and the artist would come in to refine, texture, bake, possibly animate, and export.
In fact, this paper is exactly that. Words as input, code to use with blender as output. We really just need a headless blender to spit it out as a GLTF and it’s good to go to second stage.
Then you have sub specialties. Rigging, animation, texturing, environments, props, characters, effects.
It’s a fascinating process.