StarVector: Generating Scalable Vector Graphics Code from Images and Text
starvector.github.io
starvector.github.io
Note: StarVector models will not work for natural images or illustrations, as they have not been trained on those images. They excel in vectorizing icons, logotypes, technical diagrams, graphs, and charts.(Also would make a great SaaS... for $X/month ($9.95, $19.95, ??.??) generate unlimited icons...)
Congrats to the team for their pioneering hard work in this nascent area of LLM/Transformer research!
Well done!
Would be interested in learning about your workflow. Is it a logo generation app?
I feel like this is an example of "Machine learning is eating software". Raster to vector conversion is a perfect problem, because we can generate dataset of infinite sizes and can easily validate them with vectorize-rasterize roundtrips.
I did have an idea of performing tracing iteratively. Basically by adjusting the output SVG bit-by-bit until it matches the original image within a certain margin of error. And optimizing the output size of the SVG by simplifying curves if it does not degrade the quality. But VTracer in its current state is oneshot and probably uses 1/100 of the computational resources.
VTracer seems to perform badly on all the examples. I suspect it can be drastically improved simply by upscaling the image (via traditional interpolation, or machine learning based) and picking different parameters. But I am glad that it was cited!
> logo generation app
For logo generation, I would actually prefer code gen. I thought of this problem when reading about the diffusion language models recent (if there is lots of training data available in form of text-vector-raster triplets).
Seems like this could be incredibly valuable, but I'd argue there needs to be validation steps in place to confirm it's actually generating the right thing, for the case of image -> vector generation.