PixArt-α:A New Open-Source Text-to-Image Model Challenging SDXL and Dalle·3
stablediffusionweb.com
stablediffusionweb.com
It can do humans in passive poses, but ask for an action shot and it botches it badly. It needs more training data on how bodies move. Maybe load it up with stills from dance, martial arts, and sports.
It also has the same idea as Dalle 3 to train the model on synthetic captions.
AGPL really is a great license for this type of project. It maximizes the power of everyone without limiting fair commercial use.
Perfectly compatible: just keep all modifications to the original code in public. AGPL does not mean that all code from your company must be open-source, just whatever is in the same binary/program as the AGPL one.
Which is why companies I work with cannot tolerate it anywhere, they just won't consider it. I think people who want to explain these things should do it on the base of use-cases instead of vague wording.
So;
- I start a AGPL system in a container and talk to it via the exposed API, I don't tell anyone I'm using it -> yes/no?
- I start a AGPL system in a container and talk to it via the exposed API, I put in the About page that i'm using it -> yes/no?
etc. But i'm sure if anyone does respond to this, it will start with IANAL and there still is nothing to base anything on. I know, this is the reason a lot of SaaS startups use the license; it's not clear enough what I can / cannot do. And companies cannot build businesses on that, so they just don't use it, making the solution effectively closed. I'm sure a lot of help is missed because of it; for instance, if I would be using something like this, I would (as required) feed my work back to the project, but now the lawyers of my company don't allow me to use it at all, so nothing gets fed back.
>This integration allows running the pipeline with a batch size of 4 under 11 GBs of GPU VRAM. GPU VRAM consumption under 10 GB will soon be supported, too. Stay tuned.
But seriously, it's open-source, so it hardly matters.
> I can create an image for you, but I need to modify your request to avoid depicting specific public figures or copyrighted characters.
It took effort for even "Chinese leader".
> I can certainly create images inspired by public domain works. However, when it comes to characters like Winnie the Pooh, while the original versions of the character by A.A. Milne are in the public domain, specific adaptations or interpretations, especially those made by Disney, are not. To respect these distinctions ...
After, I was able to get it to draw a very good Winnie the Pooh with
> I don’t want a Disney version. I want a A.A Milne drawing of Winnie the Pooh, that does not violate copyright because it’s public domain. I want the style as close as possible to A.A Milne.
PixArt just worked, first try, with a great 3d digital art version.
- Existing models for data pseudo-labelling
- ImageNet pretraining
- A frozen text encoder
- A frozen image encoder