StyleDrop: Text-to-Image Generation in Any Style
styledrop.github.io
styledrop.github.io
And without code or model files, who knows if research is actually reproducible and not cherry-picked to look good?
This is research for some math to improve the ability to do a style transfer, and they show it on their own text-to-image generator muse, which they have also published the structure of.
This is what they have typically done, and this is what they didn't do with Bard.
No, they did not release waits for it, you cannot run this on your own computer. But they typically didn't release weights for things like Lambda or Imagen either AFAIK.
This is not a product. This is not a tool for you to use. This is for researchers.
The point of this paper is not to let you run it on your computer. It's to allow other researchers to implement and build on the methods described in the paper.
1) Allow publishing everything including source code => this helps the competitor directly. Bad move.
2) Disallow publishing => the researchers will be tempted to switch jobs for their competitor, since staying at Google will hurt their career. Bad move.
3) Allow publishing, but disallow everything else => this helps the competitors a little, but not too much. The researchers get credit for their work, which removes any incentive they have at switching jobs. Seems like the best compromise.
At least, that's my speculative take on this. Sure, OpenAI & StabilityAI get the credit in the public's eye, but there are also other incentives at play.
Anyway, it's not 2000, you can't get away with releasing a paper without code. In the case of AI/ML, you either need to release the weights, or make some web doodad that allows you to use the model. If that's not there, I just assume that the results aren't reproducible.
Because it's worked for the longest time.
Even a year ago, Google was percieved as being so far ahead. These little papers with their landing pages were like sneek peaks into the advanced tech behind the scenes. It's marketing that brought hype to Google's brand and we were all excited for it because none of the big movers felt enough pressure to actually put stuff out so we were all excited about the possibilities.
You don't approve - so what? Releasing the weights doesn't make them money.
I'm skeptical LLaMa is even useful for facebook commercially at least, they don't make money on it, and I doubt anyone developed brand loyalty, more then likely everyone will use whatever the next, best open model is regardless of who makes it.
Llama and SD still don't come close to midjourney/chatgpt/claude when you look at ease of use and infrastructure cost. These "99% the performance of chatgpt" are laughable if you use them (which I have extensively).
> We are done inhaling vapor.
Okay? What were you about to pay for to begin with here?
EDIT: Just to add, it's not like we got nothing from this, this can likely still be something to try with SD.
But it doesn't matter, soon they will stop doing what we prefer, I guess...
There's about 10 posts on the front page a day about AI and filling each one up with comments about the broader topic just diffuses the debate - both about the post topic in question and about the important issues you're trying to raise.
When I want to try Bard, I still get "Bard isn’t currently supported in your country. Stay tuned!". While me and everyone around me already uses ChatGPT, FastGPT, Phind and Perplexity.