Asking robots to design stained glass windows
astralcodexten.substack.com
astralcodexten.substack.com
I dunno, I think a stained glass window of a finch with Charles Darwin's head is pretty great, personally
Tycho Brahe should definitely be an Owl.
Alexandra Elbakyan is evidently going to be a raven, however could probably ask her which bird she would want her head on.
Since this is virtues we can also go looking for help in the symbolism of birds https://worldbirds.com/bird-symbolism/
which tells us for example that the Swan symbolizes "Light, twin flame, purity", so I guess the Swan would be lightness, but then who's head should be on it? I propose W.B Yeats because of the Wild Swans at Coole https://www.poetryfoundation.org/poems/43288/the-wild-swans-...
Not sure what birds the others should have their heads affixed to, but I think this is definitely the way I would go as soon as I saw Darwin's head on the finch.
It seems like a pretty good general way to work toward specific requirements, though I haven't tried it.
The same can be said with respect to how a great many treat most of their relationships, both personally, professionally, politically...
Unfortunately the impact of Imagen will be low like all the other Google models because, as keeps happening, they don't trust people enough to actually release it, not even in demo form:
"Imagen relies on text encoders trained on uncurated web-scale data, and thus inherits the social biases and limitations of large language models. As such, there is a risk that Imagen has encoded harmful stereotypes and representations, which guides our decision to not release Imagen for public use without further safeguards in place."
This seems to be emerging as a clear pattern - OpenAI trains an AI and then opens it to a group of select invitees (the usual suspects) with lots of opaque T&Cs that restrict what they're allowed to say about it [1]. For example, OpenAI forbids users from sharing any outputs that contain realistic faces even though the AI generates such pictures all the time, even when not requested [2]. This is reminiscent of how RDBMS vendors forbid users from publishing benchmarks and usually results in people giving OpenAI some stick for not living up to their name.
But OpenAI is still light years more open than Google, which routinely announces just months later that they've trained an AI far better than anything OpenAI produced, but doesn't provide any evidence beyond their own cherry picked examples. The justification given is always the same: the AI has learned un-woke things like the fact that certain jobs tend to be done by specific genders, and would generate examples based on that learnings. And because the Google researchers are crazy they think that seeing pictures of white male builders or young female nurses would be actually dangerous to the public, so they just don't provide any API access at all, not even to little select lists of friends.
IMO this is turning into a serious problem for AI research. It's already notoriously non-reproducible, but OpenAI's work is at least auditable. People can do experiments on DALL-E and see that it's real for themselves, explore the limits and come up with ways to handle them, as is happening here. They could in future build apps that use them. In contrast Google, who should by all rights be at the forefront of this space, keeps getting eclipsed by OpenAI again and again because they're now so woke that they've retreated into a tiny little purity bubble. They claim that they'll make their AIs available when they figured out how to brainwash them to have the right views, but they were claiming this is an "open problem" for years and never seem to release anything. So all the talk is about GPT-3 and DALL-E and Google's equivalents just get forgotten.
[1] https://www.lesswrong.com/posts/uKp6tBFStnsvrot5t/what-dall-...
[2] https://www.lesswrong.com/posts/uKp6tBFStnsvrot5t/what-dall-...
The link you posted seems to reinforce my point though. It's a reimplementation of DALL-E, not anything done by Google. That seems to be how it goes, everyone tries to duplicate the OpenAI work even though Google's models are (or claim to be) more advanced. The mindshare value of making their demos available appears to be phenomenal.
Is it reproducible - well, only for flexible definitions of reproducible. If you try to follow the method in the paper and don't get results as good, is the issue the method or that you didn't follow it well enough? You can't follow it all that closely because they often aren't detailed enough and the training sets are constantly changing.
The issue here (for Google) is a bit different. If they were keeping their research proprietary because they wanted to turn it into a product and sell it, that's one thing and totally understandable. It's very expensive and needs some way to financially justify the cost. Their justification is quite different though and raises questions about their whole AI initiative. What's the point of creating SOTA AI trained on the internet if you're afraid of what people will use it for? Their efforts to make ideologically acceptable AI don't seem to have worked yet, nor has OpenAI been embarrassed by abuse. So OpenAI is powering ahead here and everyone talks about their models, whilst Google's languish in obscurity. Like, Scott Alexander is talking about the limits of DALL-E 2 that Google claim they already solved, but it's irrelevant because Scott can play with DALL-E and not Imagen.
It takes 2 weeks to adapt to a loss though.
Humans are strange.
Lucky for me the guy on the other team was so positive and a much better bowler than me and he asked "That's ok, would you have been happy if I told you at the start of the game you'd have 107 in the 5th?" and I realized hell yeah I'd be happy. Same applies here. Of course I should be happy with a free $80 even if I know everyone else is getting $100.
That said, I think it’s also a compliment to a tool when it does such a good job, that people criticise it for not doing a perfect job. The fact it’s actually doing the job at all is pretty wonderful. I thoroughly enjoyed seeing the results.
https://en.wikipedia.org/wiki/Tycho_Brahe#Career:_observing_...
Edit: replacing 'Tycho Brahe' with 'Zombie Tycho Brahe' in the prompts could be one effective way to fix this
I think his nose was more or less skin-colored though, so the current pictures are okay. (DALL-E is supposedly programmed not to be able to generate real peoples' faces anyway.)
Also: if ML tech stalled now I expect we'd grow to have many Query Engineers, but if it doesn't do we still? For how long is this sort of experimentation what getting the best performance out of an AI will look like?
"And when your surpassing creations find the answers you asked for, you can’t understand their analysis and you can’t verify their answers. You have to take their word on faith—Or you use information theory to flatten it for you, to squash the tesseract into two dimensions and the Klein bottle into three, to simplify reality and pray to whatever Gods survived the millennium that your honorable twisting of the truth hasn’t ruptured any of its load-bearing pylons. You hire people like me; the crossbred progeny of profilers and proof assistants and information theorists…
In formal settings you’d call me Synthesist."
—Peter Watts, Blindsight (2006)Like, since perfect synthesis is undetectable; it just sounds like what you're hearing is straightforward common sense. You only notice the process when it fails. So for sure, the book could have played with that a lot more, especially since Siri's tools are meant to be broken towards the end, when he's implied to be penning this. I think there's bits where the book suggests that's what's going on ("I can't tell you what it means, only what it says"), but it never came across well to me.
But then again, maybe that's just how a failure of synthesis would look. Like, you still want to understand what the book is saying, but if the mechanism is working it's invisible, and if it's not working, then you're still relying on it working to produce the thing you're reading. So it's just inherently hard to shine a light on the thing itself with itself, which of course plays into the central theme as well. Ie. because Siri is implied to not be fully conscious, it's hard for him to catch himself in the act of translating for us readers.
(Of course we may say that a book designed to be hard to read is still hard to read.)
I don't see how anybody could reasonably hire someone for this position under this condition. How would one possibly discern between someone who can truly suss out what's going on and a techpriest who just recites learned dogma? I imagine it will definitely be an important skill, but I imagine that just like 'knowing how to google', it will be something that's just implicitly attached to a normal job.
Some people are better at it than others if you just watch the hashtags on Twitter. There will very clearly be value derived from it before we understand it. It’s very likely, in my opinion, that people will be hired for this type of work.
[0] https://theglassroom.co.nz/wp-content/uploads/2019/04/23f059...
Edit: Actually, I see what you were referring to now. The two pieces at the top of each window in the door. I can't say that I know how that was accomplished. I've never seen something like that before. It might be revealing to see an image with a higher resolution.
"nobody depicts a moose in stained glass. A man scrying the heavens through a telescope is exactly the sort of dignified thing people make stained glass windows about. A moose isn’t. So DALL-E loses confidence and decides you don’t really mean it should be stained glass style."
If the styles of the moose and the man were reversed, you could use the exact same explanation but end with "So DALL-E decides you don't really want the moose to be in stained glass style.".
So there is no need for pluralised word.
And we are back to moose!
I think this comment under The Charles Darwin Finch is one of the funniest things I've read this year.
[1] https://github.com/nerdyrodent/VQGAN-CLIP [2] https://github.com/CompVis/latent-diffusion [3] https://imgur.com/a/DjQYLUz
* https://jossi.avkrok.net/Escher_0.png
* https://jossi.avkrok.net/Escher_1.png
* https://jossi.avkrok.net/Escher_2.png
The prompt was "A stained glass image of M.C. Escher working in his studio in non-Euclidean space", 600 iterations, 1024x768, and these are the first three ones it came up with. It's not perfect, but it definitely knows who he was!
I.e. no deepfakes…
I assume from the "this X doesn't exist" genre that these are in some sense new creations, but I'm slightly suspicious that they're just a real photo with some very subtle tweaks. Is there a write up that dispells this notion and/or addresses the legal/copyright aspects?
How different does a face need to be from an original before you can legally use it in pornography seems like a question that will come up soon if it hasn't already.
Should DALL-E play around with interpretations or not?
It seems to operate like a huge multidimensional map of an image space with some verbal tags stuck into it rather randomly, and with style transfer applied as an operator to merge different features in the results.
What it needs is an equivalent map for the concepts behind the signposts, and also a mapping between concepts and images.
The concept map would be very hard to train because a concept space that includes all of culture and history would be huge and messy.
A large concept map with a huge set of associations is one of the features of educated humans. But the concept mesh is grounded in subjective experience, so it would need an equivalent layer for that.
It's the CLIP text and image encoders of DALL-E 2.
> and also a mapping between concepts and images.
It's the "prior" module.
You guessed right but DALL-E 2 already has these functionalities in its design.
https://i0.wp.com/bdtechtalks.com/wp-content/uploads/2022/04...
Also, if you could feed it some kind of "knowledge graph" of words and images rather than a sentence, you wouldn't be misunderstood the same way, but you'd presumably run into the limit of what its embeddings can contain.
Yes there needs to be an evaluator to guide and contextualize the results, but when most of the mechanical parts can be done so quickly (often most of the effort of the creative process), that's bound to lead to far higher output and willingness to get creative work done. You're right that it in itself is not strictly creative because it has no intentionality, but I've no doubt it will lead to more creativity.
There's some connection but they're largely separate.
Whether something is on a good path or not is part of the curation process.
There's a lot of 'creativity' to showcase, but that doesn't seem to be what's causing these failures.
I guess 'creative' will always have intentionality built into it. Maybe an imagination unconstrained by purpose (e.g. dreaming) is not creative either.
This happens with or without creativity.
There are failure modes that showcase creativity, and there are failure modes that don't showcase creativity. These seem like they're largely the latter to me.
I'm unfamiliar with this area, but it seems impressive to me that a lot of astronomical quirks are resolved:
Also, NASA has flown shit to pretty much every planet. You think they could do that if the heliocentric model was wrong?
You get what you pay for.
Anyway - I’ve lurked both for ~6-7 years. The best summary I can give is that Scott is an excellent and insightful author at his best, prone to certain idiosyncrasies, but who attracts a pretty thoughtful and generally pleasant commenter community.
The bias in astrange’s position is grounded in some factual truth, but I would argue that Scott’s commenter community is not entirely in the “rationality” camp, even if that’s a big part of how Scott reached his audience. The bigger part, imo, are the incredible series of articles he wrote circa 2014-2016 on a breadth of topics, many that still hold up.
I’d say check it out for the articles and see how you feel about the comments. Take astrange for what it’s worth, with a liberal dose of salt
It's part of "online rationalism" which is supposedly a community based around lesswrong where they practice thinking about things logically, with some techniques that work and some that don't.
More accurately it's a new religion that ascribes untrue magical powers to Bayes' theorem, probability theory, computers and AI research, gives people a life mission called "effective altruism", thinks if you program computers really well they'll become "superintelligences" and take over the world which needs to be stopped, and so on. This is why he says things about "rationality" in the post and is so approving about that Yudkowsky guy.
SSC years ago had some undeserved drama where the kind of online humanities-degree activist types who call you "tech bros" yelled at him, so I think the community also has people intentionally trying to make themselves into their image of a "tech bro" as a reaction to that.
(nb "religion" is not meant as an insult or a metaphor, starting new religions is what people do in Berkeley and this is just another one of them.)
Untrue magical powers? Just because people appreciate those things it doesn't mean they ascribe magical powers to them. That's like saying that HN ascribes magical powers to startups or chip architectures.
They invoke Bayes' theorem for its property of being optimal, but it's optimal (and trivial) when you know the priors. The hard part is knowing those, and they don't, so they just use it in a metaphorical/mindset way that doesn't actually mean anything, pretend they don't have actual beliefs so they don't need to take responsibility for them, and say "prior" a lot. (https://metarationality.com/bayesianism-updating)
In general, using probability theory as a way of thinking doesn't work because it can't reject infinitesimally likely but infinitely bad futures, which is why they're overly concerned with imaginary future AIs torturing them, and it ignores "unknown unknowns" and can only handle enumerated futures you already thought of, which is why they have that odd tendency to start polycules because it seems logical and then their girlfriends unexpectedly leave them.
(That's also how you know it's a Berkeley new religion - weird sex practices.)
>In general, using probability theory as a way of thinking doesn't work because it can't reject infinitesimally likely but infinitely bad futures
Another thing frequently discussed in those circles.
>(That's also how you know it's a Berkeley new religion - weird sex practices.)
Some people being sexually adventurous in California, the horror.
That’s the problem, they shouldn’t be discussing it ;)
> Some people being sexually adventurous in California, the horror.
I could’ve said “joining cults like Leverage” instead of polycules but that would be mean. However like I did say, the problem with being a young man who gets into an open relationship because you rationally determined it was optimal is, you’re wrong and your girlfriend is going to leave you.
If you’re going to have weird sex, at least have a more fun reason for it. (There’s a tweet I want to link here about a rationalist nude sex party where they all sat around discussing immigration politics instead of having said adventures, but unfortunately I can’t remember how to spell Aeolla to find it.)
If you had the right priors to start, you wouldn't need probability theory because you'd already know everything.
> it can't reject infinitesimally likely but infinitely bad futures
Sure it can, because that's value theory, not probability theory.
As the page I linked says, the real lesson is “don’t be too sure of stuff” but that doesn’t require pretending you’re doing math or joining any religions.
But I was going to build a custom religion designed to make motte & bailey arguments in favor of it easy, it would probably look a lot like "rationalism". "It's not a religion, we just apply these precepts of rationality rationally! Look, the rational application of rational principles, rationally, is all but in the name!"
I actually agree that not everyone there treats it like a religion. There are people like me who just cruise by, pick some fruit, learn some things, apply some useful things, and get on with it. But there are absolutely also people for whom it is their religion.
Ironically, as IIRC some of the thought leaders have themselves pointed out (usually in the form of an agonized, hair-pulling sort of "Why, Bayes, why?" sort of post), there is very little rational reason to believe that improving rationality will lead to a lot of massively better life outcomes; the evidence of this is thin on the field. (Picking up some things around the edges, absolutely. But the hope that you suddenly Solved Life because you're "more rational" seems to be pretty poorly founded in reality.)
> improving rationality will lead to a lot of massively better life outcomes
IMO, to defend this, I'd argue for most people there's maybe oom of ten decisions over a lifetime that in hindsight could turn out to be really important, and if you can keep a calm head and think about it in those moments, you can potentially get a lot of benefits. Now usually there'll be lots of advice for those, and the "rational" (in the advertised, not actual, sense) answer is not always the best one, but even just stepping back, thinking about it, trying to find out how to gather evidence about it, can already possibly help out a ton.
Most people get by fine doing well enough. And it doesn't take rationality to tell you not to invest your savings in scams, so the upside is limited. But for those few times where it really matters, there are some genuinely useful tools in there, IMO.
(Though personally, I'm in it for the philosophy.)
Unfortunately there isn't such a thing as an ideal thought process - that is, if you want to make a correct prediction about the world, there is no demonstratably correct method of doing that. This sort of thing has been tried before (logical positivism) and disproven (Wittgenstein and Gödel's theorem). It's notable that AI research is mostly getting places by just doing gigantic matrix multiplications instead of this.
It can be a useful exercise but it's not a good life philosophy - it's a bit robotic. And it gets cult-like if you add things on like "…and that's why you should live in our group home" or "…and that's why you should get rich and donate it all to evil AGI research".
This isn't very important to me though, it's just something I've noticed. It seem to come up a lot more lately because people reply to every ML story with "oh no the AIs are going to take our jobs" stuff, but still the largest impact it's had on the real world is Roko causing Elon Musk to date Grimes.
I think I meet more people who've developed life philosophies in reaction to knowing about LW ("post rats" or "meta rats") - mostly they explicitly decided to join religions instead of accidentally doing it - and for some reason all of them became Buddhists. That one surprises me.
AIXI has some correctness proofs. In an admittedly esoteric sense, it does make "optimal use of information". I do think those parts philosophically hold up.
For an interesting view on Gödel, see Garrabrant's logical induction paper; you can bypass some aspects of that by working probabilistically.
> "…and that's why you should live in our group home"
People try things because they think they work, and that's why LW is a cult.
> "…and that's why you should get rich and donate it all to evil AGI research".
People recommend things because they think they're correct, and that's why LW is a cult.
Also, anti-evil AGI-research. That's kind of an important aspect.
I don't know what's going on with the postrats either.
I'd prefer a real world stained glass robot though.
Even if it can't do the physical work (which can't be too hard), where a program told you what colors to buy, or you told it what you have, and it designed using that and printed/projected out a numbered tracing where to cut.
There's a Dall-E value add, turn it into a stained glass pattern and sell them.
When people talk about UBI and artists doing stuff... we have the tech now to tell a computer what we want and have a stained glass window come out of a factory. Maybe more expensive than humans today.