I used DALL·E 2 to generate a logo
jacobmartins.com
jacobmartins.com
It seems like the most valuable thing this could do is get some of that early exploration out of the way faster and easier than a human can do it, get to two or three concepts that feel like they’re in the neighborhood, and then let a human expert take over and turn it to something final quality. That’s pretty cool.
At the end of the article I also described a bit how I would see the evolution of such a tool, and it looks like we're thinking very similarly.
---
Though I think the real breakthrough will come when Dall-e gets 10-100x cheaper (and faster). I would then envision the following process of working with it (which is really just an optimization on top of what I’ve been doing now):
1. You write a phrase.
2. You are shown a hundred pictures for that phrase, preferably from very different regions of the latent space.
3. You select the ones best matching what you want.
4. Go back to 2, 4-5 times, getting better results every time.
5. Now you can write a phrase for what you would like to change (edit) and the original image would be used as the baseline. Go back to 2 until happy.
Do you like this? What about this? You simply nod or reject the solutions that you don't want.
Pretty soon somebody's expertise and experience is not going to be enough to continue paying them what they used to get before this magic blackbox appeared.
One day enterprises will realize they can just outsource that expert who's been reduced to simply typing prompts and nodding yes or no.
I am worried that the middle class is rapidly disappearing. We will own nothing and be happy seems quite ominous. The question is then what field is safe from advancements in AI?
The only field I can think of is doctors, lawyers, executives, buy-side money managers. Even their jobs will be partially automated but it will be safe as long as they generate revenue.
Every art director at an ad agency just shrieked!
This is ok for a logo like this where it’s fair to say the base level expectation is not super creative. This logo is cool, but it doesn’t really stand out or make the product ver distinctive. If I am running a hobby or OS project that’s fine, but if I was investing a lot in sales/marketing then paying a real artist to make something interesting and novel is a rounding error.
Q: Are there really logos out there that are "interesting and novel" and that "stand out or make the product [..] distinctive"? Which ones?
EDIT: (perhaps more importantly) are there interesting, novel, distinctive logos that actually contribute to profitability?
A lot of GPT iterations of the design has left the article author with something which is quirkier than your average logo, but also looks like clipart and probably doesn't scale up or down well or work in monochrome. Which is fine for OSS. (He might get more users from blog traffic about using GPT-3 to design his logo than he ever could from any other logo anyway)
But when it comes to bigger companies, the design agency are the people that sit in meetings with execs persuading them that a well chosen font and a silhouette of a much simplified octopus will work much better ("but maybe the arms could interact with some of the letters etc etc, now lets discuss colours). The actual technical bit of drawing it is the bit that's already relatively cheaply and easily outsourced, and plenty of corporate logos are wordmarks that don't even need to be drawn...
That's not utopianism. The new jobs can't always be filled by the people kicked out of jobs. It really sucks to be them.
But it does mean that it's not irrational for people to want to automate other people's jobs. The net amount of stuff generated increases, rather than decreases.
This pattern may not last forever. There's already some thought that we've generated more than enough stuff to guarantee a decent standard of living to everybody (at least in the developed world) without working, and plenty more for luxuries if people choose to work. Even if we haven't reached it, we appear to be heading in that direction sooner rather than later.
That may cause a radical re-think at some point. And it won't be seriously delayed by making sure cartoonists have jobs.
It's not a zero sum game. There's still growth in us. We'll go to space and expand 1000x more, the space has plenty of resources, and humans will have jobs working together with AI.
Q: Am I the only one thinking of Golgafrinchan Ark Fleet Ship B?
In the past, fast automation has led to badly distributed wealth, and job loss. This situation has lasted until the unemployable people died off (yep, that was part of it), and enough wealth was redistributed through violent means.
Today we know better, and have really no reason to repeat the violent means of our previous revolutions. But it's really looking like the people in power want to repeat them.
there were no instances of violent redistribution of wealth that ended better for the average person than before. Only that a different group of people ended up with wealth.
Automation makes stuff cheaper, even for people who didn't obtain any of the financial wealth via redistribution - because there's more than just financial wealth that get created with automation. New availability of services and goods (think internet of today - this is a wealth that couldn't have existed before, and one can benefit from it even if they are poor today).
I have a few qualms with this app:
1. For a Linux user, you can already build such a system yourself quite trivially by getting an FTP account, mounting it locally with curlftpfs, and then using SVN or CVS on the mounted filesystem. From Windows or Mac, this FTP account could be accessed through built-in software.
2. It doesn't actually replace a USB drive. Most people I know e-mail files to themselves or host them somewhere online to be able to perform presentations, but they still carry a USB drive in case there are connectivity problems. This does not solve the connectivity issue.
3. It does not seem very "viral" or income-generating. I know this is premature at this point, but without charging users for the service, is it reasonable to expect to make money off of this?
Edit: Ahh, it’s the Dropbox comment of HN fame. Never mind.
Just like in Star Trek. They really knew what the end goal was didn't they.
> enterprises will realize they can just outsource that expert who's been reduced to simply typing prompts and nodding yes or no
Tbf a program averaging the market for a fact gives better returns than most of the financial industry, yet they still exist. Even if we can automate something doesn't mean we will, usually for pointless emotional reasons.
But on the other hand it's hard to say if in a 100 years humans will still be employable in any practical capacity for literally anything.
Some-when in the 00's I read an article about him that he was putting advanced networking stuff into the castle and had the intention to start something like a "think-tank" (doesn't really fit it, but I don't know what I'd call it) where he and others would hang around and code stuff.
I found the article [1] from July 2002, "Lord of the Castle Kai Krause presents Byteburg II".
> So that 's Kai Krause's long-cherished plan: Now the software guru has finally opened a center for founders and developers from the IT and software industry in Hemmersbach Castle near Cologne -- the Byteburg II
I really wonder what he's doing to these days. His plug-ins were legendary, as well as the User Interface for Bryce [2]
[0] https://de.wikipedia.org/wiki/Burg_Rheineck
[1] https://www.heise.de/newsticker/meldung/Schlossherr-Kai-Krau...
[1, google translate] https://www-heise-de.translate.goog/newsticker/meldung/Schlo...
Some really interesting reads. I especially appreciated his articles on the passing of Douglass Adams (apparently a close friend of his!) and Then vs Zen.
Love him or hate him (and I do both), Kai was all about cultivating his adulating cult of personality and dazzling everyone with his totally unique breathtakingly beautiful bespoke UIs! How can you possibly begrudge him and his fans of that simple pleasure? ;)
In the modest liner notes of one of the KPT CDROMS, Kai wrote a charming rambling story about how he was once passing through airport security, and the guard immediately recognized him as the User Interface Rock Star that he was: the guy who made Kai Power Tools and Power Goo and Bryce!
Kai's Power Goo - Classic '90s Funware! [LGR Retrospective]:
https://www.youtube.com/watch?v=xt06OSIQ0PE&ab_channel=LGR
>Revisiting the mid 1990s to explore the world of gooey image manipulation from MetaTools! Kai Krause worked on some fantastically influential user interfaces too, so let's dive into all of it.
>"Now if you're like me, you must be thinking, ok, this is all well and good, sure, but who the heck is Kai? His name's on everything, so he must be special. OH HE IS! Say hello to Kai Krause. Embrace his gaze! He is an absolute legend in certain circles, not just for his software contributions, but his overall life story." [...]
>"... and now owns and resides in the 1000 year old tower near Rieneck Castle in Germany that he calls Byteburg. Oh, and along the way, he found time to work on software milestones like Poser, Bryce, Kai's Power Tools, and Kai's Super Goo, propagating what he called "Padded Cell" graphical interface design. "The interface is also, I call it the 'Padded Cell'. You just can't hurt yourself." -Kai
But all in all, it's a good thing for humanity that Kai said "Nein!" to Apple's offer to help them redesign their UI:
http://www.vintageapplemac.com/files/misc/MacWorld_UK_Feb_20...
>read me first, Simon Jary, editor-in-chief, MacWorld, February 2000, page 5:
>When graphics guru Kai Krause was in his heyday, he once revealed to me that Apple had asked him to help redesign the Mac's interface. It was one of old Apple's very few pieces of good luck that Kai said "nein"
>At the time, Kai was king of the weird interface - Bryce, KPT and Goo were all decidedly odd, leaving users with lumps of spherical rock to swivel, and glowing orbs to fiddle with just to save a simple file. Kai's interface were fun, in a Crystal Maze kind of way. He did show me one possible interface, where the desktop metaphor was adapted to have more sophisticated layers - basically, it was the standard desktop but with no filing cabinet and all your folders and documents strewn over your screen as if you'd just turned on a fan to full blast and aimed it at your neatly stacked paperwork.
The Interface of Kai Krause’s Software:
https://mprove.de/script/99/kai/index.html
>Bruce “Tog” Tognazzini writes about Kansei Engineering:
>»Since the year A.D. 618 the Japanese have been creating beautiful Zen gardens, environments of harmony designed to instill in their users a sense of serenity and peace. […] Every rock and tree is thoughtfully placed in patterns that are at once random and yet teeming with order. Rocks are not just strewn about; they are carefully arranged in odd-numbered groupings and sunk into the ground to give the illusion of age and stability. Waterfalls are not simply lined with interesting rocks; they are tuned to create just the right burble and plop. […]
>Kansei speakes to a totality of experience: colors, sounds, shapes, tactile sensations, and kinesthesia, as well as the personality and consistency of interactions.« [Tog96, pp. 171]
>Then Tog comes to software design:
>»Where does kansei start? Not with the hardware. Not with the software either. Kansei starts with attitude, as does quality. The original Xerox Star team had it. So did the Lisa team, and the Mac team after. All were dedicated to building a single, tightly integrated environment – a totality of experience. […]
>KPT Convolver […] is a marvelous example of kansei design. It replaces the extensive lineup of filters that graphic designers traditionally grapple with when using such tools as Photoshop with a simple, integrated, harmonious environment.
>In the past, designers have followed a process of picturing their desired end result in their mind, then applying a series of filters sequentially, without benefit of undo beyond the last-applied filter. Convolver lets users play, trying any combination of filters at will, either on their own or with the computer’s aid and advice. […] Both time and space lie at the user’s complete control.« [Tog96, pp. 174]
METAMEMORIES:
https://systemfolder.wordpress.com/2009/03/01/metamemories/
>Anyone who has been using Macs for at least the last ten years will surely remember Viewpoint Corporation’s products. No? Well, Viewpoint Corporation was previously MetaCreations. Still doesn’t ring a bell? Maybe MetaTools will. Or the name Kai Krause. Or, even better, the names of the software products themselves — Kai’s Power Tools, Kai’s Power Goo, Kai’s Photo Soap, Bryce, Painter, Poser… See? Now we’re talking.
Macintosh Garden: KPT Bryce 1.0.1:
https://macintoshgarden.org/apps/bryce-1
>Experienced 3D professionals will appreciate the powerful controls that are included, such as surface contour definition, bumpiness, translucency, reflectivity, color, humidity, cloud attributes, alpha channels, texture generation and more.
>KPT Bryce features easy point-and-click commands and an incredible user interface that includes the Sky & Fog Palette, which governs Bryce's virtual environment; the Create Palette, which contains all the objects needed to create grounds, seas and mountains; an Edit Palette, where users select and edit all the objects created; and the Render Palette, which has all the controls specific to rendering, such as setting the size and resolutions for the final image.
MACFormat, Issue 23, April 1995, p. 28-29:
https://macintoshgarden.org/sites/macintoshgarden.org/files/...
https://macintoshgarden.org/sites/macintoshgarden.org/files/...
>He intends to challenge everything you thought you knew about the way you use computers. 'I maintain that everything we now have will be thrown away. Every piece of software -- including my own -- will be complete and utter junk. Our children will laugh about us -- they'll be rolling on the floor in hysterics, pointing at these dinosaurs that we are using.
>'Design is a very tricky thing. You don't jump from the Model T Fort straight to the latest Mercedes -- there's a million tiny things that have to be changed. And I'm not trying to come up with lots of little ideas where afterwards you go, "Yeah, of course! It's obvious!"
>'Here's an easy one. For years we had eight character file-names on computers. Now that we have more characters, it seems ludicrous, am historical accident that it ever happened.
>'What people don't realize is that we have hundreds more ideas that are equally stupid, buried throughout the structure of software design -- from the interface to the deeper levels of how it works inside.'
In other words, if another person needed a logo and used the same phrase how long on average until they get a duplicate of your image?
DALL E 2 is like a low or no-code tool in that way.
The outcome may not be a "finished" product, especially as viewed by a professional designer (or web dev). However, its a heck of a lot better than a tersely written spec.
And in some cases, the product will work well enough to unblock the business, get customer feedback and generally keep things moving forward.
editing: not that it's intentional, but these things will have the same effect; way too much product even for creative works. No one will be able to make money off the product but the tools.
As someone who is incredibly terrible at graphic design but knows what they like this could be a game changer as iterations of this technology progress. I can imagine going further than images and having AI/ML generate full HTML layouts in this iterative way where you start to define your vision for a website or app even and it spits out ideas/concepts that you can "lock" parts of it you like and let it regenerate the rest.
I'm not downplaying designers role at all, I'd still go to one of them for the final design but to be able to wireframe using words/phrases and take a good idea of what I want would be amazing, especially for freelance/side-projects.
I think your art/design/craft is pretty good. Some people use pencils, some use Adobe products, you have gone out there and tried the new Dall-E medium.
Glad you thought out the usage, I am sure that when the novelty wears off that you will have that neat-as-octocat logo sorted out.
I appreciate that you appreciate the value that highly skilled designers bring to a product with their visual expertise.
However, I would like to see you A/B test the Dall E logo versus the winning designer logo. You could show odd IP addresses one logo and even addresses the other.
I think the designer would edge the robot for what you need (a logo), however, the proof is in the pudding and conversion rate.
If you lined up 100 resulting images, 99 from weekend beginners and 1 from an actual artist. I guarantee you that you would pick out the artist every time.
It might be simple to trace over an image but you are probably better getting an artist to spend 2 hours on it, it will most likely look better than 2 weeks of tracing.
People are already doing by combining DALL-E 2 with gfpgan for face restoration. So there may be a role in understanding how to combine these tools effectively.
An experienced human designer, right away, is going to ask how you want the logo to be used. That's going to have a major impact on how it's designed.
So yeah, this may be like working with a doodler, but, as the author intimated, this is far from an ideal experience in getting a professionally designed logo. This is more like "Hey, you, drawing nerd, make this thing."
Nevertheless, astonishing technology in its own right.
There will probably be less need for designers of 'lower quality' simple images though.
If dale decides what we see, it might become what the next generation likes and considers “good taste”.
Taste is very complex: it's hierarchical, social, not fixed, not absolute, not rational, is specific to audience and has irregular overlaps across groups, much of it (all?) derived from human sensation and context-specific situations.
The path to something being considered as good taste is generally not simple: much of it flows through lines of power/desire/moment whose branches are not easy to trace as they're being formed. Much of taste is the hidden "why" which most of us never see.
It's realistic that Dall-E could understand what trends are on the rise, or in good taste … it's much harder to say if Dall-E could create something of originally good taste.
Assuming the status quo, true. As we evolve our lives around emerging AI tech I think we will at first be the curators and creative directors of AI, but eventually a creative agency will defer to the AI as it knows more about our tastes, market, audience, and the ENTIRE HISTORY of art, design, marketing, tastes, trends, and so on.
Eventually it won't make sense to have a stupid human rubber stamp what the all powerful AI suggests. Just as it does not make sense for Facebook to curate news feeds.
Maybe one day product advertising will look different depending on who looks at it. Pepsi logo "just for you".
That aside, a great use of these tools is to generate N spit-takes of wildly varying styles that you can present to the customer very quickly and very cheaply. Once you pin them down to a particular range of styles you can get down to the carving out the details by hand.
Right now, the input to DALL-E is all human generated.
What will happen is that DALL-E will generate something "close enough" that gets used and promulgated, so now the input to DALL-E will become increasingly contaminated with output from DALL-E.
We're already starting to see this in search engines where you get clickbait that seems to be GPT-3 generated.
https://en.wikipedia.org/wiki/Semi-supervised_learning
https://towardsdatascience.com/semi-supervised-learning-how-...
When you run a phrase, you get four images. Those images will stay in your history, but the ones you like you will save with the "save" button, so that they're in your private collection.
With this, you already have a great feedback system: saved - good, not saved - bad.
I'll keep it mind, as I might still end up choosing a different one.
The chosen one is closer to my original vision, but you do have a point that the yellow ones look more polished.
Also, since time immemorial, databases are cylinders and data comes in cubes.
For logo purposes, these are both strong, while the second adds “personality”:
https://i.imgur.com/j6P4Oh4.jpg
https://i.imgur.com/kM23GZV.jpg
I really like the design breaking out of the strong circle, and your hard hat idea was great. That last one could have been your logo “as is”!
Though you could consider replacing the green cubes with cylinders, or simply hand add rubix cube lines to these green cubes to make them data cubes.
https://duckduckgo.com/?q=data+cube&t=ha&va=j&ia=images&iax=...
Thanks for sharing the process!
Those iterations suck. I'm not worried for my colleagues and I.
That being said! Many, MANY clients have questionable taste, and I can, indeed, see many who aren't sensitive to visuals to be more than happy with these Dall-E turd octopus logo iterations. Most people don't know and don't care what makes good graphic design.
For one thing, that final logo can't scale. For another, the colors lack nuance & harmony. The logo is more like a children's book illustration, and not something that is simple, bold, smart, and can be plastered on any and all mediums.
Just my 2 cents.
I bet in another 10-15 years, though, things might get a bit dicier for fellow graphic designers/ artists/ illustrators, though, as all this tech gets more advanced.
A more obtuse example, how many lift operators do you see today?
For those who do not know the reference: https://m.youtube.com/watch?v=BKorP55Aqvg
The thing you’re missing is AI generated content can be refined by AI. If Disney promised their meh looking movie would improve on its own over time, people would be line to it because it’s new, not just streamlined copy-pasted design we see all over media now
Painting the Titanic wasn’t the hard part. The hard part was organizing the process that produced its structure. That’s were AI content is now.
We’re generating the bulk structure pretty competently at this point. Refining the emotional touches will come faster.
Ultimately the average person (who is likely the target audience anyway) won't notice anything wrong with most of those iterations and given that they're basically free in comparison would make me worried. I wouldn't be surprised if they manage to make it output svgs soon.
Weirdly, with the advent of AI, we might start to see exactly what it is that makes human beings special.
I never understood this logic, where the creator of something does something seemingly stupid and people are like "Well, don't use their project then if you don't like it". Instead of constructively calling the problem out, so the creator can try to make it better.
If my logo sucked, I'd like people to please tell me...
Because they define what the success is. If their goal is to make money they may want a logo which is the closest to optimal for getting clicks. If they want a private project, they may want it to be fun. And many other scenarios... You're welcome of course to do constructive criticism, but in the end it's up to them if they want to apply it.
However, once you reach a certain budget, it's much more involved to *choose* a logo that "fits" how the company wants to present itself, than it is to generate candidate logos of sufficient quality. I can assure you that the "many-chefs problem" for a high budget design project is very real, and the major cost driver. You have a mix of "design by committee", internal politics, what designers wants on their portfolios, etc etc.
That's a long time. I expect within a decade or two, "AI" should be able to generate an entire animated movie given nothing but a script.
I don't think it would be technically hard to build a model with current technology which can generate logos with the attributes which you mentioned. You could simply fine-tune a Dalle-E style model specifically on a smaller dataset of logos. This would just take a small dedicated team of domain experts to work on the problem.
I will say, though, I think DALL-E has opened up a new market for artists. I've gone to freelance graphic designers before, and been generally happy with the results, but it's pricey. So pricey that I honestly can't justify it for a new project I intend to sell or for an open source project I don't expect to make money from. It's usually much more cost-effective to even hire lawyers or even UI/UX people.
If I were an artist, I'd be experimenting with DALL-E, trying to run my own pirate version and learning everything about it. An artist empowered with DALL-E could give quick options to a client, iterate with them quickly, and test out some ideas before making the final work product. I'd guess a good artist who made good use of DALL-E could get a project done much faster and cheaper, and this would likely mean a lot more people hiring artists (if I could spend $100-200 for high-quality assets within a few days rather than $1000-2000, I'd gladly hire artists frequently).
I'm sure this will make some artists feel cheapened, but the reality is that art & technology have always evolved in dynamic and unpredictable ways. ML being essentially curve-fitting means that genuine inspiration and emotion is still far beyond our capabilities today, and that, ultimately, these models will only give us exactly what we ask for. A good (human) artist can go beyond that.
EDIT: Also, I agree with your assessment of the "work product," if we can call it that. I was unimpressed with the iterations, and especially the final product. I guess it's good the product is an open source tool. Nothing about the generated logo helped me understand what the OctoSQL tool did. Honestly, the name (which also IMO isn't excellent) is much more evocative than that logo. Why is the octopus wearing a hard hat? Why is it grabbing different colored solids? I guess the solids are datasets? But then the octopus is just exploring them? No thanks.
I can't think of a single well known logo that is even remotely close to what a company's product is. Photoshop, Firefox, Chrome, Microsoft, Facebook, Apple, Netflix, McDonalds, Ford, Ferrari, Samsung, Nvidia, Intel, RedHat, Uber, Github, Duolingo, AirBnB, Slack, Twitter, IntelliJ, Steam.
I guess the Gmail logo does tell you it has something to do with mail though, so I did find one example.
So whereas Ford's brand is just a name, "Mustang" has a logo that really does tell you something about the car. You kind of understand when you see the galloping horse what it's meant to do.
Intel brands its CPUs with the name inside a square, which is colored to resemble (abstractly) a CPU.[0]
And Photoshop once had a logo that communicated what it did.[1]
As a brand becomes more established, it tends to be more abstract. Whereas Starbucks was once an elaborate siren (I interpreted it to be the siren call of espresso), details have been simplified over the years.[2] This is similar to the Photoshop magnifying glass logo becoming "Ps".
After the Apple I and Apple II, Apple sometimes used apple varieties (plus Lisa) to brand it's products (e.g. Macintosh, Newton). However, this largely stopped in the late 90s when Steve Jobs returned. Macintosh was shortened to Mac, and 'i' was prepended to various product names. Most new ones were descriptive e.g. iPod, iPhone, iPad, Apple Watch. The computers have retained "Mac" in the branding, along with "book" for notebooks (a convention predating Steve's return). The logos for all of these are just the names of the products typeset in its own San Francisco font; whenever Apple appears in a product name, the Apple logo is used instead.
So, yeah, I think it's reasonable to communicate what a product does or why a project exists with its logo. I didn't really see that w/ OctoSQL.
EDIT: I should also address Firefox & Chrome.
Firefox started as Phoenix (i.e. rising from the ashes of Netscape Navigator/Mozilla). Phoenix had a trademark conflict, so was renamed Firebird. This also had a conflict, and Firefox was chosen after. In the Zeitgeist of the early aughts, Phoenix made a ton of sense: instead of extremely bloated chrome around the page as had been prevalent in Navigator and Internet Explorer, Phoenix gave you a tab bar (truly revolutionary), the navigation bad and the bookmarks bar. It was simple and clean, like a reborn Phoenix.
Chrome is interesting because the name is not related to traveling or navigation. It's telling you it's just the container for what you care about. But the logo is a bit more like a sphincter or an all-seeing aperture. I've never gotten the logo for Chrome outside a spyware context, but it has become successful.
[0] https://www.intel.com/content/www/us/en/products/details/pro...
[1] https://logos-world.net/wp-content/uploads/2020/11/Adobe-Pho...
[2] https://miro.medium.com/max/2418/1*tJf7O6FPOmnErngygbBQDQ.pn...
But isn't the logo created for most people? Does it matter that, you as a designer, think it's bad if most people don't? I see it like modern fashion shows. I look at them and think the clothes are insane and I would never wear them, but obviously other fashion designers think they look good (I'm guessing?).
I do agree that the logo isn't super practical though, it's too textured and won't scale. I would take it to /r/slavelabour or Fiverr and pay someone to vectorize it and see what they come up with.
I understand your argument but I don't think that's the problem - the problem is that even most users don't understand what a good logo looks like (even if they like them) the same as users don't know what they want. It's a known fact that you shouldn't ask users of a software how it should be designed because if you'd let them design a software they want it would be shit.
What's really interesting about this class of AIs is that they unbundle the two and you can play with them independently for the first time.
In my personal opinion as an (admittedly junior) ML engineer and lifelong artist, we've got <10 years before the golden age of human-made art is completely over.
If nothing else, inspiration is just a click away. No more searching for ideas, just talk to the AI and it will pump out numerous ideas for you.
Even if you get close to what you, the human, may like--it's difficult if not impossible to articulate what you like about it and iterate. Black box, keep trying random keywords... May as well grab a marker (read: hire a human)
The article says that octopi is the plural of octopus, but it's actually octopuses. Octopus is originally Greek, not Latin and thus does not get the Latin plural -i, but instead would get the Greek plural -odes. Since it ends in a way English can deal with, the commonly accepted usage is octopuses (English) over octopodes (Greek) with octopi being the least correct.
https://qz.com/1446229/let-us-finally-resolve-the-octopuses-...
While “octopi” has become popular in modern usage, it’s wrong.
I would argue that it used to be wrong, but language, unlike physics and code, is what the majority say it is.I used to be a stickler for correct vocabulary usage and then I saw a documentary about dictionaries (can't remember what it was) and someone from OED said basically this (from https://www.oed.com/public/oed3guide/guide-to-the-third-edit...):
The Oxford English Dictionary is not an arbiter of proper usage, despite its widespread reputation to the contrary. The Dictionary is intended to be descriptive, not prescriptive. In other words, its content should be viewed as an objective reflection of English language usage, not a subjective collection of usage ‘dos’ and ‘don'ts’. However, it does include information on which usages are, or have been, popularly regarded as ‘incorrect’. The Dictionary aims to cover the full spectrum of English language usage, from formal to slang, as it has evolved over time.
Now I think it's something that is just fun to argue about, but I don't take any of it seriously.(edited for formatting)
Meanwhile, if you think that sounds interesting I’d highly recommend the documentary Helvetica.
I haven’t watched it, but the subject is fascinating.
I really dislike the latin plural rule, that some misguided but powerful people decided on centuries ago.
"Indexes" is much more natural English than "indices", and we should, when possible, use those those forms.
I don't think a particularly convincing reason was advanced other then "technical things are more Latin-adjacent".
To wit: A blog post from Merriam-Webster: https://www.merriam-webster.com/words-at-play/the-many-plura...
What a silly thing to say! Where does this poor fool think language comes from?
This is one of the cringiest Well-Actually-isms. It tries to look pedantic while completely missing the point.
The logo was created for OctoSQL[0] and in the article you can find a lot of sample phrase-image combinations, as it describes the whole path (generation, variation, editing) I went down. Let me know what you think!
And btw. if you get access take a look at [1] before you start using it. A ton of useful bits and pieces for your phrases.
TLDR: DALL·E 2 is really cool, though takes quite a bit of work to arrive at a useful picture. Moreover, some types of images work better than others ("pencil sketch" is consistently awesome). As with programming, it's difficult to realize how much pieces you have to specify if you're not an artist - you don't know what you don't know.
[0]: https://github.com/cube2222/octosql
[1]: http://dallery.gallery/wp-content/uploads/2022/07/The-DALL%C...
edit: found it in the article: "From a monetary perspective, I’ve spent 30 bucks for the whole thing (in the end I was generating 2-3 edits/variations per minute). In other words, not too much."
It gets expensive fast.
https://jacobmartins.com/images/dalle2/DALL%C2%B7E%202022-08...
I didn't prompt anything specifically, it came after a line of variations from a definitely-not-icon-looking picture.
Though I'd try tags like "iOS icon".
thanks for the writeup. I looked at your other blog posts and I would like to read more about octosql (needs/specification, architecture, development strategies, challenges, DBMS protocols/interfaces/libraries).
And thank you for adding outer joins after I recently mentioned that they are missing!
There is no technical documentation available right now other than the readme. I'm planning to write it around September-December (together with a website for them).
You can share your email at jakub dot wit dot martin at gmail and I'll let you know when it's available.
https://labs.openai.com/s/z1PVd5v6td9PsiY20Y5GdxDf | https://labs.openai.com/s/yxX49BjX07BztYgMjm49iXKc
[1] https://labs.openai.com/s/x2UP0MEmj2qNnKWTbko8rrso
2. “cute baby dragon, logo, digital art, in a dark circle as the background”
[2] https://labs.openai.com/s/JmOXAqjpR2ctmraDxEkB7twF
Thanks for this post, it helped me tailor my own search queries. Because of your post, I was able to discover a whole new realm to DALLE-2. For some reason, repeating the same query parameter at the end yields some rather interesting results.
Something strange about DALL·E is that if you just type gibberish by pounding randomly on your keyboard, it will still "work", i.e., produce an image.
Alternatively I guess it could just pull harder towards the prompt, idk.
The second is more 'advanced' to me than the first, possessing an actual style, but neither is anything I would consider high quality enough to serve as a project/company/site/personal logo.
Art will become a commodity. Human art and ai art will be indistinguishable, "artists" will become as common as "photographers" since the inception of digital photography and social media.
Movie and TV scripts will be iterative with a creative director and AI working together.
Animation will become a lot easier, less people needed, fewer creatives.
Software will become easier and easier as developers will simply guide AI. This is already beginning to happen, but imagine paired programming with natural language interacting with an AI.
Architecture, civic planning, engineering, medical, law, policy, physics, it's all gonna change, and rapidly. DALL.E 2 shows how a leap in sophistication can revolutionize an industry overnight. Microsoft has exclusively licensed DALL.E 2, I can only imagine the myriad of creative tools it will serve the creative industry with.
The working in real-time will be the biggest leap. Asking DALL.E for an image and refining it as you talk is going to be nuts.
There were still ~60% as many employed photographers in 2021 than in 2000 with higher real wages (data from BLS - https://www.bls.gov/oes/current/oes_nat.htm).
For camera operators, the employment is flat, again with rising real wages.
>imagine paired programming with natural language interacting with an AI
Mostly it will get in the way. AI "programmers" are only good if they are able to generate correct code from spec/pseudocode and in first 1-3 number of tries (otherwise it will be faster to write it yourself).
This is simply not true. I use GitHub Copilot and it's already made me faster and shows me ideas I would not have thought of myself. And that's just Copilot. When you can talk to an interface and say "I want to update the vote count by one when I click this button" I think you'll change your mind. The AI will know the entire codebase inside out, it will know the intention of all the code, all the data models, know how users use the application intimately, be aware of problems instantly, able to run hotfixes without user intervention. Got a slow query? No problem, here is some SQL that follows all the business rules and is 10x more efficient. And that's just a start. Every single aspect of software development from management, engineering, and marketing will all be transformed.
As for photographers I have 99% more friends and family pumping out thousands of high quality photographs than I did in 2000. Go look at all the professional looking shows made by regular folk on YouTube. To deny that camera phones transformed photography seems silly.
Regular folk have access to drones to do wild tracking shots in 4k that were only possible with helicopters and huge cameras 20 years ago.
The future is here, it's happening all around us so rapidly we have a hard time keeping up with how dramatic the changes are.
I can see the intent side of things, but I just can't see the 'glue' side of things as well.
Who says the AI will "know the intention of all the code, all the data models, know how users use the application intimately"? Are you aware that language models do in fact have token input/output limitations that will not go away? Are you aware that there is such a thing as diminishing returns when it comes to improvements due to increased number of parameters/training set size that are already evident? Are you aware that the training set of codex pretty much includes all available public code, so it will be impossible to scale it by a factor > 3 in the next several years at least?
Your assertions are full of wild assumptions backed by nothing.
As for photography, the fact is there has been no job apocalypse because your "friends and family" are "pumping out photos". And the point of your initial post, even if it was implicit, was "you are going to be unemployed in 5 years". This will have an impact on your dev flow and will be used by managers to try to reduce salary premiums for software engineering but your wild assumptions stated with so much confidence may never happen.
P.S: At this point, I find Intellicode actually slows me down, that's why it's permanently turned off. Current copilot will at most save me 2-3% of my working time each week if I am coding in a language it can actually do something in (it's worse than useless for Scala).
> Your assertions are full of wild assumptions backed by nothing.
I use Copilot, you admitted it currently saves you 2-3% of time. Well, that's just Copilot, you think Microsoft will just sit on that? My assertions are based on what is happening today and extrapolating an exponential increase in that performance for tomorrow.
Digital cameras definitely revolutionized photography and made it much more accessible to regular folks. Not everyone wants to be a pro photographer though, and the number of wedding shoots available has not changed. People still need to be paid to take photos because no one is going to do that for free. However, we can all take pro level photos with much more ease than when all we had was 110 and 35mm film with a really crappy lens.
There are more "photographers" than ever, the same number of pro photographers seems reasonable given the burden of people's time to money ratio. So the net result is billions more family and friends photos which previously were not taken, the same will go for art. I want to create art, but I have little skill, but given the opportunity to make a comic strip just by talking to an AI will allow me to do so. I imagine some people will do this extremely well as a profession until it's no longer useful.
I don't know, I understand why you are being dismissive and playing down my wide eyes, but I think you are also wrong and remaining uninterested because it's too "religious" to speculate wild things in light of wild real world changes is head in sand territory.
Already AI is being used for comic book backgrounds. It's just a matter of time before all of this becomes commonplace.
When you look at AI and what it does, it is no different to what humans do. We are trained on a model (experiences and other minds), and we make derivative decisions based on the model. If you can do this in software and take advantage of light speed learning then of course all we can do will be done by AI faster and better. In time humans and AI will be the same, AI will design all the tools and tech to make this possible. It's the only natural conclusion to humanities' ultimate goals.
To say this revolution is not going to happen is to say humans have hit a hard technological limit, and I don't see any evidence to support that.
If I was less enthused I might make my opinions more philosophical than religious, but I feel overwhelmed by the possibilities of real world changes. This is no longer a philosophical thought experiment, it's happening. We are careering toward surpassing a Turing test for goodness sake. Uncanny valley apex of animation; go look at what cutting edge AI can do in terms of producing lifelike animated avatars, it's so close you have to double take.
https://www.youtube.com/watch?v=G-7jbNPQ0TQ
CGI artists have been trying to get to this level of realism for as long as the industry has existed.
Unlike a religious pamphlet, this god is tangible, it's here, and dismissing it because it sounds too spectacular is putting your head in the sand. AI is so out of this world it is a religious moment for humanity.
Civilization has seen people like you sitting comfortably and scoffing at the very idea of an aeroplane being remotely viable, and yet within 50 years of the first powered flight we had international airports.
Singulatarians really are funny until it becomes tragic.
That doesn't mean that it will make artists obsolete. It will give them more time to e.g. actually think about what kind of background would fit there best. It's a tool, not a replacement.
The key would be knowing the context of a situation. AI took over chess first, because chess always has limited context. Logo design on the other hand, needs understanding of the product, the target market, the feeling of the brand, and so on. So it'll probably be a mix between photography and management.
I fail to understand how the AI is any more vulnerable to creativity in a vacuum than a fellow human artist.
> Artists are people that sample the probability distribution of human experience
Seems that you are agreeing that human artists need to tap into human experience and the world around them, so yeah, the AI will need to be able to take inputs from the external world too.
I see no reason for an AI not to be continually training on inputs from the outside world. How difficult can it be to hook an AI model up to inputs from the internet, or even putting cameras on drones or robots and letting it explore and get "inspired". I think it's myopic not to see how an AI can learn and evolve using the exact same mechanisms as humans. I mean we are building AI in our own likeness, it will operate using analogous mechanisms. There is also no reason why AIs won't talk to each other and be inspired by other AIs rather than humans.
What will the art of an AIs living together without human input look like? When are humans basically surpassed by AI and no longer have any relevant input? Just like Alpha Go humans will see stuff no one has thought of, stuff so wildly creative that human art will look naïve in comparison. That move Alpha Go gave to the world is waiting to happen in all forms of human endeavours.
When you say something like "if anything it may increase the demand for artists" all I can think of is the dozens of times throughout history that man has seen a revolution on the horizon and thought that the status quo will still be effective. We've always been wrong. Who would have thought selling books online would replace book stores, let alone become one of the world's most successful commerce platform period. Who would have thought that broadcast/cable TV could be replaced by people making their own shows at home and distributing them via personal computers building audience numbers that surpass network TV?
Whatever happens, however this plays out, we are in for a huge shock.
Funnily enough, reading this made me less worried for artists. It seems now there are more photographers than ever, possibly because more people care about good photography than previously (despite the fact that modern amateur photography is probably on par with yesterdays professional). Maybe art will go the same way, something everyone can do, but with more respect for professionals. I imagine it'd be the same for those other fields as well.
Or AI will take all jobs and we'll end up in a Manna situation, which would work even better for me
I get theres a feeling that anything but a crushing reality of grind is living like a "pampered pet", but the second half of the book is really saying that a humans skill is in our ability to create, not our ability to work. We outsourced that to primitive machines before we even had a language to speak. We create, the AI works, replace AI with tractor or computer and the concept is the same but doesn't sound so bad, because we accepted it as alright many years ago.
As for the humans that don't use their imagination, maybe they never want to talk to an AI artist, just as many humans don't care about art at all. Millions of humans don't care about social news, and yet FaceBook algos pump out content for people all day long.
[1] https://raw.githubusercontent.com/cube2222/octosql/main/imag...
All of this was just me finding a practical purpose to go for while having fun with Dalle. If I was really serious about a logo, I would definitely go and pay an artist. Both for monetary, as well as esthetic, reasons.
Though as far as an app icon goes, I think it's actually sharp enough. It starts looking bad when you zoom in a bit.
In its current state it's not a viable logo because, for one thing, it won't look good in black & white.
That sounds like a concern that stopped being relevant for many software companies a decade ago at least.
These days app icons and hero images are more important than whether you can fax or print the logo.
Ignoring this issue is the mark of an amateur.
I've run several different types of businesses and even those that required print work never required or even benefited from black and white, or even monochrome as another commenter mentioned. We _always_ had the means and preference for full color: emails, brochures, documents, websites, t-shirts—it didn't matter. There was _never_ a time we needed to degrade the logo so significantly. From talking with others that appears to be extremely common in modern businesses, especially software, since the majority of our presence and revenue stream is online, and not glass silhouettes in our office.
As I said, outside of a fairly narrow range of real world use cases, this comment is outdated: "Ignoring this issue is the mark of an amateur." If you have one of those rare use cases, check that box, but otherwise it shouldn't be the norm or a requirement.
Maybe in the US but not worldwide.
Seems to have been blurred after the fact. The version linked in the article before cropping looked fairly sharp: https://jacobmartins.com/images/dalle2/DALL%C2%B7E%202022-08...
Plus even that uncropped one is already jpeg'd, whereas DALL-E 2 downloads are pngs, so there should be an even sharper version.
They need a black and white variation, different sizes, and the underlying component assets.
So Dalle2 might actually be able to provide that in the future as well.
But for now - it's going go give you an 'image' which you have to get an artist to then clean up int a proper logo with assets.
I'm playing with DallE-mini on hugging face and am generally unimpressed, I'm not sure if its' the same Dalle.
I tried the main DallE website sadly don't have an 'invite'.
The software used was Topaz Labs Sharpen AI. How they define "AI" I can't say for certain, but they're apparently using models so I'm assuming there's some kind of machine learning involved. Their software does a really good job on photos and videos well beyond what a standard sharpen filter does. The upscaling features are also pretty awesome. (no I don't work for them)
It’s a good first draft and something to give to a designer, but can’t stand by it’s own as a serious app logo
Generative ML is going to destroy the internet one day.
is someone generating paraphrased clones of articles appearing on HN?
why? for ad revenue?
Do people trying to read GPT3 generated English translated into their own language have more difficulty detecting generated trash?
"Each person has heard in regards to the most up-to-date frigid ingredient™, which is DALL·E 2"
I'm not too worried
That's my main gripe with DALL·E as well. This missing feature makes it impossible to use for stories where the same character goes through an adventure and is present in different settings, doing different things.
Although I don't know much about how DALL·E works, I have the feeling it shouldn't be too hard to add this possibility. That would make it so much better / more useful.
No offense, but this gives me flashbacks to bad clients and non-technical managers :D
I have to sit here and watch everyone else play with the fun "open" ai tools... company needs a name change if they're going to keep this up.
On the upside (for MidJourney), you're seeing a HUGE stream (they are hitting the 1 mil Discord members ceiling) of generated pictures and that kinda grows your appetite and you want to also try more and more prompts..
It's also an interesting way of balancing what I assume are high operational costs on the server-end by pawning off some of the hosting of assets onto Discord.
I know at least one artist and one relatively popular youtuber (with over million subs) who applied to a waiting list much earlier than me and are still waiting.
If you generate an image with Dall-E and there's a face that is distorted, you can use this tool to restore the facial features.
I've always been fascinated by how artists abstract the core notion of an image. It's stunning to see a computer do that.
And indeed, seeing what Dalle will draw when telling it to visualize stuff like "data streams" was very interesting.
So you end up using language that's sort of reminiscent of that, creating an emotional picture. It usually takes multiple passes to transfer the whole idea from your head to theirs.
I'm told that animation directors end up doing exactly the same thing. A digital model really can do what human actors can't. You could say "make that eyebrow curve 10% more" to an an animator. But it won't work unless you tell them why and what it means.
This will make it pretty hard for freelance/solo entrepreneur designers.
In retrospect it makes sense, since the visual domain has been the one with the most focus in AI.
If this gets applied to the other top domain, speech recognition and generation, then I could foresee this doing the same to the call centers, eventually also phone reception in a very small and relaxed business.
Otherwise I love this article. We spent an hour at work going back and forth with different generated logos.
Could you please link to the specific ones you liked most? That would be very valuable to me.
https://jacobmartins.com/images/dalle2/DALL%C2%B7E%202022-08...
https://jacobmartins.com/images/dalle2/DALL%C2%B7E%202022-08...
The selected one seemed a little too detailed and in need of editing.
Of course, DALL-E 2 is not the end of of text-to-image research - it'll be interesting to see where we are a year from now.
As you say, an expert can do far better.
But having something artistic created that well exceeds the average ability is gobsmakingly astonishing. And for quick blast variety generation, it is world class.
-percentage of the entire drawing that the image you want to draw should take; a lot of times I think the object I want is too "zoomed in" or large; a circle background is a good way to limit it but I think it should be more obvious
-No way to fix the color of the background so that it can fade in easily to other images or design
-Reuse drawing styles to generate further image to explore further and maintain consistency
A syntax could be: Octopus juggling blue database cylinders, digital art, cute, image-size:40%, background-color:#304324. With image-size, and background-color being keywords in the definition
It would be different if all the training data was art that was explicitly licensed for this.
Would watching a lot of animated movies in order to learn how to create good animated movies yourself be unethical as well?
Human brains also use anything the human can see, feel, hear for training. And what you produce in terms of creative outcome is a result of your experiences. But you don't owe anyone anything for training your human brain -- even if you use your brain to sell paintings, music etc.
And I stand by my point, it should be artists who decide whether their work should be used as training data for networks that get commercialized. If your work is used as training data, it is essentially an integral part of a product that is being sold without consent. Does this sound ethical?
And concerning creating a logo with such tools: Is there any consensus on an eventual copyright of such works?
https://www.smithsonianmag.com/smart-news/us-copyright-offic...
that's not all there is to this though obviously
I can see an immediate use-case for an AI layer in apps like photoshop, figma, sketchapp, gimp, unreal engine, etc that works in the background to periodically fill-in based on the current canvas.
You could prompt for inspiration, then start cutting, erasing, moving things around, blending manually, hand-drawing some elements, then re-rolling the AI, rinse-repeat.
I'm sure someone's working on it already but it seems there's a lot of scope for integration into current workflows.
On the bright side the result may be better, it may be easier to become and "AI usage specialist" than specializing in many different areas, the result may include many intermediate results that a specialist would find too much work to do and, with a bit a patience (like in the presented case), the task can still be done without the need of an "AI usage specialist".
Currently, I think the problem is an UI one. There should be an option to allow the user to do something like: "from the last drawing, just add this..." or "in the last drawing, change the color/size/style of this and that...". This would be probably enough to achieve what the author wanted in a much smaller number of iterations.
There is also on more thing: the costumer doesn't know exactly what he/she wants from the beginning. So, it is normal to have a few iterations until something pleasing is achieved.
Create a logo generator site, allow users to pick something very limited like industry/field from a dropdown or something, generate say 9 logos with AI generated text discriptions that fit this selection and remember which one the user picked and use that data to build a network that generates good text descriptions to feed into DALL-E 2 based on a singe item selected by the user.
This matches my view of the idea that AI will replace programmers. My value isn't in the typing and the syntax, it's in my ability to turn a spec into an internally consistent design by resolving conflicting instructions and clarifying edge cases; and sometimes in knowing what the user wants when they are unable to express it themselves.
Even if AI winds up writing all of the code, someone with the programmer mindset still needs to define the problem in a concrete manner. They'll always have a job as a "machine-talker."
We produce a lot of content and the biggest hurdle in graphic creation is the back and forth with the designer, plus the lag between writing, designing, and publishing. This would make it easy enough that the writer can include a prompt for the illustration right in the text itself.
More than the costs, I’m excited about the efficiency gains and smoother workflows.
[0]: http://dallery.gallery/wp-content/uploads/2022/07/The-DALL%C...
Maybe your perception of "logo" needs more reference points. For example, this gallery of classics in Brand Identity will be a good starting point(use the triangles on top to navigate): https://www.joefino.com/logos_html/L01_Xpand.html
There is no doubt in my mind that the next iterations of neural networks will remove all "overpaid" and "overconfident" design professionals, that's why I adapted to the reality and moved to frontend development. All of this with clear realization that everything humans can do for a production processes will be augmented and removed. The nasty "humans" always want to be paid, more and more. They want to have rights and privileges. What a hassle.:)
If someone else just needs something simple and passable there is Dalle.
And I’m sure there is every option in between where someone can use Dalle as a starting point and pass it to a pro, or a pro would even use Dalle as a way to brainstorm options.
Dalle is a tool that has empowered everyone. It shouldn’t be seen from a stereotypical luddite perspective as in your first post.
We'll hardly be able to tell the difference, if at all. Maybe it doesn't matter as long as the conversation is engaging for the human.
And what when people are certain that the machines are better in everything, who will want to chat, listen to music or watch paintings from the "lame" humans, when the robots will be the ultimate solution for every human need?
I imagine it would make it very easy to “seed” a website or a platform with initial “users” and content.
I also imagine it will be (and likely is already) being deployed to create the impression of popular support (or lack thereof) of a politician, business or policy.
I'm unsure if it's confirmation bias, but I find myself noticing weird abberations in online comments that don't seem to be ESL related. (edit: it's probably just mobile swipe typing at play)
I feel like I understand where you're coming from, but often the phrase I hear by experts (I even use this myself in my space) is, "Sure, it only took 20 minutes to do this wiring/write this code/draw this logo, but it took 5 years to know what to make." Sure, the results aren't what you'd get if you paid a professional logo designer, but if you can get close enough, it's really cutting out the X years training necessary to get to that point.
This is exactly my point. With repetition and solid design foundation comes the intuition what is the right direction towards the accomplishing of the given task.
Some will say the design is a subjective, I would argue that designers' role is to move towards objectivity and away from the idea of "personal taste".
That's why I give a link to the works of the master in this craft. This is exactly the same argument with the Copilot case. Is it capable to give some "boilerplate" solution - yes. Is this solution mediocre at best - yes.
The thing I think I like about this is I can meander through a few different concept on my own time.
Dall-e seems to have this concept embedded really well.
simple logo of software engineer octopus using green yellow blue and orange databases
Time to buy more credits.
The fact that you can integrate on it seems to make it much more useful.
For example see here[0], where I've combined a picture of a flying whale, a tardigrade in space, and a bunch of flying turtles.
But overall, you have to think about the context Dalle has seen similar images in the training set. If it's seen them on an art sharing site, then it's probably good to mention such sites and tags it could hypothetically have there. Or if it's more like a photo in an article, think about what could be written about it in the article.
That's my intuition about it at least.
[0]: http://dallery.gallery/wp-content/uploads/2022/07/The-DALL%C...
latent diffusion: https://replicate.com/laion-ai/erlich
vqgan + clip: https://replicate.com/ml6/julius
https://www.ml6.eu/knowhow/can-ai-generate-truly-original-lo...
How do you have the multiple figures arranged in a div in markdown -- Is that using tables?
I also didn't want to be tied to a CLI so I write all my markdown files in PCloud (because Google Drive is an ass) and have a webhook button on my phone that grabs them all and deploys.
Also been very happy with https://typora.io/ which has pasted image settings to move them to a folder in the file's directory.
My blog source code is hosted on GitHub[0] and deployed to GitHub Pages. Everything is done automatically by GitHub Actions.
You can see the image arrangement code in the source of the article - it's just rawhtml with inline css. A very ugly approach, but it works.
For the images, I just changed the download directory of my browser for the time I was writing the article so that it put the images into the right folder automatically.
Good luck with the article!
[0]: https://github.com/cube2222/cube2222.github.io
[1]: https://github.com/cube2222/cube2222.github.io/blob/main/con...
You can also ask for "black and white vector art" to limit the color palette.
May be better idea to learn how to prompt the future AIs.
Before we know it, we will have an AI making an entire movie (about 200-400k frames)
This does make me wonder if it would be feasible in the future to run these kinds of solutions on your PC, even with pretrained models. Or will these AI solutions generally trend towards being hosted in the "cloud" as consumer PC will never catch up in required resources for them?
You search the way you think you should at first, and don't get what you want.
But that search informs you of the "terms of the domain" in which you're searching.
So you then refine your search to include those terms, and iterate until you find what you were looking for in the first place, but weren't an expert in (SME)
Also, perhaps it can be smart enough to ask questions, like "What database should this be written for?" in the above example.
well, that's pragmatic! I think they should go back into their image editor and simplify it themselves though
I guess it’s useful for family friendly Disney-Esq corporate media, but for what I would consider real impactful art, it is lacking a great deal.
Don’t worry artist, your jobs are safe.
IANAL but my understanding is the ability to copyright the output from something like DALL-E 2 is questionable at best, due the lack of human authorship.
(See "Monkey selfie copyright dispute" on Wikipedia for more info.)
I think someone taking the time to touch this up would make it copyrightable and trademarkable, and I'm okay with that.
AI generated patents though? Only if we allow AI generated prior art.
This may or may not be the case but I get the feeling that most ITs haven't heard of diminishing returns.
It's like saying "steal from artists, but only the good ones"
Not sure about "iteration". I mostly did a lot of experimenting.
If you have something like "let's add a helmet to this", then that's basically 5 minutes to a good result.
The whole process took a few hours if I remember correctly, with the main hurdle being to come up with the phrase for sensibly laid-out pictures (the circle background). It went quite quick from then on.
A couple of years ago there was a list of jobs that were going to be in danger from being automated in 10 years, I don't recall if designer would be on the list but it looks as though that moment has gotten significantly closer.
Assistive technologies like this are awesome!
It would be extremely cool to see the same process spent on generating a logo.
Open-ai has nonsensical censorship. Dalle might be popular right now, but open-ai won't survive if they keep up their ridiculous attempts at trying to control culture. I've already got something running on collab, new models are coming out, and midjourney just got a v3 update that blows dalle2 out of the water.
I'll have to find a different option.
Just because some asshole uses the peace symbol, does it mean the peace symbol is hateful? Anyone who claims so is dishonest
However as it gets better, even that won't be needed and I'm concerned then what that means for the average person, or for me trying to get my skills up so i can increase income, only for it to get wiped out by AI at some point in the future.
I think it makes much more sense for simple illustrations for articles, presentations and books ("pencil sketch" style). For logos, especially since you'd usually want simpler shapes, less detail, with a lot of readability, I'd go pay an artist if it was for a company I was building.