Google Duplex might make my design job redundant
thenextweb.com
thenextweb.com
This was one of the original points of HTML and the original design of the web. The Semantic Web meant that 3rd party and automated assistants would be able to control it. Burners-Lee described it like so:
> I have a dream for the Web [in which computers] become capable of analyzing all the data on the Web – the content, links, and transactions between people and computers. A "Semantic Web", which makes this possible, has yet to emerge, but when it does, the day-to-day mechanisms of trade, bureaucracy and our daily lives will be handled by machines talking to machines. The "intelligent agents" people have touted for ages will finally materialize.
Everything goes full circle, and there is nothing new under the sun.
Where voice assistants are concerned, this is a largely unexplored UX space. Voice assistants in the future are not going to be the same as they are today, because our design trends are going to evolve as we learn more.
If those trends move towards predictable interactions over interpreted ones, and if we decide that there are times we want to use assistants without physically talking to them, will we see a resurgence of text-based assistants? Will we eventually go full circle all the way back to command lines, just under a different name?
Automation and people trying to have everything 'cloud-based' and 'secure' will probably rely on services not just like this, but like time-saving and squared-away solutions like this.
Businesses had no idea of this back when HTML was invented, and now that we understand how that works and that has become such a task that we would benefit greatly from implementing something to get rid of it, who knows!
I think it's worth a shot to automate something like this, which is already tech focused and surrounded by talent, that automating other things will become closer in scope.
That's what it's all about really, time-saving and automation.
Secondly, I have a hard time thinking design is anywhere close to obsolete. Google Duplex sounds like an adapter that will make reservations for you upon voice command. This will not obsolesce anything except a few forms on most websites. Design encompasses _much_ more than how a form looks or functions.
Unless you can explicitly enumerate the "small" subset of tasks the cli is simply better at, I don't think that's a valid out. The CLI (specifically, the unix philosophy) is IMO the best platform for general purpose computing. It's the specialized stuff, mostly applications with such a large feature/configuration surface that it would be foolish to try to learn all the commands and arguments, that is best left outside the cli.
CLI works incredibly well for anything related to controling systems (local, web servers), or running any kind of processes at scale (web scraping, photo metadata). Discovery is the biggest problem here, though I do believe it's solvable. If I want to do any one-off task, such as resizing an image, that's incredibly easy to locate and do within the typical operating system GUI, whereas to do it in CLI I'd have to resort to man-pages or search engines to figure out what command to execute.
For everything else CLIs are more consistent, flexible and extendible.
But why not use both? GUIs are amazing for editing a video, command line is amazing. If you want to convert a thousand videos into different formats based on their meta data the CLI will be the only option that doesn’t makes you sit there for a week.
My kids learned how to use touchscreens as babies. They learned how to manipulate GUIs around the time they entered school. But they still regard my terminal window as black magic, and I'm not sure how to even start explaining what's happening when I punch in commands.
That said, I agree the article is overstating the impact this will have: computers talking to computers has been a thing for a long time, they just use APIs (which are carefully crafted to be as umambiguous as possible) instead attempting to parse ambiguous human speech and pipe it into arbitrary web pages.
Though perhaps not trivial, it seems easier than, say, explaining what you're doing when you fix/manipulate things on a car engine.
The terminal uses a form of language to give commands, and the result is the execution of the command and/or some sort of printed output.
`ls -lah` could be explained as 'a fast way of inputting' the equivalent of "Alexa, tell me in detail what files are in this directory". (Yes, you would have to explain something basic about files and directories, but that still seems reasonable.)
You can try to re-experience that frustration for yourself by playing an interactive fiction text adventure game. For example in http://adamcadre.ac/if/905.html, the first few things I tried were: "left", "go left", "move left", "map", "where am i", "help", "?", "tell me what I can type", "fuck you", "exit" (you've now exited the room and are in the living room)
Meanwhile I can play hide and seek with the tools I need in Gimp for hours.
When you see your kids do something with a GUI ask them to write down a list of instructions so that you can do it too. That list of instructions is basically what using a CLI is, it's just less discoverable.
(Seems i can't link the actual image but in the article there's a chart that breaks down the speed/information density of different types of communication.)
Maybe someday AI will be a legitimate user interface, but I don't think we know what that looks like yet.
Are we going to see an industry shift away from the belief that a physical human needs to be at a website in the future? Is there going to be some kind of back-door where Duplex won't need to solve Google Captchas?
It's hard for me to look at Duplex without thinking that it's something of an admission that bots and automated assistants are a legitimate way to interact with the web, and that we should be trying to block behaviors, not agents.
I believe this is the entire point, and how they end up capturing an entire market that nobody else can compete with.
Recaptcha requires you to prove you're a human. Unless you're a google bot, then you're fine.
If designers then stop creating these forms, then what will the AI use? Seems like some kind of API would be needed, but typical REST API is not detailed enough to support this.
So the API would somehow need to be created (potential job for UX person) or inferred from something.
These things have been around for years. There was a MIT Media Lab demo long ago. It's basically Amazon Echo for services, right? Google might make this work by insisting that services offer an API that they can call to get business from Google users. Of course, Google will want a cut of the revenue.
They'll probably get it going for food delivery and car services, then bail on anything that isn't basically online ordering. Airline reservations, maybe. Doctor appointments, probably not. Appointments with important individuals, unlikely.
To use the example in the article, I can imagine making the request to Siri on my watch, having Siri read back a reservation, followed by “is this OK?”
Too bad I hate wearing watches.
No thanks