Chatbots-Are-AI-Antipatterns
hello-jp.net
hello-jp.net
The chat/voice interface is like Skinner's Box: we keep "using" because we don't know what will work and what won't, and when something works it seems magical. And it keeps being pushed on customers because it's beneficial for corporations: customers will fruitlessly navigate chat support rather than realizing they should give up and call, businesses will invest in products with open-ended interfaces because they can convince themselves it's flexible enough for their needs.
Businesses tend to thrive when consumers choose to keep using them, not when they don't. How does users fruitlessly navigating chat support serve that end?
There is no long term in modern capitalism.
No-one's getting promoted for "declined to ruin our customer support, thus avoiding loss of X recurring revenue per year over the next decade", but they might get promoted for "reduced cost of customer support by X with AI!!!"
In the long run the invisible hand will dole out rewards and punishments here, but thinking about the short term is just _easier_ than thinking about the long term.
Only in competitive markets.
There's so many studies done on human computer interactions that at this stage, I'd believe it's just pure malice. Corporations want to push the idea that their tools will let you do your task without being trained in it, which have never happened even with physical tools. Even game console and smartphones require some guidance from other people.
> It seems we're exploring why Jakob and I think chatbots might not always be the best solution! I see you've already been introduced, so let's dive deeper. Given our discussion on chatbots, what are your thoughts on the hybrid interfaces Jakob mentions towards the end of the article?
But at some point in real work, you always need to move up to a higher bandwidth medium. If that medium is going to be text then it's got to involve rigorous specification to avoid ambiguity. If it's going to be quick and casual then it has to have other channels to convey meaning than just the words used—real humans use body language, tone of voice, and back channeling to ensure the message is getting across with high fidelity. There's no such mechanism in chat, which makes it frustrating to use with people or with robots.
So rather than the machine benefitting from that extra info, it's actually just making it worse by introducing another link in the telephone chain; hence users getting exasperated and giving up to go back to conventional HMI paradigms of keyboard/mouse/touch.
If you’re communicating purely technical information about a subject everyone is already aligned on, text is fine.
But the information density present in facial expressions, vocal inflection, body language, etc. is significant and much is lost without it.
You are right, though. If all your meetings are about persuasion, video and audio will be best.
If someone is in a technical role and most of their meetings are like this, that's a bad sign.
This, admittedly, means it's very often not excellent.
I discovered that eating nothing but granulated sugar for 48 hours made me quite sick.
Watching companies force chat interfaces into everything is like watching a bunch of kids discover that you can’t live on a diet of sugar. They won’t listen until it makes them good and sick.
Went outside, found the ant hill right next to the building window and dumped the bag of sugar on the hill. Half of it was gone in a day! The next day a little more, but there were NO more ants in the house. The next day there was still sugar left and the day after that. The ants started stopped coming out of the hill much, and I think they had filled up their nest with so much sugar they weren't getting any other nutrients. After a week there was still sugar, but they didn't come back into the kitchen for the rest of the term, when I moved out.
I agree that chat isn't the end interface for AI powered applications.
That said the ability to describe desires in "fuzzy" ways is a usability value unlock.
My take is that natural language should be use surface functionality, translate between structures, and generate personalize defaults.
Its a hybrid approach and why my co-founder and I started https://tambo.co
LLMs handle the query understanding part, but it is still up to the backend developer to implement the functionality to fulfill the various possible requests. And the opaque nature of the interface gives the user no clue as to what requests are supported. We all experienced this problem with smart speakers, with their ever-changing features, and resigned to using an increasingly limited set of features.
Ctl-Shift-P does the job where available. Not entirely sure how AI makes it any better.
What we need are designers who can help establish a foundational structure (information architecture) that leads to discoverable and simple UIs to nudge users in the right direction. Once users are at a place where they know what’s possible/available, then perhaps you can allow some fuzziness to help them cross the line to accomplish their task.
[1] https://www.interaction-design.org/literature/topics/discove...
Don’t trap me in a chat window https://austinhenley.com/blog/trappedchat.html
Natural language is the lazy interface https://austinhenley.com/blog/naturallanguageui.html
These friends probably never had to deal with Google support.
Like good grief I know we're insanely politically polarized at the moment but I think every single person in America except for like 200 executives would get behind a law that says "cancelling a subscription must be exactly as easily done as getting one."
There is no reason at all to require a market to give us permission to make a better society.
Better societies but without good market couldn't compete technologically and militarily, so ultimately they lost. Borrowing a line from Rick and Morty: that was always allowed.
I trust Amazon to make society a better place than lead-brained boomers/millennials who can't figure out how to cancel a subscription
That's not how the free market works. The free market works companies prioritizing profits over people. So what's "important" for people is only visible in the market when it's also profitable.
Cancellation policies fall into two well-known failure modes of markets to actually indicate what they want: information asymmetry and inconsistencies in time horizons between counterparties.
You have no ability whatsoever to assign market preference to this without at the very least ensuring consumers are actually aware at time of purchase of the friction they'd face at time of cancellation AND the chances they'll want to cancel.
That's all leaving aside that a person's preference for Product A over Product B obviously does not mean that a person prefers every dimension of Product A over Product B.
2. Competing on being easy to cancel is a bad thing, actually. However wins that competition is losing money. That's why you see the competition go in the other direction. As in, who can make the shittiest cancellation interface (Planet Fitness, btw).
3. People who cancel a service are cancelling BECAUSE of the service. Why would they go back to a service they disliked enough to cancel, if the service stays exactly the same? They wouldn't. Meaning, "good cancellation" doesn't count for anything.
So it doesn't help you retain customers, it doesn't help you get back customers you lost, and it doesn't help you get new customers.
Oof sounds like the base is pretty divided in this one. Better send it to committee where it can be studied forever.
Would you like to sign up for our new personalized support for only $4/mo? The first month is free with credit card - I mean the Nigerian scam works because of self-selection, amiright?
We gotta compare apples to apples.
If these companies had good design, good processes and clearly defined processes, they wouldn't need a chatbot. They could add one to make querying better/easier.
Let's say you go to your ISPs website, because your internet connection is down. Ideally their frontpage would have a direct link to "operational status", but let's be real, it doesn't. You eventually find the status page, it doesn't mention anything about outages. So you call or ask the chatbot, in neither case will you get an answer, because neither the person on the phone, nor the bot has ANY clue that you're internet connection is down. You then switch to a different ISP, who has procedures and systems in place and a link to their status page and they update it before you even think about contacting them. No amount of AI chatbots will fix the problems for the first ISP, because their internal systems suck and can not provide the bot with the required information anyway.
Most interactions with ChatBots and customer service could, and should, be solved with better processes and better systems. We're not even at the point where we can compare the well designed UI to the well designed chatbot, because most businesses can't provide the data and processes for either.
GUI is great when I have to do something and I need guidance to do it correctly because in GUI I can see what is expected from me in a glance, I can see required parts and parts I can gloss over.
Let’s not forget about CLI which is perfect when I exactly know what I want to do and I do know how to do it in least amount of motions.
People who think chatbots are great think that natural language can be efficient like CLI but it is not.
I admit I lean more on the "command line pilled" side of things, but to me this analogy implies that the chat interface, specifically prompting, are fundamentally intractable as an interface to AI.
Every operating system is necessarily built on compiled/interpreted text, and therefore has the ability to "fall back" to a command line interface to have finer grained control over basically everything.
Even mobile devices, completely 100% GUI, where native terminal emulators are unusable (except for the truly dedicated) are wired this way as soon as you scratch the surface of app development.
I think this analogy has to hold for AI, no? The entire current ChatGPT renaissance is built on the concept of prompting. There are no examples of breaking the mold in this article that don't rely on an underlying "chat" paradigm. Voice agents? Just turning a vocal interface (something predating GPT3) into... prompts for an LLM. Gemini summarizing documents? A button that has a prompt pre-written and ready to go.
Carrying forward the analogy, there will be a "GUI" where your AI agent will provide a streamlined experience that is extremely powerful in matching your context most of the time, but then if you want to tweak anything yourself one of the most powerful ways to do this is to get down into the "command line" of what chat interfaces control what or how these underlying prompts are structured/ordered.
The article goes on to claim that hybrid UI's are the future. I disagree, the future is figuring out how to automate prompt generation so I don't need to even interface with the UI to begin with. Hybrid UI's are a patch over something that can't be efficiently automated yet.
The opportunity is the "yet".
To take a concrete example from the article, the Cursor screenshot includes a prompt "Please refactor these components ..."
When was the last time I told an engineer on my team to do that? They automatically do that because they feel like it's a good idea. That's a no-interface magical AI.
I may still try it in an app I'm building.
Docs -> RAG -> tooltips / onboarding sounds like it might work out ok.
I have tried showing help text in an internal app using tooltips when the user would hover over the target element (or show a small icon on touch devices), and the feedback was good as the tooltips were never in the way but easily available for help (accessibility for keyboard users needed some thinking, but for the limited audience for that app, it was not a problem). And while you're at it, may be make it more engaging than a simple text only tooltip (which can be done without any intelligence), and let the host to customize and offer complex workflows.
The point is OK (ie, a car with no steering wheel just chat) but pretty obvious and seems like the author is a hypocrite in this regard