Apple and Google avoid naming ChatGPT as their 'app of the year'
techcrunch.com
techcrunch.com
ChatGPT is a boring app. It doesn't have a ton of platform integrations, it doesn't show off the capabilities of the platforms, it doesn't push any boundaries on what an app can be. The product might be great, but that's not the point.
This is my personal opinion based on years of following Apple's awards, I don't have any inside info in my role about this and this is not my employer's opinion.
I think it's understandable that these platforms aren't giving ChatGPT any awards. It is not helping to make the platform stand out against other platforms.
I agree, but I think the crux of my point is that this is because ChatGPT the app doesn't stand out from a platform perspective, rather than because advertising ChatGPT would go against the platform goals.
Essentially I have no real reason to question the integrity of the awards processes here, because they seem to be awarded for the same reasons as they always are, and that ChatGPT just doesn't tick the boxes they're looking for.
I'm not naive enough to think some exec hasn't been happy with the convenience of this, but it seems a little conspiracy-theory like to assume malice here.
The revolutionary part is that you're chatting with a language model running on a server rather than a human. But that doesn't make the app impressive, it makes the back-end impressive. App of the year isn't about celebrating impressive back-ends.
Messages and WhatsApp are better designed and the communication experience is better.
But Apple - and Google I imagine - don't let 3rd party apps replace siri / "hey google". Apple and Google docking points because chatgpt doesn't have a lot of platform integrations feels like a very petty move.
edit: With that said, if Chat GPT came before iOS, I'm pretty sure Apple would have had to design iOS with that use case in mind and get these APIs exposed somehow, maybe even through a direct integration with OpenAI or Anthropic.
Now, that would definitely still take some work from Apple. Right now as I understand it the device listens for “Siri…” all the time using some on-device ML. They’d need a way for that to work with chatgpt and hand over to the audio stream. And have a chatgpt conversation from a lock screen without breaking sandboxing.
But all of that would definitely make my phone a much more useful device. I won’t hold my breath though.
I think it's reasonable that they don't let you replace the logic behind their branding – it would be quite a serious issue if ChatGPT started giving offensive results when you ask "Hey Siri".
Whether they should allow replacing the voice assistant entirely is an interesting question. On Android OEMs can of course do this, and have done so with mixed success, see "Bixby".
The issue with LLMs is that they are slow. People expect a response immediately, so much so that Amazon Alexa built much of its popularity as a kitchen timer over the alternative because it's just fast.
I think it's a much more reasonable argument that the reason Alexa, Siri, Google Assistant, Bixby, and Cortana (not sure if that exists anymore?) don't use LLMs is not that Amazon/Apple/Google/Samsung/Microsoft are asleep at the wheel, but that LLMs are not suitable yet. I'm sure this is coming, but I'm not surprised none of them have it yet.
Fake take.
No one believes this. Or at least no one who is both honest and intelligent.
This is mere mental gymnastics to explain this away. I'd take a hard look at what is going on in your brain to cause these kind of apologetics for Google and Apple. Its kind of scary to see from the outside.
And your reply is incredibly hostile and aggressive. Before you suggest to others they should "take a hard look at what is going on in your brain" — you may want to take your own advice.
I welcome good faith critique and discussion, but feel that you may be engaging in bad faith arguments and taking my points out of context or purposefully misrepresenting them. Perhaps not, perhaps I could have clarified my position, but I hope you understand the point I'm trying to make.
Not sure why this is news, ChatGPT is an interesting service but their app is barely a thin layer over their API. So why would it deserve the mention?
I mean, do we need to give say Google, "App of the year" every year because they're the most popular search engine?
The ChatGPT app itself was a very thin veneer over an online service. It utilizes almost no facilities of the smartphone client, has no local AI or unique intelligence of the device. Indeed, OpenAI left the field open and before they came out with their own there were a myriad of competing facades that were effectively interchangeable.
I don't see why that would earn app of the year. I wouldn't even put it in the general conversation.
I remember seeing here that his default system prompt is also optimized for oral conversations.
Voice2voice is great, but I find the way chatgpt constructs responses in oral conversations to be very soulless. It feels like I'm talking to the personification of a bland corporation. I'm not sure if it feels worse through voice because it sounds close to human, or if its the system prompt. I'd love to find a way to fix it.
Does anyone have any suggestions of a good system prompt to add to make the AI loosen up a bit?
Even the neutral POV policy of wikipedia is constantly worked around by various actors by marking "bad" sources as biased or unverified, while "good" sources canonical and neutral.
A little bias is inevitable. But there's no reason to cynically throw the baby out with the bathwater. We can still strive for fairness and balance. We should still complain when we see blatantly biased source, and seek out good journalism. We'll just never fully achieve it.
But thats ok. We can try anyway.
I still struggle to find any use to it in my daily life. It is a cool demo, but no one wants to read AI generated text from other people.
But for quickly creating some template when i want to write a big report or email for something, then yes it's very useful.
ChatGPT is about as accurate as random websites on the internet, and you don’t get obliterated with ads.
Simple example, ChatGPT will give you a clear recipe for whatever you want, sans life story designed to make you scroll past a million ads.
In general, the context of search gives some insight into the credibility of the source.
Only way to know if a recipe is good is to look at it.
The only reason those blobs of text exist is to get you to look at more ads. Put more things in your head against your will, sell you more garbage, and manipulate your feelings.
If it wasn’t true, why are the recipes always at the bottom? Why not put the most valuable part right front and center? These websites have no respect for you and likely copy pasted the recipe anyways.
I could specify for it to use MDN exclusively but at that point I might as well use search.
In addition to that I could judge the quality of search results (a lot vs little mentions of a technology, shady vs reputable site etc.) to make educated guess of the output I'm getting from search. Can't do that with GPT.
These are key differences off the top of my head.
You don’t. But i’dtrust a top rated Stackoverflow answer over whatever LLM spits out.
There is no “confidence score” from an LLM output. You cannot tell whether it is making things up (and potentially make very bad decisions based on it’s output)
https://community.openai.com/t/new-assistants-api-a-potentia...
Try something more complicated! Ask for a gingerbread recipe without sugar, for example.
I think I'll ask it for a calzone recipe this weekend. The one I use now makes the dough a little too bready.
It's not very good at it, but it doesn't need to be, to be far better than I am.
People have some sense that someone giving them information may be an {expert, charlatan, idiot}, or that a website they’re looking at is run by a university vs a blogspam content farm, but many have not developed a sense for when or how much they can trust LLM output, which is delivered with the same tone and confidence regardless of whether it’s entirely fabricated.
There is probably a component of personality involved in how people approach this. Collectively we are all learning how to interact with this new source of information and people take varying paths.
I've had ChatGPT return very serviceable "true" results and I've had ChatGPT return utter fiction.
1. you don't know the answer, but you can check yourself and easily whether a given answer is roughly correct
2. you don't know the answer and wouldn't be able to check how valid a potential answer is
LLM-based tools are great for 1 to synthesize various sources into one coherent answer, since in this case, you won't become a victim of their hallucination. E.g. "write a one-off Python script to do this": you can quickly check if it does the job, even though you couldn't say whether that's idiomatic Python.
I would say it is not good at giving a sophisticated answer to anything that requires a lot of nuance. And I've also asked it questions with fairly objective factual answers that it gets hilariously wrong.
I still believe anyone using these tools on a day-to-day basis should have a sense of "trust but verify."
Example: Using a small part of a new, big, unfamiliar library. Rather than digging through the library docs, I can ask ChatGPT about it, which often points me to the relevant parts, which I then can still confirm in the official docs.
But the code it gave me was a great starting point. I found it much faster & easier to rewrite the bad code it wrote than pore through documentation and figure out how to solve my problem from scratch.
- if I'm working in a field that's new to me or that I don't understand, I ask for help understanding the basics and vocabulary of the field. it does very well at this.
- if I have a well defined problem, but am simply not familiar with the libraries for a given situation, it tends to do a good job translating my english queries into the right code. I do examine and test the code afterwards to make sure it's correct. You can really see this in action when you ask it to do data analysis; and the REPL loop in that mode is also great at catching bugs.
- if I have like a copy-paste from a documentation site, I can ask it to transform that into code or into a better-formatted version. this saves a lot of time, and I don't have to remember regexes or vim keybinds
I also do this, but I am careful about being confident that "it does very well at this". We can't actually evaluate what it's putting out, other than that it sounds plausible, which is something LLMs are truly great at.
Like random formal letter to AI? Maybe it is a societal problem more than a technical one if we all hate writing and reading these.
I don’t get the search documentation part. It has obvious blind spots on many things and hallucinate on others.
No, I'm talking about things like writing SQL queries (even complex ones), CMake files, Docker configs or plotting stuff in Python. Of course, if you're not already an expert in these things, you'll have a hard time distinguishing useful replies from hallucinations - that's why I said it's mostly for senior devs. Without expert knowledge, you will likely not be able to benefit from it in its current state. But if you have that and know how to write efficient queries, it can easily up your productivity by a factor of 10 (i.e. going back and forth for 6 minutes with GPT4 to make it get your requirements can save you an hour of work looking through documentations).
Sounds like what a junior developer would say, given they tend to depend on it even when it hallucinates the wrong answers badly.
Also explains the rampant title inflation that is going around in the tech industry these days.
Sometimes it comes up with really good techniques that are different from my usual approaches, other times is plain wrong and I'm correcting it.
I am more productive with the ChatGPT in my life. Whenever stuck on some weird error, instead of googling I paste the log and discuss with the bot what can be going wrong there. We can talk about pointer analysis, performance bottleneck comparisons. Could I do it all on my own? Sure. However it is boring and quite certainly would require 10x more time to perform all calculations on my own.
In the end of a day it is just another tool. Brings advantage when used properly.
But I use the web interface, not the app.
That said, sharing the output with others is not necessary to get value out of it.
For example: "Help me work through an [idea/plan/problem] by asking the next Socratic-method-style question."
I tend to think those awards are probably not "what is the most popular" but "what is the best showcase for our platform".
If you're judging purely on the app, it's not surprising. The app itself is about as basic as they come.
On the Apple side, AllTrails won, and it's deserving of that win. It's made some incredible updates this past year with some topnotch design. Flighty would have been another very deserving win.
Things that give my phone mobile-centric improvements winning instead of a website wrapper