1-800-ChatGPT
help.openai.com
help.openai.com
It also brings back memories of trying random (and known) 800 numbers from payphones.
In the days before we had pocket supercomputers, I used both of these services occasionally while out and about.
They know who's calling, and the greeting was something like "Hello again". They are catching up at building a competitive database of persons and their preferences at the scale of FAANG. They're moving over from collecting info for their models to collecting info from their users for their agents. This is what they need to offer good agents.
But I might be wrong and it's just phoneme collection, as you speculate.
List a business with a Google voice number and you can call in, check messages, and _dial out_ from Google voice. Free international calls!
I was in school in Canada where we had a payphone in a hallway. People heard me randomly saying "Funny Business Name, City State ... Connect me" into the phone so much, it became a running joke.
When I eventually got my own phone, I transferred the number and I still have it.
Goodbye to an old friend: 1-800-GOOG-411 (2010)
https://googleblog.blogspot.com/2010/10/goodbye-to-old-frien...
3 questions that Gen-Zers probably have never heard asked and will never ask themselves
Alaska Air has a whitelist of messaging services that you can use for free during the flight. WhatsApp is on that list.
So if you want to research obscure plane movies on an Alaska flight, you can connect to their wifi and message either WhatsApp's built-in LLaMA or now ChatGPT.
You can spend 2 hours watching a moving, emotional story that teaches you something new about the human condition and the choices we make in our lives.
Or you can spend 2 hours that turns out to be full of plot holes and inconsistent characters, where nothing makes sense in the end and you've utterly wasted your time.
In what universe would you not want to have that information before watching? Especially if you're generally a busy person and only get to watch 10-20 movies a year.
I truly don't understand the attitude of "just pick one", whether it's for movies or other things. That reviews are "micro-optimization". Like, do you just not value your time? Do you not care about quality?
It's not like reviews are always right. But one film with 98% on Rotten Tomatoes vs. one with 45%... that's a really strong signal. Why on earth would you choose to ignore that?
Watching a bad movie is not going to harm you. Maybe you'll take something away, maybe you won't.
Much like having a bad day is unlikely to ruin your life - it'll just give some nice context to the good days.
And we're talking about watching them on the plane, so the "busy person" argument really doesn't apply here.
Where did I say it wasn't? That's a straw man.
But if you're going to watch a movie for the next two hours, then yeah -- your life is going to be about that movie. So why not choose wisely?
> Watching a bad movie is not going to harm you. Maybe you'll take something away, maybe you won't.
Straw man again. And again -- why not choose quality instead of choosing ignorance and rolling the dice?
> Much like having a bad day is unlikely to ruin your life - it'll just give some nice context to the good days.
Again, straw man. Nobody's talking about ruining your life. But why intentionally choose a bad movie...?
> And we're talking about watching them on the plane, so the "busy person" argument really doesn't apply here.
To the contrary. For a lot of busy people, the plane is one of the few moments they have time to watch a movie. So it sure does apply.
You're arguing in favor of choosing bad things, because it's not going to ruin your life. Huh? Shouldn't we have a higher bar for the things we try to choose to spend our time on? You're describing standards that are the lowest of the low -- as long as it doesn't harm you, it's fine. Don't seek anything better. Yikes. I've rarely come across a life philosophy more depressing.
I can’t take this knee jerk response seriously. Why wouldn’t good movies be more worthwhile than bad movies? How is this even controversial?
You can just pick a better movie and make life better.
What is your objection to that?
Usually tho I just watch trailers and judge the actors' chemistry to decide if I would enjoy watching those characters. What other people thought of it is not especially relevant. Particularly on flights I've watched some amazing foreign content that I just would not have stumbled upon if I was just watching whatever topped rotten tomatoes.
Reading a movie review or just asking a friend that’s seen it if they liked it has been a thing… always.
When in reality if you ask ChatGPT for 10 good movies from this year you will get this.
Anora - Directed by Sean Baker, a compelling drama about the life of a sex worker in Coney Island.
Challengers - A provocative tennis drama directed by Luca Guadagnino, starring Zendaya.
Dune: Part Two - Denis Villeneuve's continuation of the epic science fiction saga.
Furiosa: A Mad Max Saga - An action-packed prequel exploring the origins of Furiosa, directed by George Miller.
Inside Out 2 - Pixar's sequel that dives deeper into the complexities of human emotions.
Wicked - A musical fantasy adaptation directed by Jon M. Chu . The Zone of Interest - A thought-provoking film about Auschwitz, directed by Jonathan Glazer.
The Idea of You - A steamy romance starring Anne Hathaway.
Hit Man - A comedy thriller starring Glen Powell.
The Outrun - A powerful drama about a recovering alcoholic, starring Saoirse Ronan.
Let me know if you'd like more details about any of these!
Which is a great list.
Originally the problem was supposedly that it would hallucinate complete and utter gibberish, but now here we are quibbling over one example and insisting that maybe it's not quite as good as alternative descriptions.
The gap between what was produced and what you're looking for is small enough that I think it could be covered with some slightly tweaked prompt instructions.
I'm not saying you're wrong but want to note how the goalposts keep seeming to shift whenever we talk about these capabilities.
That point is that the information provided above about these movies is worthless. It does not add any new value beyond what would already be available in the streaming interface. Several of the descriptions are nothing but the genre and one person involved in the making of the movie. And yet even with these descriptions being incredibly short and vague, they still manage to contain at least one misleading summary.
I know that any mention of fallacies, valid or otherwise, causes instinctive eye rolls, but in this instance I agree with them that this amounts to moving the goalposts.
You don't even seem to be disputing the actual results here, just gesturing towards a kind of philosophy class exercise of whether we can ever "really" verify its accuracy. I see Wittgenstein's name increasingly tossed around in these parts (a good thing!), so I'll just note that one of the reasons he's hailed as one of the great philosophers of the 20th century is because he felt these puzzles about "really" knowing were frivolous.
I don't think I agree that what's needed here is some new and extra process of verification. I think the same usual quality control criteria that are already being used are good enough in this case.
In general you can't, but surely it's not that big a deal if ChatGPT offers an inaccurate summary of a movie you're about to use to kill time on a flight? I suppose it becomes important if, e.g., you're relying on it to tell you whether a movie is appropriate for children, but, if you're just asking it whether a movie is worth watching, that's a question that doesn't have an objective, factual answer anyway, so a hallucinated answer is probably about as useful as that of a not-previously-known reviewer.
My wife wanted a pair of boots for Christmas that I couldn’t find in her size. Google was a wasteland of SEO, but ChatGPT found 5 sites and was able to tell me current stock levels.
' As of my knowledge cutoff in January 2022, the last movie I have information on is "Spider-Man: No Way Home", which was released in theaters in December 2021. It was one of the most highly anticipated films of that year, marking a major event in the Marvel Cinematic Universe (MCU) and the Spider-Man franchise. '
I pasted the same initial prompts in both, but Meta AI needed more clarification. When ChatGPT found multiple entries with similar titles, it gave information about all of them.
https://gist.github.com/appsforartists/004bafe11a9e23a418fd5...
The first thing I fact-checked, the Rotten Tomatoes scores are actually 66% and 51% respectively[1]. Probably not enough of a difference to sway any opinions, but an excellent example of the type of inaccuracy that the previous comment was referencing.
Hilariously it often believes that it can’t access the web and then hallucinates reasons for how it can know things beyond its knowledge cutoff date. But in any case, it works very well for this use case.
I'll definitely give it a go, I wonder if this lands better with those aged 50+ who are more used to phone calls rather than chat.
I’m well into this group and still make a lot more api calls than phone calls.
Fresh out of college I recall vividly thinking, I’ll need to build an impressive list of side projects to overcome preconceptions about how much I can truly offer at my age. Maybe nothing has changed.
50+ are going to be so addicted to this thing its not even funny. My parents are not reaching for AI immediately yet, but thats just a yet. This is the wave that could come at any moment.
(Though I see that as mostly a failure of our financial industry. Credit card numbers should be obsolete by now.)
our industry is old enough that the first generation of pioneers has died of old age.
Do you really think someone who grew up with computers in the 80s is incapable of using a smart phone? These are people who are still in the workforce today. These are your most skilled colleagues.
Some of them probably designed the device you think they're too old to understand
I know people who are in their 20s and 30s who seem to be uncomfortable with computers, cloud technology, and especially AI.
In some ways I'm one of them. I will never let an always listening AI helper be in my home. And I'm <40.
But as you can probably tell from the other replies, the idea that older people don't know how to use internet-era technology is a meme that was wearing thin 20 years ago already.
People who haven't had ChatGPT "land" for them yet are likely just people who don't find themselves asking a lot of questions they need a chatbot to answer, regardless of the medium. That probably has some age skew right now, but isn't really about the medium at all.
When I dabble with chatgpt it always feels like I’m playing with a toy as I don’t really have a use case I’m taking to it. I’ve used a few websites creators and code generators which have been useful but also I don’t think they saved me much time overall. Web design, graphic design, etc and creative stuff are things I suck creating so it gives me a new power and is easy to iterate on. Otherwise, I’ve not found much actual value from it yet.
If it makes you much more efficient in your job, like it does for professional software devs, many of HN users, then i think you’re more apt to be excited by the tech
I was thinking earlier today that an agent listening to my calls would be helpful. I was on the phone with a financial institution that will require some followup. Being able to sync in an agent to transcribe and remind me would be valuable.
I understand this isn't that.
If anyone else reading this is in a place like you've described, try 1900hotdog.com
My biggest use case for ChatGPT voice mode is when I _need_ or _want_ to be handsfree. Think working around the house/yard, Driving, Walking around the grocery store, cooking, etc. I find that I end up using my iPhone's voice-to-text then simply communicate with text mode (in the case of driving, I stop). After all, once I have to touch my phone, it's just faster to work in text mode.
All of my devices know how to make calls. All of my devices know how to make calls from a voice command. All of my devices know how to hang up a call. This is really nice.
How ironic that it's not actually Apple delivering that despite being in the perfect position to do so (they have a deal with OpenAI for ChatGPT using Siri, have all the contextual knowledge they could ever need etc.) – my iOS 18.2 Siri + ChatGPT experience has been extremely disappointing so far: It seems to completely forget all context between questions, ignores me for follow-up questions 80% of the time etc.
Im still hoping one of Open AI's 12 day announcements is they are creating a AI Phone with Microsoft called GPT and or an Phone AI OS.
“Hey Siri, call 1-800-CHATGPT” will have to do for now :)
I want that front and center in my AI phone as my personal assistant. The AI Phone's UI would be sparser.. wouldn't be a lot of UI (some app icons but not as app icon driven). While the image/video of your AI chat bot personal assistant you look and talk with could be a celebrity to a deceased relative or loved one (they live on and help you through your day to day). There's so much things to innovate and move forward from the boring iPhone!
Hopefully Open AI makes an even bigger announcement of getting into the personal device business soon (or later).
This is the old owner of the number. Some carriers are still routing it wrong.
Not very confidence inspiring.
For the past 8-10 years it has all felt like a bunch of apps that just aim to be mediocre middlemen/gig economy brokers with bad customer service.
And I'd wager that there are silent revolutions happening all across colossus that's the tech industry that will become apparent in the next decade.
Jeff Bezos put it best during his recent interview at the 2024 NYTimes Dealbook Summit, "We're living in multiple golden ages at the same time." There's never been a better time to be alive.
It can sometimes be useful to input a more "human" search and have something get spit out but 60% of the time it completely lies to you. I'm talking about questions related to web specifications which are public documents. Section numbers, standards names, etc.. will be completely made up.
So much this. So many times I've argued with hired experts saying "can't be done" just to see yes, it can be done.
I'm glad ChatGPT didn't lead you astray, but I'm not seeing what it's added here besides shuffling up the user interface in a way that you presently and subjectively prefer?
My general rubric is: “would I trust someone on Reddit to correctly guide me on this”. If the answer is “yes” then ChatGPT is likely going to do well. If the volume on a particular subject is low / susceptible to false information then it’ll lie.
Recently it lied hard about how to configure MikroTik routers. I lost many hours. But for a large construction project recently it completely balled out.
Are you doing cutting edge / complicated stuff? Have you examples of where it lies?
No specific prompts, but most were related to the XHR/Fetch specs and behaviors within. It would say "X.Y.Z sections defines this" but that section didn't exist at all and the answer provided was not accurate.
> My general rubric is: “would I trust someone on Reddit to correctly guide me on this”. If the answer is “yes” then ChatGPT is likely going to do well
I see. Well, I don't know if I find that very valuable but if others do, then so be it.
- completely made up books
- real books that were only marginally related
- real books with really bad reviews
I'd estimate that only 30-40% of the time did I find the results at all useful.ChatGPT lies a lot about RouterOS, I don't know why. Claude helped me a lot on the other hand with all things MikroTik.
In the past week I have used it for helping write a script in a framework I'm not super familiar with (OpenSCAD), I was able to finish a project in 5 minutes that otherwise would have taken me hours. I have used it to help make movie recommendations (none of them were hallucinated). I have used it to translate a conversation with a non-english speaker, etc. There are other tools that can help me do all of these things, but none quite as fast or painlessly.
It might not be useful for your use case of asking questions related to specific web specs, but that doesn't mean that the technology has no value. Horses for courses...
They can make stuff up, but saying "60% of the time they lie to you" hasn't been true for years.
If you're using them to fill knowledge gaps, what scaffolding have you set up to ensure that those gaps aren't being filled with incorrect-but-plausible-sounding information?
Imaging being graded on your ability to quote exact line numbers of particular parts of your codebase as a senior software engineer without being able to look at it!
LLMs are not, in isolation, a search product.
This is such an exhausting conversation
I had the same impression about the hallucinations 2 years ago. The reality is in at the end of 2024, you can get incredible value from LLMs.
I've used copilot to code almost exclusively now for the past few months. Anyone still comparing it to text completion I feel is operating on completely out of date information either intentionally or unintentionally.
The hard part is, despite actually having some "real" value delivered, you still have to sort through the 99% of bullshit that comes along with it anyways.
I'm also going to stand up for AR/VR here. I'm in a long-distance relationship and me and my partner spend an hour or so in VRChat around two to three times a week. The power that has to reduce the badness of an LDR is well well well well worth the three hundred bucks I paid for a Quest. That and some of the golf games on it are fun.
I've had an HTC Vive and an Oculus Rift 3 (Walkabout Mini Golf is one I tried!) and while I wouldn't try to argue NOBODY has found a use for it (somebody somewhere found uses for all of the things I mentioned, just not me and just not the majority of people like big new things are promised to) it never really ticked the "new value" box before they ended up in the closet for me.
Lots of engineering involved
Isn't this the new LLM playbook?
I pay Claude/ChatGPT trivial amounts of money for metered API access to their models, and they in turn provide it to me.
Middlemen/marketplace models like "Uber for x" or "Etsy for x" or "Betterhelp for x" is a totally different business model.
Yes.
> and adding value.
No. The only breakthrough innovation LLMs gave us is the ability to speedrun the making of racist pictures. Not sure the world really benefited.
EDIT: There is one limit at zombo.com. The limit is myself.
The scary thing is it's actually conceivable to somehow integrate GPT into those things.
Maybe GPT-4 is the 1080p of LLMs: Noticeably better than 720p and 480p models, and not bad enough to warrant additional improvements.
Sure, 4K, 8K, ... are technologically available, but for the majority of use cases, 1080p is enough. Similarly, even though o1 and other models are technically feasible, for most cases the current models are enough.
In fact, GPT-4 is more than enough for 80% of tasks (text summarization, Apple (un)Intelligence, writing emails, tool use, etc.)—small models (<32B) are perfectly fine for those tasks (and they keep getting better too.)
Alternatively, does it not seem more likely that they have different product groups? Surely the folks working on ChatGPT are an entirely different beast than the folks working in model development?
Nothing I said was absurd in response to making an unsupported idea that model development has plateaued.
o1-pro is that model. Expensive and slow, but significantly better at many tasks that involve CoT reasoning.
My experience is that o1 is extremely good at producing a series of logical steps for things. Ask it a simple question and it will write you what feels like an entire manual that you never asked for. For the most part I've stopped caring about integrating AI into software, but I could see o1 being good for writing prompts for another LLM. Beyond that, I have a hard time calling it better than GPT-4+.
How have you been using o1?
Even writing, where it is supposed to be worse than 4O, I feel that is does better/has a more solid understanding of provided documents.
This is what worries me. Aside from programmers and few other professions, most jobs in our civilization are prime for automation...
I know it was that good, because I got it to do that for me… and then the UI kept getting better and the expensive models became the free default option and I stopped caring.
Google's offerings here are still a huge mess. OpenAI is crushing them right now at building products that people want, and making them accessible.
I have very nontechnical coworkers get excited about cool new things ChatGPT can do, but I'm not certain any of them even know we _have_ Gemini in our Google Workspace.
This would hardly be the first time Google has produced innovative technology which eventually fizzles because it never captured much mindshare outside of the tech news circles
Isn't that exactly (part of) what they were saying in the comment you replied to?
> […] building better tech. Gemini will be faster/better and it will have more features
At first I thought you meant that literally. ;)
I can now ask my phone to call ChatGPT. 100% hands-free. It’s only a few steps less than using the app, but there’s a lot of incremental value to not needing to touch my phone.
Concrete example: I’m driving. I ask Siri a simple question, but it can’t answer it. Previously, if I wanted to use ChatGPT, I’d have to stop, pickup my phone, unlock, open the app, get my answer, then start driving again. I’d never do that. Now, I can just ask Siri to call ChatGPT
https://support.apple.com/en-mo/guide/iphone/iph00fd3c8c2/io...
We had a video monitor in front of us showing the live feed, that kept distracting me personally.
I meant no disrespect, but from 2' or so, the conversation sounded more natural and things got smooth. Interest feature and I liked the 80's banner with the phone # like in the old TV ads!
I was always curious how things worked when I saw announcements on HN. So happy to share to satiate the next generations curiosity :)
You could text a question to CHA-CHA (242-242) and someone would google it for you as a human search engine!
It’s great to ship fast. But you need to maintain things as well. And that requires even more time and engineers and money in the end.
There’ll definitely be projects within OpenAI that will be shutdown in a few months, just because it hasn’t caught and/or engineers want to work on something new.
That’s how Google worked in the 2000s - shipping new things fast - but then there was Reader and now they lost everyone’s trust.
I'm not sure if I'd use term 'diversifying'. At least not in the sense of spreading themselves wider across more projects to reduce overall company risk (if that's what you meant).
I think that we're still very early into AI and because we're still not sure what kind of applications people will want to use in the future, it makes a lot of sense to experiment.
I'm aware that the way "caller ID blocking" works is that it just sets a flag on the call metadata, and it's up to the receiving carrier to observe it and not present caller ID to the callee, but I'm not sure whether bypassing that is a common feature carriers (Twilio in this case) provide to their users. (It's also possible that the only thing Twilio exposes at the API level is a "recurring caller" boolean, of course.)
In any case, even skipping the disclaimer based on having called before seems like a problem: Different people can be using the same phone line at different times. Wouldn't it be required to still read out the disclaimer every time?
* Limit user to 15m a month * Greetings unique to user state
It also needs access to Model/DB etc… which is all not exposable to internet
I've been looking into options for our non-profit tech startup (Ameelio) and about the best pricing I can find is about 1.35 cents per minute. It surprises (and saddens) me that it's still so expensive. I'm sure at a bigger scale you can negotiate better pricing, but based on the quick conversations I've had with vendors it doesn't get significantly cheaper.
[*] limited bandwidth (8 kHz), providing a valuable opportunity to enhance and specialize models for telephony applications, ensuring better performance and user experience even with low-fidelity audio inputs.
But being able to blame the user's phone line probably goes a long way to avoiding unhappiness due to testing :)
An e-mail version of this would also be nice.
https://www.therecycler.com/posts/82-of-german-companies-sti...
Here is a demo of it https://youtu.be/14leJ1fg4Pw?t=805
> am based on OpenAI's GPT-4 model. Specifically, you are interacting with an instance of GPT-4, which is designed to understand and generate human-like text based on the prompts it receives. My responses are influenced by the extensive training on diverse datasets, but I do not have access to real-time data or events beyond my knowledge cutoff in January 2022.
But the linked page suggests knowledge cutoff date is Oct 2023. It hallucinated an answer even to that....
You can get your own number and customize the agent.
Is this still an issue? Maybe I have had too high hopes for AI.
Then again, local noise reduction on modern phones/earbuds probably goes a long way to avoiding that problem.
it does the same as the chatgpt whatsapp chat, but well you can forward images to it, it can send your reminder emails in the future and can manage todos for you (some kind of memory)
if it would have gotten more traction i would have extended it that you can also forward emails to it and it responds to the original email as your assistant
(and hey, if someone from openAi is reading this, feel free to offer me a position as a product manager)
(ring ring)
> Welcome to ChatGPT!
Hi, what are you wearing?
> I am not physically capable of wearing anything as I am a digital assistant.
Hey! What kind of chat line is this anyway?!
> This is a virtual chat line where you can communicate with me, a language
> model AI, to ask questions or have a conversation on a variety of topics.
WTF? (hang up)
> ...
> ... hello?
chatGPT: ~ This may be recorded...
chatGPT: ~ You agree to openai terms and conditions...
Me: What's the square root of two?
chatGPT: What number do you want to know the square root of?
Me: Two
chatGPT: The square root of ten is approximately 3.1...
<click>
If they wanted to show how very non-understanding and un-intelligent chatGPT is, they are doing a great job. So much quicker to see in a voice interaction than through online query submissions.
I did notice some weird VOIP noise on my first call, so maybe it was receiving a bad audio stream.
My background isn't AI so I can't contribute to that. My background is WebRTC/telephony so I could build this. Even if I was involved in 'AI stuff' I would have zero impact, but I can build this!
It's likely the people implementing the WhatsApp feature, are not the ones working on the LLM models.
If they believe AGI is around the corner and they are competing with others to get there, seems silly to invest resources in standing up a phone line, etc.