Before you believe the ChatGPT hype, ask it something you know well
nwn.blogs.com
nwn.blogs.com
When ML models are trained on whether they give correct answers, rather than correct-sounding answers, I'll be interested in the bot they produce.
Thus it's not the UnemploymentMaker5000™.
There's a number of naive voices that believe it just needs a little this or that and it will abound forward in capability. That little thing they're asking is not actually little for AI. It just seems simple because humans develop the requisite concepts (such as compositionality) so young.
It's weird to me that people seem to expect current AI to be as good as an expert, especially at something as subtle as the intersection of society and technology. The hype around chatGPT is not that it's an expert in all things, it's that it's a novice in all things, and better than that at some.
For instance, I asked it to list common defects in injection molding. It did a pretty good job of listing and describing the things that I'm usually looking out for. On the other hand, I asked it about how much a solid tire is typically stretched when placed on a rim (e.g. for an electric scooter). It came up with a typical percentage range, and then I tried to get it to reason out how the cross-sectional dimensions would change based on that stretching. It made several really basic errors that I managed to get it to correct, but in the end, I didn't really trust its ability and or "knowledge" in this area.
I have found it to be pretty relaxing to use as a recipe source, since it will just give you a recipe and the basic instructions without the fake history and 17 ads. It's pretty good at taking some vague requests about having certain ingredients and wanting something in a particular direction, and giving a reasonable result. Maybe if I was an expert on cooking particular types of food, I would find its answers comical or pedestrian.
> basic instructions without the fake history and 17 ads
if this were ever to be truly commercialized as a mainstream source of information, do you think that whoever runs this will forego the trillion-dollar opportunity to still show you 17 ads? purely out of goodness of their heart?
In addition to that, at least a tiny fraction of those 17 ads' revenue went to the people who actually create recipes. In the brave new world, why would anyone bother putting out information if this chat middle-thing charges its users but doesn't ever pay to creators?
ChatGPT's answer (summarized)[1]: "Memorize the initial state, and track the complete state of the cube in your head as you solve a cube normally".
If you aren't a cuber, the answer sounds plausible enough (It does work for chess after all), but it's humanly impossible with a Rubik's Cube. The only way you can memorize a random state of the cube _and_ operate on it, is to keep your changes to a minimal so your memorized state gets incrementally solved - you solve a few pieces at a time.
[1]: https://twitter.com/captn3m0/status/1598301110955311107
Wait, how is ChatGPT like a famous robot secretly operated by a human?
Anyway, a lot of the recommendations it makes, he dismisses as being unfeasible or ineffective, but that doesn't mean ChatGPT is wrong. For one thing, it's based on a fallacious appeal to authority: I'm an expert, I disagree, so it's wrong. Or, in some cases: We tried a version of what ChatGPT suggests, and it didn't work, therefore it's bad advice. That's not necessarily the case, since there are plenty of reasons something might have failed in the past, like execution or timing.
Anyway, I know ChatGPT should not be trusted (it makes no claim to trustworthiness), I just don't like this reasoning.
For one thing it's being trained by a bunch of Kenyans behind the scenes:
Systems like wikipedia and stackoverflow have (imperfect) social mechanisms to filter out useless or false information. ChatGPT have none of that. No miracle.
That's the crazy part. This is pretty much how Google has worked for like 30 fucking years -- links to the relevant content that has the most links!
But people are all, "OMG ChatGPT is an amazing breakthrough that's going to argue in front of the Supreme Court!"
Even better: ask it something you just made up.
For GPT, there's no real difference and you shouldn't expect any.
I think it gives you more insight into what GPT actually is than just seeing it being wrong about something you're knowledgeable in (which shouldn't really be surprising).
The correct conclusion is not that 50% of kids are dumb and 50% smart. Actually, none of the kids knew but they had a 50% chance of picking the right answer at random.
Knowing what you don't know is an "advanced" skill many adults haven't even mastered. I'm not at all surprised chatGPT struggles to not-answer questions on things you just made up.
I can remember being a child and having this belief, but it wasn’t based on a guess. To me, it was logical.
I thought that because fish and chips accompanied each other on the plate, they probably also accompanied each other in the sea. I can distinctly remember picturing fries swimming around with fish, in what I imagined to be their natural habitat.
Those are things that I don't need to have a deep and accurate understanding. Usually, they are topics that I barely know how to start searching about them. So, for taking me from zero to enough so fast and intuitively, Chatgpt is really impressive to me.
I tried to have a conversation about S1000D governance, and it was making up whole new bodies of NATO governance in about two replies.
Imagine what GPT6 is going to look like
What?