ChatGPT Isn't as Good at Coding as We Thought
pcmag.com
pcmag.com
It’s like drinking the complementary table water at a restaurant and then leaving a one-star review saying “tasteless and flat”.
I seriously can’t believe this is maybe the tenth such “paper” I’ve seen making headlines.
It brings shame on not only the institution that produced it, but the journalists that failed to notice and published the clickbait headline without even reading the paper.
Plus, I've mostly given up on journalists. I get clickbait, that's why titles are there for. But clickbait doesn't have to equal badly researched. They could have just as well gotten clicks if they wrote an article "Researchers fail to prove ChatGPT is bad at coding because of this lazy mistake". But they mostly either have no idea of the field they are reporting about or are pushing an agenda. Research? Ain't nobody got time for that.
Maybe the logic is to study the most popular product? Would make sense to me.
This is especially painful considering the ludicrous pace of advancement in AI. You really have to aim for where the puck will be, not where the puck is. Within a year GPT4 will be out-of-date.
Fundamentally, the entire point of this "research paper" is to compare ChatGPT with Stack Overflow answers, coming to the conclusion that people prefer SO.
Yeah, well, meanwhile Stack Overflow usage has dropped off a cliff since ChatGPT become generally available: https://observablehq.com/@ayhanfuat/the-fall-of-stack-overfl...
I certainly prefer to ask ChatGPT basic coding questions because I get an answer immediately with no argument.
Instead, the only version information appears to be buried in section 3.1.2 ("ChatGPT 3.5 Turbo API is used").
I still think GPT-4 has an edge but its frustratingly dumb and loses context so often that I'm not sure how much of an edge it really has any more.
I fail to understand why we would want automatically generated language at all, except for the fact that you can sometimes make money off of it.
Considering we're where we are as a civilization because people could make money off of poisoning the Earth and causing catastrophic ecosystem and climate change for profit, I think we should examine our motives more closely.
This reminds me of what Bohr said to Oppenheimer - in the movie - when some doubt arose around his mathematical ability: "Algebra is like sheet music. The important thing isn't 'can you read music', it's 'can you hear it'. Can you hear the music, Robert?"
I'm very much wondering if GPT can hear the music, but time will tell if and to what degree this even matters.
What would be the equivalent of “making music” in this case? Talking haphazardly about your business?
I can see this working if something like “making music” is possible without knowledge of said music. Which in actual music is quite possible because it is an intuitive art. Programming.. less so, but I never say never.
I would counter this with:
- What value is it to your company to know that the developers who churn into and out of your company... while they're there, they actually know a section of your codebase quite well?
versus the alternative: the degree to which your devs know where your code lives or what it does is low, since they do not engage with it enough.
Sure, one could say: "AI will navigate our codebase for them".
But I think it becomes a bit of a slippery-slope, regarding the question of: "When/to what degree can really we take the developer out of the picture?"
Imagine being a developer, and no one on your team really knows how your product works.
Is AI building your product at that point then? Is that product actually going to exist & be funcitonal?
Extreme example: if a random algorithm picked out a few words from a dictionary and somehow cobbled a sentence together using those words, would you say "meaning" has been transported?
The "generation" of these symbols should be the result of a process equivalent to whatever we are doing when we "experience" or "cognate" or whatever, otherwise the results will be only very superficially useful.
So you're basically dismissing the whole field of genetic algorithms?
If not, what do you mean?
If you read a sentence (whose source is unknown to you) and it has meaning to you and affects you, what difference does it make if the sentence was written by a human or a machine?
I struggled with this too and because I am fond of defective analogies this made me think of the meaning of a personal note to, say, a lover.
Does it matter by whom or indeed what it was written? The experience of reading occurs solely in the recipient’s mind.
Surely it does not affect him/her differently if said note was produced by a romantic slot machine.
Before I meander towards my confused and shaky conclusion, let me clear one thing out of the way: I did not mean to imply that mechanically generated sequences are meaningless.
They absolutely carry meaning, because meaning is ultimately created by us. We are meaning-creation machines. We can find meaning in tea-leaves. Surely tea-leaves carry meaning, but sadly not the will-of-the-gods-kind. Although that doesn’t stop us from thinking they do.
What I am tentatively suggesting is that the meaning of the lover’s note is as much found in the experience of the relationship they share as it is found in the symbols on the paper. The symbols reference that experience.
Because it is ultimately a human that reads and interprets I think failure to take into account the origin of the material will be the source of a subtle and highly impactful, but inexorable error not dissimilar to that of the tasseographer: what you see doesn’t mean what you think it does.
Language used to be a domain dominated by humans only. We are wired for it and biased by it. If it sounds intelligent, it is intelligent, right? This bias could make us defer to the machine quicker than we should.
What’s the difference between GPT4 generating a good piece of code and a human?
Maybe I should shut up now. I am open to interesting reading suggestions by the way.
If there is no inherent message, if it is just a byte array likely to get a desired result in a specific context, you are losing traction of mind over reality and letting an awful lot of noise into our delicate systems.
Isn't letting accidental complexity aka noise into our systems bad? If you don't know the purpose and the context of the message, you already let that happen, and it will only get worse as the subsequent modifications to the systems described by the code or behaviors enacted due to the message starts to become an opaque box due to the fact that there is no meaningful blueprint behind the code or the texts underlying it.
Then you get a bunch of humans that interpret this soulless message, but fail to comprehend the inherent emptiness sufficiently.
I’d say that’s not good, but I have no solution either.
If you realize code and language are ways to communicate asynchronously and conserve information across time and space to enable developing and maintaining systems, the apparent value of generic concatenation of words that looks good and functions but has not meaningful message conveyed drops off precipitously.
I have maintained code, and even wrote some terrible stuff myself. All I'm saying is that once you strip all the lofty ideas about what code should be (which are all very good ideas!), then you are left with the very simple core idea: code is supposed to run and do stuff.
its not your new AI Friend steve, its a LLM. And one with a knowledge cutoff in 2021 too
Holy crap, you can do that? I need to learn about ChatGPT plugins.
https://chat.openai.com/share/fbd14b6c-0ba0-473f-9e94-db97c6...
Providing a prompt, and/or know how to properly prompt is key.