Personally I rather have a more limited but reliable tool than a more powerful but unreliable one.
Personally I rather have a more limited but reliable tool than a more powerful but unreliable one.
How many terrible teachers are allowed to continue to teach after decades of disastrous results?
How many ideologues with no interest in teaching, but every interest in indoctrinating young minds, are tolerated because the alternative is "no teacher" and the class wouldn't run?
How many teachers who would fail state exams, teach, despite relying on answer sheets to be "competent"?
I agree with you that on the top end of education, this is no replacement and at best a supplementary tool. For the poor kid in a bad neighbourhood whose teacher is more interested in "de-colonizing" mathematics than teaching mathematics, this is a Godsend.
Currently AI seem to have 2 weaknesses, it’s quite bad at reasoning and it hallucinates.
Yes, humans screw up reasoning too but at least they try. AI like ChatGPT seem to skip critical thinking altogether - there are examples of ChatGPT contradicting itself multiple times in a single conversation.
Then there are hallucinations. When people don’t know something, most just say they don’t know. AI just can’t seem to help itself but make up complete BS that it confidently try to pass off as facts.
P.S. Frankly I think the 2 might be related. AI can’t “sanity check” it’s own output.
Counterpoint: Why "hire" a known-bad teacher? Is a teacher that is known to give fake information better or worse than none at all?
I don't know the answer to this, but I hope that Khan Academy thought about this and decided that the good of GPT is better than it's bad. Or, they can't know so thats why they're doing a limited study before rolling this out.
It's worth discussing if we want AIs replacing them - maybe the scalability is better, but these are usually worse.
I would disagree. I think that at this stage, ChatGPT is very predictable, and we have excellent statistics on how likely it is to respond correctly to a particular type of question. I would have much less predictability about any particular human teacher, who might have an especially good or bad day.
Also, ChatGPT can be much more accountable, in the sense that when issues pop up, a filter can be implemented to prevent these particular issues for arising again. I don't see how humans are better in this regard - even if they are "accountable" in some vague sense, it can take many months to get rid of an underperforming teacher, and can take decades to change something across the system.
"What happened between 1850 and 1960 regarding civil rights"
A: "Oh nothing, it was great for everyone"
History is a mess here. That said it would be nice if it were incorporated with written texts, lessons, and the AI so you can bounce questions off of it.
Most humans have the decency to not do the latter while AI (at the moment) freely makes up complete lies when it doesn't have the answer instead of just admitting it doesn't know.
It's not usually that you ask them a question they can't answer, and they just confidently make up something entirely wrong, and possibly also an incorrect explanation that few or no other people have seen.
It's a different kind of wrong. One can write a book addressing common misconceptions picked up in school, from human teachers, and cover much of it. Does that work for lies AIs tell? They could be about almost anything, they could be broadly wrong, they could be subtly wrong, et c. "Show me some things my teachers may have been wrong about" seems to me an easier request to adequately address than "show me some things my AI tutors lied about", because the former tend to cluster around a kind of common space and form an experience with a great deal of overlap between students, while I'm not sure the latter do.
What's the analog of that with human teachers, and how long would it take to implement a solution there? Even if you had all the budget in the world, how would you go about making all teachers stop misleading students about how airplane lift works?
How about we fix the unreliability first and not put the cart before the horse.
> The elitist mindset needs to die, let's give access to advanced knowledge to everybody and stop putting it behind some unreachable walls.
What are you talking about? We have high quality university courses on the internet for free. We have things like the Khan Academy and similar sites. How is advanced knowledge behind unreachable walls?
With context, which Khan academy has in abundance due to their lesson plans and transcripts, accuracy will be higher than even the best teachers and tutors.
Once you give context and known-true facts to best-in-class LLMs like GPT-4, the output is shockingly good.
As far as I understand the preferred way to use LLMs nowadays for domain specific information retrieval is through embeddings that insert the related context in the prompt. GPT-4 is specially good for this since they increased the prompt size almost by an order of magnitude.
This means that the model can be given a very specific task: to extract information from the context or avoid providing an answer at all.
The answer doesn't rely on the neural memory of the model, since it doesn't need to store information, just understand the task, and they are really good at that.