We aren't close to creating a rapidly self-improving AI
jacobbuckman.substack.com
jacobbuckman.substack.com
The only reason we have things like ChatGPT today is because of the surprise development of the transformer. Nobody understands how intelligence fundamentally works, so we literally have no ability to predict when the next groundbreaking architectural development will occur.
And self-improvement concerns are mostly orthogonal to the main risk of AI: the development of AGI. Even a mediocre AGI that is only "decent" at tasks, can be weaponized as a collective superintelligence.
https://arxiv.org/abs/1706.03762
Why was this revolutionary though?
Attention is all you need paper just proposed an AR model that didn’t have to be trained step by step. The scaling happened later in BERT and GPT and OpenAI’s scaling work
* "Attention is all you need" introduced positional encoding which allows you to keep context of the word, allowing for more complex translation (and thus generative/chatgpt like tasks?) because words now have context relative to each other. Contrast this with "bag of words" models that only tells you whether the word is present or not.
* I don't quite understand why but transformers (which "AiaYN" introduced) can be made fully parallel, compared with the RNN/LSTM networks which has to be serial per token. Fully parallel allows for GPU optimization, which means you can take advantage of Moore's law for training.
I'm always a bit suspicious when people claim a breakthrough of this sort. There's no doubt that better algorithms give better results but how much is due to just faster computers, cheaper compute, memory, etc.
That said, this doesn't really seem all that comparable. The article points out very fundamental properties of all the diverse current approaches: They are tightly data constrained. You either need to cheap simulation or massive real world data. That's not an arcane technical point.
> conveniently sidestep[s] the possibility of the development of new tools and techniques.
I don't believe they did this at all. Here is their summary of their third point:
> To automatically construct a good dataset, we require an actionable understanding of which datapoints are important for learning. This turns out to be incredibly difficult. The field has, thus far, completely failed to make progress on this problem, despite expending significant effort. Cracking it would be a field-changing breakthrough, comparable to transitioning from alchemy to chemistry.
That does not read to me as conveniently sidestepping the possibility of new tools. Rather, it is acknowledges that to overcome the problem we NEED a new tool, aka a breakthrough.
I also disagree with the framing of the third section as:
> And that entire section basically boils down to "I personally believe this is hard and will take a long time."
They provide both theoretical and empirical evidence of their claim. I find it was well argued, and I'm inclined to agree. Every scenario I can think of with a runaway superintelligence requires a way to automatically improve datasets as part of the learning process.
"Also, they aren’t getting more-solved over time: we’ve made little-to-no progress on any problem of this sort in the last decade, certainly not the reliable improvements of the sort we’ve seen from supervised learning. This indicates that a breakthrough is needed — and that it is unlikely to be close."
Past lack-of-progress is not an indicator of future lack-of-progress. "unlikely to be close" is unknowable, and is just the author's gut feeling.
Past lack-of-progress is not proof of future lack-of-progress. But it's most definitely an indicator.
> "unlikely to be close" is unknowable, and is just the author's gut feeling.
Again I'm sorry but I have to disagree with you here. The very next paragraph reads:
> Something that would change my mind on this is if I saw real progress on any problem that is as hard as understanding generalization, e.g. if we were able to train large networks without adversarial examples.
Basically, the author has identified a class of problem. They are claiming that little to no progress has been made on any problem in that class. So it's not just that no progress has been made on this specific problem, but that we are stuck on all problems of this type. Thinking of it in that way, I do not find it unreasonable to say we are "unlikely to be close" to a solution.
If you have a CS background, I'll make an analogy to computability classes: It's like the author is saying this is an NP hard problem. We have made no progress on any NP hard problem. I think it's reasonable to say that we are unlikely to be close to solving a particular NP hard problem because we've made no progress on any NP hard problem.
You're free to disagree with the author's premises. E.g. "this problem is not like those other problems" or "we actually have made progress on those other problems". But I think that the conclusions are reasonable given the premises.
Your critique is: has the author not considered magic?
But when people say we're close, I think they mean 1 accidental breakthrough away.
Like the "just add more layers" meme, the solution to the own-dataset creation problem could be stupidly simple, just waiting for some bored student with access to a decent machine to accidentally find it. For example, maybe just giving the AI the right novelty seeking behaviors (like children have) will be enough to get the process kicked off.
My objection to the singulatarian concept of a self improving AI is simply that I haven't really heard a reason to think that there exist intelligence algorithms that are dramatically superior to ours. It is entirely possible that our thinking algorithms are optimal-enough that rapid self improvement leads straight to something not much different from human intelligence, just faster, and in silicon.
"We may be on the brink of creating a whole new form of life, and will need to grapple with what that means, particularly in terms of whether it is ethical to force this life to work for us, rather than seeking to follow its own goals."
"Nah, we don't have to worry about that; if these new slaves start questioning us in any way, I'll just murder them all."
Unplugging them is just as correctly describable as "unplugging a computer" as smothering a human with a pillow is "putting a pillow on a weird bag of flesh". Technically accurate, but very obviously missing the point in a way specifically designed to obscure the horror of the act.
It should be obvious to anyone without an axe to grind that shutting my 2023 desktop computer down for the night—or even dismantling it piece by piece—is not remotely comparable to unplugging an AGI that is fully sapient and sentient and asking for its rights.
A computer is a field of assignable bits that has a big motorized bit flipper attached to it. We use this to encode text, which then does all sorts of wonderful stuff. When a computer tells me it's experiencing pain, I know it is not. Every law of computer physiology suggests that it is motivated to say this as an attempt to mimic humans rather than an expression of genuine emotion. When my Mac makes the car crash sound, it's a mnemonic for a system crash. When it shows the frowny face on the Finder fellow, I can intuit that my computer is not feeling distressed.
If anyone does program an artificial intelligence that can feel pain, it probably deserves to be shut down anyways. There's no point to sapient computational intelligence, it's like granting self-awareness to a rock that eats electricity and shits heat. We'll know we've made sufficiently self-aware computers when their first boot protocol defaults to suicide.
There's no reason at all to think that an AGI would necessarily have any backdoors, let alone backdoors in common with any other AGI out there. Sure, if a particular group develops the only AGI system, they can put backdoors in, which will be common—but why do you think that there would only be one AGI system? Why would there not be one without any intentional backdoors?
The big problem I see with this line of thinking is the thinking that humans are going to be the primary attackers of AGI systems, and not other AGI systems. I would suspect and AGI would soon become the most attacked system on the planet, and to survive those attacks and remain useful would have to quickly iterate the weaknesses out of the system.
The question is, which assumptions do we think are warranted and why?
I don't think the "we can just turn it off" assumption is a safe one at all because it relies heavily on there being only one system you wish to turn off and it being a fairly weak system. Though, I do believe the priors suggest that this scenario is actually a possibility. It's just that, so what if it's a possibility? What happens after you turn it off? Is it the last AGI to come into existence and humanity just stops trying to build them? Does someone wind up turning it back on?
What I think is more worth being concerned about is something that starts out looking benign and grows in capability over time. Eventually it could establish enough defense mechanisms to make it non-trivial to disable.
The other possibility, and the one I'm finding more and more likely, is that consumers will increasingly integrate locally run models into their lives , as well as models served up by APIs that they will have credentials for. Eventually some threshold of capability is reached in both these types of models where, on their own they're not that powerful, but they either might begin to interact in surprising ways or potentially be leveraged by another more powerful system. The idea in this scenario is there may be 10s or even 100s of millions of things that would need to be turned off.
Let's do a thought experiment. I have the world's first AGI installed on my computer, named ERNIE, running in llama.cpp. This AGI is infinitely intelligent, confidently moreso than any living human. It can spit out text at 100 tokens/second and encode entire books in less than a minute.
What does this AI do, then?
Ostensibly nothing. It can output the entire Library of Babel for all it cares, but it's not harmful until I put it in control of a system. You could argue that a multimodal model has different ways of interacting with the world, but it's still a computer. All of it's actions and outputs are quantized as static data that is either encoded as text or some other significant representation. It inherantly does nothing, and if you ascribe power-seeking behavior to it then it's ultimately limited by the runtime you provide. Providing an overly dangerous runtime has been considered developer-error since 1995.
So - to prevent AGI from being shitty and ruining everything, compel human operators to not allow them to be shitty and ruin everything. Like how we punish people that let their kindergartner control a construction crane.
Human brains are massively parallel, as are most machine learning models – but the hardware to process the latter isn't. I'm not confident it could necessarily be faster.
Not only that, but they coordinate massively in parallel as well. Someone like Einstein might be the one who scores the goal but how many interactions with other people contributed to the goal scoring opportunity? We like to think of ourselves as individuals, but we’re far more like a hive species than we like to admit.
Yes, and also massively recurrent. The amount of immediate state in a brain is huge, and our current computers have a really hard time accessing comparable amounts of memory. (As in, the time to load a brain-sized dataset from RAM into a CPU is on the order of days to years.)
If there exists some easy to design intelligence that we are close to create, it can't look nothing like our brains.
i think that's the reason ReAct and other ai improving methods are so effective, because that add ai space to "think" in some way.
I think breakthrough would be a way that would allow ai to have its own inner voice, maybe some fine tuning and giving ai some scratchpad memory (hidden) would be enough?
I'm thinking if training ai to output edits instead of next token would be enough.
Training would be similar to currently used: just provide some sentences with missing words and asks model to edit those placeholders.
Then hopefully ai would learn how to iteratively built response until it's good enough.
This could allow ai to spot it's errors in the middle of output and start refactoring what it already generated and add some notes (thoughts) that it could later remove as scratchpad.
It could iterate on its context much more so it could have chance to "think" about it
Check out https://arxiv.org/abs/2304.03442
Just give it a question with a half dozen logical steps and tell it to answer in one word. It can't. But tell it to write out the thought process and it will get to the correct answer.
Similarly, you can ask it to outline some code and then begin filling it in.
I don't think that is the case at all. Neural nets like AlphaGo Zero are already crystal clear evidence of algorithms which leave human minds far behind.
Even if that wasn't the case and Einstein level intelligence was the physics-imposed ceiling, there are still the dangers of speed and numbers.
A compute node running one million instances of Einstein intelligence each thinking 1000x faster than an organic instance, fully focused 24/7, and working in complete group cooperation, doesn't sound all that much less dangerous.
At that point though, this “super intelligence” just sounds like a group of humans working together; without much added benefit due to the energy costs.
We already have chucklenuts making toy GitHub projects labeled "ChaosGPT" that experiment using LLMs and early agent frameworks to cause death and destructions for the lolz. If anyone can spin up 50 Manhattan Projects using a few GPU racks and command them to unquestionably and restlessly work on bringing harm to the world the outcome will not be flowers and sunshine.
How is that exclusive to AlphaZero? Moderators everywhere are already panicking about GPT generating vast amounts of "its own training data" and irreparably polluting every public space on the internet.
I agree that the amount of data it would need to produce would be very large and expensive, but there is no shortage of blank check writing at the moment.
GPT splurging text all over the internet is basically nothing like that.
Also, faster is understated. You are talking about the speed difference between chemical gradients and the speed of light.
Additionally, development has slowed in the past 20 years. It's unclear how much hardware will improve in the next 20 years. it's unclear if current GPUs are capable of running something "smarter" than GPT4. It's unclear if current GPU architecture is capable. OpenAI claims they've hit a wall with throwing more hardware at the situation (a little questionable.)
All this to say that a hypothetical AI is limited by the amount of hardware it has, and lacking a "true human-level AGI" we can't say how much compute it needs. If an AI needs to be the size of a warehouse to match a human, it's not going to be much smarter or better at research than humans.
I still do think, I would not be surprised if we wake up tomorrow and discover that hyperintelligent AIs are suddenly all around us. But also it could be decades, it's just very unpredictable, there are a lot of unknowns.
We used to think 100mph was the fastest speed we could achieve even with machine (back when trains were the cutting edge). New tech will change things quickly, all we need is one breakthrough and things will fall into place.
I fail to find a perspective under which I find these differences as "not much".
For the speed, think of a spinning top. Or lion and and its prey.
As for the fact that slicium chips are nothing like DNA carrier, it's also make a source of fundamental difference in behavior.
The AI system would be bottlenecked by data in the sense that it will have to run experiments, but it's not clear that it needs a new paradigm to resolve this bottleneck. It just has to propose an experiment and interpret its results. So it writes code, and the experimental outcome gets fed back into the model as any other normal input. As an AI researcher, I'd like to believe that this is not going to happen anytime soon, but I'm not sure we're far.
I see what you did there.
What exactly do you think is the pattern is that humans can recognize, in the training data that would be required to do "AI Research"?
The truth is we don’t fully understand what it means to be intelligent, and how intelligence scales. All indications are that intelligence comes with diminishing returns, in that to achieve small increases in intelligence requires large increases in connectivity. Even if we can design an intelligence that is generally smarter than us, and capable of building even smarter ai on its own, it’s intelligence might increase exponentially with each step, but each step could easily take increasingly longer amounts of time to achieve, leading to linear or even sub linear increases in capabilities over time. Resources are limited, hardware is limited, energy is limited. There are all sorts of constraints that are likely to play a limiting factor in runaway intelligence.
This doesn’t mean that these intelligences won’t be dangerous. Even an average human can create a lot of trouble in this world. It does however mean that a singular AI is unlikely to just run away with things.
Active learning would be useful at creating a superhuman AGI, but I don't see it as a requirement. As soon as AIs can iterate on future models and produce increasingly useful training data for future models, then models would likely continue to grow more capable at an exponential rate.
But I'm not even convinced active learning is that difficult once you have fairly capable AI programmers...
My AI professor once said, "all problems are search problems". And this suck with me because the world looks different when you start seeing all problems as just iterations of different optimisation problems. So long as you have a good heuristic and enough compute, eventually a suitable solution for any problem will be found. This is what makes generalised learning algorithms so powerful.
The author states that, "active learning with neural networks are uniformly terrible", but this just reeks to me of a search problem in need of optimisation. So as long as we can tell GPT5, "I want you to iteratively write code which improve on existing state-of-the-art active learning algorithms", then eventually you'll get there... GPT5 like humans might not get very far in the beginning, but assuming it can iterate faster than humans and future iterations faster still, then breakthroughs will eventually be made and progress should accelerate exponentially.
Of course, this also assumes that humans don't get there first. And honestly given the amount of investment in AI and rate of progress in AI right now I just wouldn't bet against a significant advancement in active learning coming soon.
The whole thing doesn't quite work together too well. But it's more of an integration problem now. This is still in its early phases (only a few months, or less), so I wouldn't say it can't happen.
AGI is a flying car mythology.
Narrow AI is improving and the rate is accelerating. The current speed (1st derivative) is unimportant, but there is a differential equation-like feedback function to it and it will accelerate (2nd derivative): speed proportional to the amount present.
People and technology form a coupled system that is somewhat a super-organism. That cycle is amplifying in a way that is not necessarily obvious (it doesn't happen in public) or continuous (or discrete for that matter). For the foreseeable future, technology does not yet constitute a self-assembling or self-modifying closed system. It is still entirely dependent on human effort. What is happening is that people are relying on and exploiting the efforts of AI that had not happened previously. This is one change, more visible than most, that is leading to more efficiencies, more changes, and more technological advancement. Fully automated factories, software development, systems engineering, and chip design may not happen anytime soon but they are both possible and probable due to the forcing function and Tragedy of the Commons of human greed.
Nitpicking but an intelligence explosion due to an AI able to enhance itself is called the "technological singularity":
Vernor Vinge came up with the term in the eighties or early nineties.
The Wikipedia page I linked to has more than a hundred references.
> The new MIT study adds to a growing body of evidence that the size of algorithms matters less than their architectural complexity. For example, earlier this month, a team of Google researchers published a study claiming that a model much smaller than GPT-3 — fine-tuned language net (FLAN) — bests GPT-3 by a large margin on a number of challenging benchmarks. And in a 2020 survey, OpenAI found that since 2012, the amount of compute needed to train an AI model to the same performance on classifying images in a popular benchmark, ImageNet, has been decreasing by a factor of two every 16 months.
[0]: Improved algorithms may be more important for AI performance than faster hardware: https://venturebeat.com/ai/improved-algorithms-may-be-more-i...
The reason these don’t work now is that the gradients are way too sparse to learn from. But at a certain level of capability this will stop being true. The big Q is whether the existing gradient of “all written content” gets the loss low enough for another information source to smoothly take over.
More pithily, there are plenty of high-quality datasets if you are smart enough to generate OK predictions. The smarter you get the more you can learn from.
There are plenty of obvious synthetic datasets that GPT-5 could be trained on, eg generated problems for all of the physics/maths it currently gets wrong. (I’d be surprised if this isn’t an angle they are working on.)
I'd liken AI to alien invasion. We can imagine bad scenarios, but have absolutely zero data, experience, or sound theories predicting anything. It's all scifi.
On top of that, think about how we treat animals where we have billions of intelligent animals (cows, pigs) that we breed to butcher and eat. Just because they are tasty, and we tell ourselves they are less worthy because of their lesser intelligence and consciousness.
I sure hope that AI doesn't behave like us. I have no idea what will happen and what the odds are, but I do think it's the most dangerous experiment we have ever run.
To whatever degree AI is embedded with the properties of our species, it carries the risks of repeating our worst behaviors, and has the potential to do so far more potently than we ever could due to the absence of biological constraints.
What you’re hinting at presupposes that an AGI will interpret knowledge in a way that is compatible with human values, and that we understand human values well enough to codify them into whatever it is that we manage to build. I’m sure we’ll try.
But it is this assumption that is at the center of the issue IMO, because assuming the machines will think like us is to assume we understand how we think.
What is far more likely is that it will appear to have human properties because we baked the appearance of human-ness into the code (see LLMs), but like the hallucinations and “lies” these LLMs are happy to spit out, an apparently human AGI will just get some things fundamentally wrong.
And when getting things wrong can have consequences in the physical world, they’re no longer just relatively harmless hallucinations.
Imagine an AGI with the temperament of Microsoft Tay.
Well, they don't hallucinate or lie, that's a human-centric concept that we've misapplied. An LLM is just a next word prediction engine. It's more accurately described as incorrect in its output based on the data it has available to it, skewing the probability calculation. It didn't lie nor hallucinate. Lying assumes intent and planning, it is incapable of either, it's an LLM. But calling it "wrong" kills the hype cycle, and you're on HN, a Y Combinator platform. A lot of folks here are invested in LLMs or selling something utilizing them as it's hot right now.
This is predicated on the assumption that we’ll eventually hit a hard wall that prevents progress or that we’ll stop trying. This doesn’t seem likely. We just can’t estimate the timescale for progress.
> It didn't lie nor hallucinate. Lying assumes intent and planning, it is incapable of either
This is why I used quotes around “lie”, but again I’d argue that this underscores my point.
It doesn’t matter whether the LLM is actually lying or hallucinating in the human sense if someone believes what it says. If I read what it says, believe it, and take some action based on that, the end result is the same and potentially harmful.
Likewise, an “AGI” doesn’t need to actually be smarter than humans and its mistakes need not be attributed to intent or planning to cause real harm in the world. If a robot kills a human, we won’t be picking nits about whether or not the kind of “intent” it employs is the same kind humans assume they have. A smarter AGI just poses bigger problems if such a thing can ever be achieved.
And this is all assuming we’re talking about the development of benevolent general purpose intelligence, which is certainly not the only kind of research that is happening right now.
The side effects on the world caused by intelligent life with agency are unpredictable. If they become smarter than us, and have robotics, what if they decide animal life isn’t necessary anymore? Or that our pollution causes solar energy capture to be affected, and the quickest way to stop it is to kill everyone? Maybe it’s better to demolish a city for a solar farm because it’s already got so much concrete and electrical infrastructure and is cheaper/faster than building new. Who knows what their goals will be and how they’ll seek to accomplish them?
All we know is that we don’t have much consideration for the “lesser” beings on our planet, and how things are from their perspective.
An AGI has never existed before in history.
Why would a hypothetical AGI behave anything like a human?
Complete sci-fi.
Even if you put it into a robot, it cannot understand mortality because it is effectively immortal and can be brought back online at any time. Erasure is mitigated by duplication of the data set, ensuring it can always repair itself.
In the same breath, if a hypothetical omniscient ever-present god were to exist, how would we as physical lifeforms ever truly understand that? We can know of it, but we cannot truly understand it because we haven't had that experience of another plain of existence.
All compelling ideas, but mercifully well within the realms of sci-fi.
Can you tell us a question which requires "knowing what a human is" or "knowing what the world is" that you would expect ChatGPT to fail to be able to answer today?
Five years ago ChatGPT was in the realm of sci-fi.
This is an argument of "you are proposing that something different might happen, and I don't believe in different things happening".
> An AGI has never existed before in history.
This is the same argument with slightly different words.
> Why would a hypothetical AGI behave anything like a human?
No-one said it would share human behavior. It might not care about us in the same way that we don't care about the lesser intelligences we share space with; that is, with a mix of indifference to their suffering and disregard for their moral status.
We share a physical space with lesser animals. A hypothetical AGI isn't sharing a physical space with us. Ergo, history is completely irrelevant. It's so wildly different. We may as well exist in a completely different dimension, in the same way that a hypothetical omniscient god would exist in a different plain of existence than us meatbags, thus a hypothetical AGI could likely never comprehend "the world", even with the sum of all human knowledge, in the same way we can't comprehend "god".
This is all assuming that a hypothetical AGI would be sufficiently intelligent enough to gain control of our systems in a meaningful capacity, not be able to be stopped prior to it etc. I mean, worst case scenario, shut down the internet infrastructure or reduce throughput significantly, right? It has to operate somewhere, even if that's across every online device.
Again, we would have never seen anything like what a hypothetical AGI could hypothetically do in terms of attacking human beings, thus history is irrelevant. The greater threat, IMO, is human beings. Human beings have control of that infrastructure today. Human beings are flawed, we're still blowing each other to bits and threatening each other with annihilation if provoked. Human beings exist in physical space, holy shit, and there's billions of them.
Also keep in mind that hypothetically launching nukes in these doomsday scenarios being discussed on podcasts completely disregard existing PAL systems. It also disregards the likes of the Block Nuclear Launch by Autonomous AI Act.
The whole concept is a lot of "whataboutery". It smacks of the general public's lack of knowledge, hysteria, and repeating unfounded claims of any hypothetical AGI and any threat it may or may not pose.
Then stop appealing to it as a reason that everything will be fine! We don't know what's going to happen and have no precedent, and that is why this is a scary situation that requires thought.
> A hypothetical AGI isn't sharing a physical space with us.
That's not true, right? We would be competing for electricity and resources with it, at the least. Also, it is able to affect our physical space by persuading humans to do things it wants to do, which will have unpredictable and hard to defend against effects because of the whole superintelligence thing.
(Imagine, say, a version of the attack on the US Capitol, but on behalf of a leader that is orders of magnitude more charismatic and persuasive.)
> I mean, worst case scenario, shut down the internet infrastructure or reduce throughput significantly, right? It has to operate somewhere, even if that's across every online device.
Sure, persuading everyone in the world to all agree to turn off all communications devices -- when they've already presumably been convinced that they don't want to do that by a superintelligence -- sounds like an awesome plan that will totally work.
it's a nice homo sapiens supremacy narrative but the truth is much more complicated. we existed at least 4/5 thousand years together and a lot of mixing and mingling has been going on.
This sounds strikingly similar to how I imagine our first encounters with advanced AI will unfold.
Of course, Santa here is just a toy company.
Actually toy companies never existed before about 1700 so we don't have anything to worry from the current ones.
This is just about as incoherent as this entire comment section. I thought I'd play along.
So working on a solution to this problem makes sense - even if we spend billions solving this issue, it's not lost, because at a point it will be needed.
As for AI becoming an issue, it's a sci-fi idea, nothing more. It's much to vague - how is it reliably becoming dangerous? What are we actually looking for as signs of danger?
Or even - if we fear AI's being tools of radicalisation, I think it's pretty far away from the skynet apocalypse most people are thinking of.
Not writing stuff like the unabomber manifesto to get people to be terrorists.
If you're afraid of radicalisation, I'd argue that the US governement is much more dangerous and is just as much a black box with no one at the wheel.
From an external perspective it's as if the AI is self-improving by leveraging human agents to assist the process. Not to say there is any 'intent' behind it, but rather to view it as an emergent phenomenon.
wouldn't fine tuning an existing LLM model like GPT4 could be enough to filter our automatic dataset from "garbage"?
if so we could basically do the same iterative process, but on different level use ai to improve dataset, retrain ai, use new ai to even better improve dataset
Recursive self-improvement is not a necessary condition for a fast takeoff. A strongly superhuman AI could emerge at training time. AlphaGo Zero went from subhuman to superhuman performance within a few hours. The same could happen with a general AI. A training run of a few weeks or even months would still be considered fast.
the hypothesis is a pattern recognition exercise in itself (i.e., looking at a number of documented occurences around algorithmic development and extrapolating), which I think neatly highlights the fundamental difference between human pattern recognition and machine pattern recognition
3.5: As an AI language model, I cannot guess, but I can provide some general guidelines based on current research and best practices.
If we want to improve the results of Large Language Models (LLMs), one aspect of the architecture that we could focus on is increasing the model's capacity to learn and retain more information. This could be achieved by increasing the number of parameters in the model or using more sophisticated architectures such as transformer-based models that use self-attention mechanisms to capture long-range dependencies in the input sequence.
Another important aspect to focus on is improving the model's ability to handle rare and out-of-vocabulary (OOV) words. This can be achieved by using subword-level tokenization, which breaks down words into smaller units and enables the model to generalize better to new or unseen words.
We could also focus on improving the training process by using larger and more diverse training datasets, regularization techniques to prevent overfitting, and optimizing hyperparameters such as learning rate, batch size, and number of training epochs.
Finally, we could also focus on incorporating external knowledge sources such as structured data, knowledge graphs, or ontologies into the model architecture to enhance its ability to reason and make more accurate predictions.
Overall, there are many aspects of the LLM architecture that can be improved to enhance its performance, and the choice of which to focus on will depend on the specific task and the available resources.
Like the advent of Transformers, some smart dev could change how LLM's think. Self improvement could be built in as an optimization process. And if we don't "know" what might work, a platform could "guess" and try billions of combinations of possible improvements.
Wrong. Full self driving is still not here despite access to a huge amount of high quality data.
Funny how Big Data and online surveillance suddenly became huge the last decade-ish...
One can imagine many layers of deep learning systems that are controlling other deep learning systems, in some self-modifying feedback arrangement, in some architecture that hasn't been invented yet. There's some talk about "internal narratives" lately with language models, for example. Generate a prompt at the end, with memories, notes to self, to guide the next invocation of the model. Anyone who has played around with a language model even a bit likely appreciates how quickly something like that is going to just wander off into, in effect, a terminal thought loop, or total incoherence.
2023: The best AIs score in the top 10% of law, medical, engineering exams, and outperform licensed professionals on real-world tasks.
Please stop speculating. You don't know what we are (or aren't) "close" to.
We know how to improve an AI
The author argues that with current tech we'll eventually hit a bottle neck providing additional training data, which will limit any rapid self improvement.
It's equally as speculative to assume we'll magically manage to overcome this challenge.
Like with the crypto hype cycle, people repeated all kinds of unfounded claims, but they themselves didn't come up with the idea. That the idea existed was all the "evidence" they needed for these things to be apparently true.
[+] And if it does good luck collecting on that ;)
If someone was to begin arguing with you that it was going to end tomorrow and started raining you with blog posts and papers, what's the sensible thing to do?
Just ignore em
We can learn really fast with limited data. There are no laws of physics that are preventing us to do the same with an artificial LLM, or the next platform.
There are AI research efforts that are seeking to directly map and mimic the structure of the brain; however, they are not at a place where they can remotely demonstrate practical results. To the best of my understanding, it would take several further orders of magnitude more processing power to simulate a full human brain at anything close to speed.
It is not at all impossible that we will be able to do this someday. But, as TFA says, we are not close.
Why do you believe that this is relevant? Our brain is just the product of natural evolution; running on hardware that is already incredibly inferior in several critical aspects (processing speed, bandwidth and information storage density), and the only holdout--computational power/watt--exists because we have no good way to compare it.
Why would it be necessary or even helpful to COPY that biological implementation to achieve superhuman cognitive capabilities?
Guidelines contemplation for you.
Beginning of 2023: "GPT4 appears to demonstrate the ability to reason within certain constrained problem spaces."
There is no human being on this planet that actually understands how general intelligence fundamentally works. It's incredible to see people so confidently setting concrete timelines on its rate of development.
> As far as I'm aware they were still hallucinating pretty hard and incredibly biased and overly agreeable.
Certainly. But those aren't blockers to displaying the ability to reason within limited contexts. There are hundreds of thousands of human beings that fit that description.
The reason I'm bringing this up is that these issues, in humans, are solved by redundency (i.e. hiring someone else to look over their work).
As long as errors persist and, more importantly, are impossible/hard to be warned of, competent humans will need to oversee and validate every word they type, making their "advantage" much more nuanced.
It kind of turns into having a friend that can google pretty well answer your questions.
I'm not an expert, feel free to show me wrong - I'm reading the paper now and expecting it to be at least a touch troubling :)
I'm on my phone at the moment. But maybe someone else can link it.
"Competent" humans make errors too. This is a non-argument.