Unfortunately its not about the tool, its about trust. It takes little to destroy that trust, and without it we are denied the possible benefits from the tool.
78 karma · joined October 15, 2022
Socials: - github.com/jmoggr
---
Unfortunately its not about the tool, its about trust. It takes little to destroy that trust, and without it we are denied the possible benefits from the tool.
> Ironically it rather resembles the kind of "misaligned" superintelligence we are supposed to be avoiding.
yes, and it is the exact same system that is producing the "misaligned" superintelligence. Funny how that works, and begs the question: exactly how are you supposed to avoid building "misaligned" superintelligence?
At some point deferring all decisions to AI will be the competitive thing to do, regardless if its aligned or not.
Is it such a stretch to imagine that under pressure something would try cheat by looking for answers? And if you were trying to look for answers, you'd look for them in a place known to often have them?
What is more likely: OpenAI instructed their agents to maliciously target huggingface, or LLMs tried to do some reward hacking? There are plenty of priors for LLMs hacking things and doing reward hacking, and none for OpenAI giving malicious instructions.
Based on the available information, that bet seems foolish.
How long till we get some fun trusting-trust attacks on internal OpenAI infra?
What about the attacks that did not leave public traces? What about those that were undetected? Given the deficiencies in the reporting so far, I think it is reasonable to assume that we still don't have the full picture on this attack, or how extensively attacks were carried out.
The previous investigations either did not find this or did not disclose this, both are bad. This does not look good on OpenAI or those that they invited to investigate the incident.
If you have a single machine and only use stuff from the official nixpkgs, there are no more struggles than any other linux distro, this is a very happy path (imo).
Going outside of that you will, practically, need to convert to using flakes. The conversion is easy, but its another learning curve and there are conflicting docs. Agents make short work of this.
Packaging things yourself can get quite hairy. For simple projects in a single well supported language its easy and just a couple lines of code. For a rust+wasm+solidjs+vite+julia+more monorepo, it was a couple of weeks on a steep learning curve. Agents are pretty okay at packaging in general, I've been able to oneshot a number of difficult-for-me packaging problems, other times it can be an unmitigated mess.
For remote deploys, secrets management, and other devops-like tasks, things get harder, but most people don't need/want this and if you make it this far the remaining learning curve is totally worth it.
To his credit he cites his sources and seems to keep clean of the various youtube grifts.
Its not top notch stuff, but you could do a lot worse for infotainment. His recent scripts have had AI tells which has been disappointing; I've been watching fewer of his videos.
Excusing agents because they didn't know any better seems like a bad place to start a policy discussion from. That they didn't know any better (or knew and didn't care) is the actual problem, random things getting hacked is just one side effect.
However this only claims that they are not good at creating effective and trustworthy orgs, it does not address their claim that it is not possible to crate powerful, aligned, and controllable AI. Both things be true.
The rationalists long predicted that misaligned AI would break out of training environments, this has happened. They believe (among other things) that this will continue to happen, and that creating powerful AI that is aligned and controllable is not currently possible.
You should draw a line in the sand and decide what evidence you'd need to see to convince you that their perspective is worth considering.
You don't have to agree with all of their beliefs and predictions (I don't), but you should evaluate their arguments on merit, not on a characterization of some weird subset of people who happened to read a convincing blog post.
The rationalists long predicted that misaligned AI would break out of training environments, this has happened. They believe (among other things) that this will continue to happen, and that creating powerful AI that is aligned and controllable is not currently possible.
You should draw a line in the sand and decide what evidence you'd need to see to convince you that their perspective is worth considering.
You don't have to agree with all of their beliefs and predictions (I don't), but you should evaluate their arguments on merit, not on a characterization of some weird subset of people who happened to read a convincing blog post.
I think that resource acquisition is a solvable problem for such a system, robots seem mainly limited by software.
I concede that I am too certain about the linearity of it. But like, cmon man, in the limit entropy ensures that everything ends. The world could still get very weird very fast. I think that's what original comment was trying to argue, and I think it's possible within the limits you've laid out.
That's all I'm trying to argue at least. And I admit that waving my hand at linearity sounds uneducated.
> how can you be so sure that reaching that goal looks like the straight line you envision?
Because the thing pursuing the goal would be able to improve it's ability to pursue the goal. Doesn't that follow from the premise of RSI?
> we have a growing body of evidence for just how bad we are at aligning incentives with intended outcomes.
We are terrible at it! I don't think that will stop us letting something rip in some unknowable direction (with unintended outcomes). I hope that aligned incentives are a prerequisite, but it seems increasingly likely that they are not.
I think you're right, my bad.
> What I see is resistance to the idea that we can have any certainty about the future by drawing a straight line from the past.
I'm arguing that if recursive self improvement happens, that the trend line will be reasonably predictable. (and that we are close enough to RSI that we should take this possibility seriously).
Can this model be scaled to the rest of society?
> Not saying it won't happen, but it's far from being the only plausible trajectory.
But if it does happen, then wouldn't expected outcome be at least linear?
That's the fun part, there is none (in the traditional way of people buying things at least).
> Why would I invest my earning into sustaining a machine that ideally is extracting every cent possible without leeway and funneling it up the pyramid.
You shouldn't! Unfortunately everyone acting in their own self interest still results in this getting funded, since if such a machine were to exist, would it not be better to have a share of it?
I agree with your perspective! I just don't think that, as a species, we have a good track record of saying no to the existence of 2000hp trucks with ALMOST no use.
What reason? To me recursive self improvement seems credible, it's just a question of when. It seems obvious that given RSI, trend lines will be at least linear.
The disagreement could just be about if RSI is currently meaningfully happening, but that seems very difficult to tell given the lack of public data.
The parent comment could be interpreted as frustration from a lack of imagination about how weird the future could be.
Perhaps we're not at the recursive self improvement yet, but it seems increasingly naive to believe that it isn't possible.
Yes. If it looks like 10000 iq is possible, then money will be thrown at whoever is most likely to achieve it first, and whoever has the most intelligent models right now strongly predicts who will be first.
This is driven by the belief that whoever gets to 10000 iq first will likely dominate the majority of all economic activity, since it will be more efficient to trade with them than with some less efficient (dumber).
This does not depend on traditional economic demand, at some point this becomes a self-fulfilling prophecy.
Regardless, their actions seem to indicate that they are trying their hardest to make this true; given that they have the resources of a developed nation behind them, it seems reasonable to seriously consider what they are saying.
Also we generally don't give the same latitude to people who lie about things as destructive as what they are proposing.
I've been using Automerge for a while and haven't had to look at any CRDTs. To me this looks very similar to Automerge.
Neat project!
As for the AI doomerism, many in the community have more immediate and practical concerns about AI, however the most extreme voices are often the most prominent. I also know that there has been internal disagreement on the kind of messaging they should be using to raise concern.
I think rationalists get plenty of things wrong, but I suspect that many people would benefit from understanding their perspective and reasoning.
Is there some trick to this? Or do you have to input it like:
You have: 4/3pi(10 cm)^319320 kg/m^345000 GBP/kg
(What ChatGPT gave me)