https://www.piratewires.com/p/compression-prompts-gpt-hidden...
57 karma · joined September 27, 2023
https://www.piratewires.com/p/compression-prompts-gpt-hidden...
In my mind the existential risks make regulation of large training runs worth it. Should distributed training runs become an issue we can figure out a way to inspect them, too.
To respond to the specific htpothetical, if that scenario happens it will presumably be by either a botnet, by a large group of wealthy hobbyists, or by a corporation or a nation state intent on circumventing the pause. Botnets have been dismantled before, and large groups of wealthy hoobyists tend to interested in self preservation (at least more so than individuals). Corporate and state actors defecting on international treaties can be penalized via standard mechanisms.
2. Accessibility of information makes a huge difference. Prior to 2020 people rarely stole Kias or catalytic converters. When knowledge of how to do this (and for catalytic converters, knowledge of their resale value) became available (i.e. trending on Tiktok), then thefts became frequent. The only barrier which disappeared from 2019 to 2021 was that the information became very easily accessible.
Your last two questions are not counterarguments, since AIs are already outperforming the median biology student, and obviously removing sites from the internet is not feasible. Easier to stop foundation model development than to censor the internet.
> What is to stop someone from training a model on such data anytime they want?
Present proposals are to limit GPU access and compute for training runs. Data centers are kind of like nuclear enrichment facilities in that they are hard to hide, require large numbers of dual-use components that are possible to regulate (centrifuges vs. GPUs), and they have large power requirements which make them show up on aerial imaging.
> Essentially you are advocating against information being more efficiently available.
Yes. Some kinds of information should be kept obscure, even if it is theoretically possible for an intelligent individual with access to the world's scientific literature to rediscover them. The really obvious case for this is in regards to the proliferation of WMDs.
For nuclear weapons information is not the barrier to manufacture: we can regulate and track uranium, and enrichment is thought to require industrial scale processes. But the precursors for biological weapons are unregulated and widely available, so we need to gatekeep the relevant skills and knowledge.
I'm sure you will agree with me that if access information on how to make a WMD becomes even a few order of magnitudes as accessible as information on how to steal a Kia or how to steal a catalytic converter, then we will have lost.
My argument is that a truly intelligent AI without safeguards or ethics would make bioweapons accessible to the public, and we would be fucked.
More banally, state actors can already use open source models to efficiently create misinformation. It took what, 60,000 votes to swing the US election in 2016? Imagine what astroturfing can be done with 100x the labor thanks to LLMs.
[1] dx.doi.org/10.1038/s42256-022-00465-9
Several of my coworkers had been dating for 5+ years, but they were only making $50k annually (early career engineers). The socially expected family-sized condo costs $500k to $1M, and the young couple is expected to buy and furnish it before their wedding.
Yes. DMCA claims are made under penalty of perjury.
> Being prevented from publishing a video someone else created, say, doesn't inhibit your right to express any opinion (aka freedom of speech).
Indeed, the inhibition of free speech is when you are prevented from publishing a video you created, because someone else falsely claims that they created it. Hence the felony.
More recently I visited a much nicer midwestern town which is planning an expansion of its bus network amd optimizing traffic lights for better bus flow. There is a new mall on the bus line. The difference is that 95% of bus riders in this second town are upper middle class college students.
The rationale is that the DMCA gives complainants the ability to restrict others' speech, and so the law wants to strongly disincentivize abuse of that power.
The relevant analytical unit at the small scale is the family: I don't want my kids to be temperant because of stability, I want them to abstain from drugs/games/$VICE because that's the path which maximizes the chance of their living a fulfilling life, or (more cynically) which maximizes their chance of bearing me successful grandkids and great grandkids. This is why puritainism is selected for evolutionarily (at least in environments where resources are limited).
To return to the large scale policy questions, I also don't want to see the continent of my children fall to a mercantilist China (using China as an example because Chinese law cracks down hard on drug sales and limits students to one hour of video games per night). Accordingly, I support policies to limit access to addictive substances and stimuli, despite the inevitable conflict between those laws and individual rights. The inequitable enforcement of those laws is another problem entirely, and one which I think would be well solved by starting with the prosecution of celebrities and thought leaders who openly partake in $ADDICTIVE_STIMULUS, and their suppliers.
More recently in the US I got a call where the other side didn't say anything and hung up after exactly one minute. Suspicious indeed.
More generally, the challenges in communication are a well known barrier to outsourcing knowledge work.