AI Is a Terrifying Purveyor of Bullshit. Next Up: Fake Science
lastwordonnothing.com
lastwordonnothing.com
"The researchers showed that GTP-4 ADA “created a seemingly authentic database,” that supported the conclusion that one eye treatment was superior to another. In other words, it had created a fake dataset to support a preordained conclusion. This experiment raises the threat that large language models like GTP-4 ADA could be used to “fabricate data sets specifically designed to quickly produce false scientific evidence.” Which means that AI could be used to produce fake data to support whatever conclusion or product you want to promote."
Note also that faking data is not only about science. I'm pretty sure we're going to see fake datasets for fake market analyses, fake polling data for fake political campaigns, fake data from fake data breaches, etc., etc.
How does society function when nobody can know if anything they think they know is actually the truth? And not just in the "deep epistemological thoughts" way. Even in the way that we used to be able to know things, with less than total epistemic certainty, but still with sufficient certainty to usefully be able to use information... we can't do even that any more.
You would get a huge amount of digital distrust, and some percentage of people would retreat to "If it wasn't printed in a book before 2020 ..." as a standard. Wild paranoia would emerge and others would retreat into a kind of indifference because "Who knows?" Previous conspiracy theories, suck as a faked moon landing, might become plausible just due to how many other wild fakes existed. How would you know who really won an election?
“Being in a minority, even in a minority of one, did not make you mad. There was truth and there was untruth, and if you clung to the truth even against the whole world, you were not mad.” ― George Orwell, 1984
It doesn't.
There is an interesting dichotomy here. This case is indeed not ideal, but many developers would be happy if an AI could create the fake data for testing that is representative of real data.
It's not the tool, but how it is used, or by whom it is used?
I agree: synthetic data has many well-intentional use cases. It may also remedy privacy issues, among other things. And you can easily enough generate synthetic data without AI, but the issue, I think, is that AI makes the generation easy also for laypeople, and I suppose you can prompt LLMs to generate custom "high-quality" fake data that no data sleuthing can detect.
We're seeing something proposed in the US around banning contributions to RiscV because it is open source and "our enemies" might benefit and avoid tech sanctions. This seems like another bad take to further isolationism in the US
It seems easy to create a list of everything bad a human can do and then write a corresponding article about how bad it is that an AI can do it too.
The same reason it's cool that AI can make a recipe for a cake or draw a picture of a boat is the same reason it's bad when it can do bad things. It's not as if people were not capable of making recipes for cakes or drawing pictures of boats. It's just now _anyone_ can rapidly do those thing on demand +/- some efficacy determined by the sophistication of the AI.
> And I could. More than anyone else on Earth, he said, because of my expertise. That knowledge I had gained in defiance of the dark could finally be put to use. I was to create a focus, a black star, a new centerpoint around which a universe of purest darkness could turn. To take dark matter, dark energy, and harness it, bring it forward into a form that could be held, used… worshipped.
> Scientifically, it was nonsense, of course. Dark energy and the like don’t work like that, not even remotely. But that wasn’t important. What mattered was that it felt like science, and that was all I needed; to do my work, to create the black star, would need a parody, an aping mockery of science. But it would also need the deepest of darknesses.
I think it can only be a good thing if AI breaks all our bullshit detectors and forces us to actually establish roots of trust and make science independently verifiable.
How would that even be possible in a world where literally nothing can be trusted?
The point on hallucinations is particular on point: when a human makes up stuff it's lying, when a machine does it, it's marketed as hallucinations.
welcome to the party, pal