This legal complaint alleges that defendants operating a non-profit entity for the benefit of humanity have committed massive fraud on donors, beneficiaries, and the public. The complaint raises concerns about OpenAI's operation, including its dual structure as a non-profit and a for-profit entity, potential insider dealings, and the exclusion of the general public from its benefits. It claims that OpenAI has used deceptive advertising, unfair competition, and fraud to develop its valuable resource for personal gain.
The complaint highlights OpenAI's mission of benefiting humanity and points out that a narrow group of stakeholders have received commercially invaluable early access to its technology. It also argues that OpenAI's for-profit operations might infringe on copyright and fair use laws, as the technology is built on large datasets, much of which is copyrighted. It accuses OpenAI of breaching trust and fiduciary duties, disrupting legal frameworks, and potentially engaging in willful and wanton negligence by increasing existential risks related to AI.
Finally, the complaint alleges that OpenAI might have engaged in banned political activities, specifically suggesting that the technology may have been used to influence the 2020 US presidential election in favor of the Democratic party.
I've never found it particularly useful for most articles which are easy enough to read/skim (the first and last 2 paragraphs will usually tell you what you need), but long complicated legal documents are a whole other matter. This is great.
My impression was that hallucination happened when it simply didn't have facts in the first place, had conflicting facts, etc.
I thought summarization was generally fairly reliable, but I'd be happy to know if this is not the case.
The failure foolishly and misleadingly called “hallucination” is only one manifestation of an attribution error. If your summarizer leaves out something very important because it doesn’t understand it the result will be quite misleading.
For your average web text which these days is 90% filler and not important anyway, this is no big deal. This particular lawsuit appears the same. But for anything important, I wouldn’t trust it.
On the other hand, if you're in a field where there's an adversarial use of text and the uncomprehended 20% might be used to nullify, contradict or make loopholes in the main body, then relying on ChatGPT is similar to using Tesla Full Self-Driving in a construction zone, near firetrucks, during a snowstorm.
The only challenge is chunking the larger bills and synthesizing the larger summary without losing out on possible nuances. Something like California's SB423, for example, is over twice the 8K token limit and that's not even a large bill.
Unfortunately, things like the US Code or Code of Federal Regulations are in the range of 100s of millions of tokens.
* "Altman and the other parties to this suit are increasing the risk of global human extinction or actual world domination by a small set of individuals for a chance to personally gain extended lifespans. It is the reasonable explanation for taking such massive risks with this technology and flaunting the law so obviously."
* "OpenAI and at least one of its partners most likely filled social media like Twitter, Facebook, and Reddit with politically charged commentary designed to push votes towards the Democrat party." (And no, the filing doesn't provide any substantial evidence for this assertion.)
* "Y Combinator is, according to ChatGPT, the most notable Tech Accelerator in the world. A screenshot of ChatGPT-3.5 stating this is included as Exhibit L."
* "Open-source and closed-source, not-for-profit and for-profit, are binary choices, or Booleans. Booleans are a form of data with only two possible values, which are typically opposites. When defendants drastically changed the Boolean values that structure [OpenAI]... the founding mission went from ‘true’ to ‘false.’"
Sure, but the kooky claims are mostly tangential to the identified causes of action except the fourth, and they aren’t the sole basis for that one, so they have marginal bearing on the overall suit.
6.4MB
Gzipping the 52 MB PDF just shrinks it to 51 MB. You got it to 6.4 and kept it PDF.
gs -sDEVICE=pdfwrite -dCompatibilityLevel=1.4 -dPDFSETTINGS=/screen -dNOPAUSE -dQUIET -dBATCH -sOutputFile=output.pdf input.pdf
Edit: I would mirror it on one of my sites, but all my sites are pay-per-gb for bandwidth.
Because it has 280 pages of newspaper articles and white papers about OpenAI, ChatGPT, and other startups attached to it as exhibits, many of which are only tangentially related to the case.
PDF embedding is funny.