The Financial Times and OpenAI strike content licensing deal
ft.com
ft.com
Ads are not enough, and readers are fed up with ads, so they use adblockers, cutting publications' revenues. With AI chatbots, people browse these publications even less, further reducing revenues.
Paywall and licensing content is the next best option, if not what it is?
content creators are going direct to consumer with their content, and there is endless opportunity for actual content creators to thrive without publishing houses as middlemen
This line of reasoning seems more or less equivalent to "movie piracy will exist one way or another, therefore I should be allowed to launch a commercial Netflix competitor which streams movies without paying the studios". Just because it's possible to appropriate content doesn't mean we should just give up and put it in the public domain, that's obviously not sustainable.
The success of the LLM were built by the work product of others, without compensation. Now those others are looking for their due compensation.
Freely profiting off the, quite literally, compressed output of others isn't really a business model that's sustainable, for either side. The only sustainable solution would necessarily involves money going from the content users (multi B $ LLM companies) to the content producers (artists, news orgs, etc). For a logical litmus test, apply what's happening to any other content/industry.
If OpenAI loses value as a company if it does not have that content, that indicates the content has value, and OpenAI should pay for the materials they 'use' to create their own value.
I just hope you are consistent in that you oppose the existing ubiquitous copyright violations done by the entire art industry, in the form of commercial "fan art", or similar.
A whole lot of content is built off of other people's works, and much of it is not done "in fair use", but of course, then the shoe is on the other foot, those same creators complain.
Do you feel that entities like The Financial Times should just willingly or be forced to give up all their data to the public?
Maybe we should go back to 14 years for use in AI models?
If LLMs are the core of 'AI' then I'm not interested, but I feel confident they're a sideshow on the way to stronger AI.
Lets avoid the huge mistake that was made with the internet, where the web happened to be the first broadly useful tool which exploded in popularity and then de facto became the internet, with everything shoehorned into it. The problem is, on the scale of possible interfaces and software, the web is absolute shite, but now its gravity is too great and we can't escape it.
I've tried to use LLMs but they're just not useful. Getting answers that may be great or may be nonsense just doesn't work for me. I know there is something better and I won't be distracted by a very impressive novelty.
I think it'll be a while before we perfect memory, neuroplasticity, the physical experimentation feedback loop, etc. but LLMs at least give us a way to represent and manipulate human language.
I agree with that, but plenty of people are talking about LLMs as the precursor to AGI, especially the OpenAI crowd. Plenty of people here and elsewhere say that LLMs are actually intelligent.
I think LLMs are an important illustration of what's possible, but it's not the one true way.
What is the alternative path that is as lucrative as taking that big 4$$ check from OpenAI/Microsoft? I mean the only real alternative is that Google, Facebook, or Amazon write you a bigger check. But you're still at the same place because they'd only write that check to train their models as well.
The industry is backed into a corner. If no one else will pay, they have to take money from the only person who will.
I can imagine LLMs tipping the balance back in favor of something like traditional media. In short: people go nuts with them, and flood the open web with garbage (which may have the added benefit of degrading LLMs). That could, compared to the last 20 years, make curation and verified provenance far more important for people who value having the chance to know real things, which would drive customers back to pay-walled media organizations and publishers.
That may not happen, or may not happen on a time-frame compatible with most current media organizations. It would certainly take some time for everyday users to develop the necessary fatigue with LLMs and their output to motivate action.
I think it'd also depend on the LLM companies losing in court on copyright grounds. If the LLM companies win, I think we'll still get the crapflood, but the islands of sanity resisting the tide will be smaller or nonexistent. If LLMs put the final nail in the coffin of the media orgs, Wikipedia will die soon after (it's got separate culture problems that may do it in, but it is also totally dependent on the editorial decisions of traditional media and publishing).
Each individual publisher isn't needed. There is lots of training data in the world. So you can either get a payday or get nothing.
I also can't believe that these media companies haven't learned the lesson they shouldn't have learned by looking what happened with Google and saying maybe we shouldn't give that away, at any cost. Like what's their bargaining position after the training has been done?
That's why they're taking the money that OpenAI/Microsoft is offering.
They tried to do it in court with google and never got anywhere with it. So this time around they just take the money from whoever is willing to write the biggest check.
https://www.bloomberg.com/news/articles/2023-12-13/openai-ax...
So yes, making agreements to license content does illustrate that there exists a market for using text in AI training, and it will do a lot of damage to arguments that it's all fair use.
We’re bringing the Financial Times’ world-class journalism to ChatGPT
https://openai.com/blog/content-partnership-with-financial-t...
1. OpenAI is trying (and failing?) to add legal restrictions on LLM (and co) training to only a few major players.
2. By entering into licensing agreements, it starts to set the standard and essentially forces the above to take place as only a few major players would be financially able to afford it.
> In addition, the FT became a customer of ChatGPT Enterprise earlier this year, purchasing access for all FT employees to ensure its teams are well-versed in the technology and can benefit from the creativity and productivity gains made possible by OpenAI’s tools.
i.e. ft is going to be written by llms now.