The success of the LLM were built by the work product of others, without compensation. Now those others are looking for their due compensation.
Freely profiting off the, quite literally, compressed output of others isn't really a business model that's sustainable, for either side. The only sustainable solution would necessarily involves money going from the content users (multi B $ LLM companies) to the content producers (artists, news orgs, etc). For a logical litmus test, apply what's happening to any other content/industry.
I just hope you are consistent in that you oppose the existing ubiquitous copyright violations done by the entire art industry, in the form of commercial "fan art", or similar.
A whole lot of content is built off of other people's works, and much of it is not done "in fair use", but of course, then the shoe is on the other foot, those same creators complain.
If OpenAI loses value as a company if it does not have that content, that indicates the content has value, and OpenAI should pay for the materials they 'use' to create their own value.
This line of reasoning seems more or less equivalent to "movie piracy will exist one way or another, therefore I should be allowed to launch a commercial Netflix competitor which streams movies without paying the studios". Just because it's possible to appropriate content doesn't mean we should just give up and put it in the public domain, that's obviously not sustainable.
Do you feel that entities like The Financial Times should just willingly or be forced to give up all their data to the public?
Maybe we should go back to 14 years for use in AI models?
Ads are not enough, and readers are fed up with ads, so they use adblockers, cutting publications' revenues. With AI chatbots, people browse these publications even less, further reducing revenues.
Paywall and licensing content is the next best option, if not what it is?
content creators are going direct to consumer with their content, and there is endless opportunity for actual content creators to thrive without publishing houses as middlemen
If LLMs are the core of 'AI' then I'm not interested, but I feel confident they're a sideshow on the way to stronger AI.
Lets avoid the huge mistake that was made with the internet, where the web happened to be the first broadly useful tool which exploded in popularity and then de facto became the internet, with everything shoehorned into it. The problem is, on the scale of possible interfaces and software, the web is absolute shite, but now its gravity is too great and we can't escape it.
I've tried to use LLMs but they're just not useful. Getting answers that may be great or may be nonsense just doesn't work for me. I know there is something better and I won't be distracted by a very impressive novelty.
I think it'll be a while before we perfect memory, neuroplasticity, the physical experimentation feedback loop, etc. but LLMs at least give us a way to represent and manipulate human language.
I agree with that, but plenty of people are talking about LLMs as the precursor to AGI, especially the OpenAI crowd. Plenty of people here and elsewhere say that LLMs are actually intelligent.
I think LLMs are an important illustration of what's possible, but it's not the one true way.