Today it's not just one industry - Western IP laws are slowing progress across multiple tech frontiers. While companies navigate complex IP restrictions (In EU and US), China's development is following a sharp exponential curve. You can already see it clearly in robotics, electric vehicles, and now these last weeks with AI.
While in the west you deal with predatory licensing (try talking with Siemens, Oracle, or Autodesk), and everyone keeps working on barriers and moats; other nations that allow a more collaborative approach (voluntary or not) are on an accelerating trajectory.
IP law is clearly no longer suitable for purpose - we need a system that encourages collaboration more directly. A complete free for all isn't ideal either - and I certainly don't advocate that- but even that appears to be better than what we have now.
Versus the content owners?!?
The way I think about IP is that if you grew up with something, by the time you're an adult it should be possible to remix it in any way you like, because it's part of your culture. Nobody should get to lock down an idea for their lifetime.
If your book hasn't made you rich in the first 14 years after publishing, I'm sorry, audience just isn't that into you.
(I am aware movie deals can languish for many years before finally landing a deal, and with 14 years studios could just wait you out and there's no incentive to write a script anymore, but copyright should be 14 years after publishing right? And movie scripts are not generally published before they get made into a movie, if ever.)
> Copyright in a work created on or after January 1, 1978, subsists from its creation ... [1]
And "creation" basically means "written down" [2]. IANAL.
[1] https://www.law.cornell.edu/uscode/text/17/302 [2] https://www.law.cornell.edu/wex/fixed_in_a_tangible_medium_o...
It won't ever change, though. No chance ever. Other than to be actually made infinite. The nature of "intellectual property" and money being speech in US politics locks that in.
This is a nice compromise in that it disincentivizes simply squatting on properties--you have to pay money to maintain later copyright so you have a strong push to make money on it.
GRRM isnt included on ASOIAF content because hes the copyright holder, but because his influence makes people want to see the shows
From the post:
"""
Our first recommendation is straightforward: shorten the copyright term. In the US, copyright is granted for 70 years after the author’s death. This is absurd. We can bring this in line with patents, which are granted for 20 years after filing. This should be more than enough time for authors of books, papers, music, art, and other creative works, to get fully compensated for their efforts (including longer-term projects such as movie adaptations).
"""
The new guidelines say that AI prompts currently don’t offer enough control to “make users of an AI system the authors of the output.”(AI systems themselves can’t hold copyrights.) That stands true whether the prompt is extremely simple or involves long strings of text and multiple iterations. “No matter how many times a prompt is revised and resubmitted, the final output reflects the user’s acceptance of the AI system’s interpretation, rather than authorship of the expression it contains,”
They are suppose to come out with guidance regarding the first question in a month or so.
Shining a torch at a plane is usually fine, shining a laser at them usually is a crime.
not sure I agree. a lot of work only get recognized broadly long after published.
Information/access to data/works should be totally free and there should be other ways to support the creators.
For example I could easily download MP3s of music and MP4s of series/movies but I don't: simply because of two reasons:
- I want to support the artist (to an extent as possible) - Using Spotify/Apple Music/Netflix is much more convenient with a totally acceptable monthly fee.
I know the article is not about entertainment but a library, same rules should apply.
And if one wants to train an LLM, let them: at its essence it's just a person who has read all the books (and access to information should be free), just the person is a machine instead of a biological human being.
If I gzip a couple hundred thousand books and distribute them freely, can I also claim it's just a person who has read those books and avoid a massive lawsuit?
Please, stop anthropomorphising machine learning models.
Netflix, Spotify, and Valve (Steam) didn’t succeed because of copyright enforcement. They won because they made paying for content easier, faster, and better than piracy.
Piracy isn’t hard, but these services solved the friction: instant access, high quality, fair pricing, and features that free alternatives couldn’t match. That’s why they still thrive today.
[1] https://www.escapistmagazine.com/valves-gabe-newell-says-pir...
It should also be de-criminalized.
At least I sometimes also get replies but many of them use fallacious arguments to the point of feeling like trolling. No idea if the same people commenting are also downvoting but I am starting to think that votes should not be anonymous.
There's absolutely no reason rich people owning ML companies should be getting richer by stealing ordinary people's work.
But practicality trumps morality. The west needs to beat China and China doesn't give a fuck about copyright or individual people's (intellectual) property.
The ML algos demand to be fed so we gotta sink to their level.
- Web pages; hard to argue that royalties are due since these are publicly available for free
- Scientific papers; these do cost money but the copyright is typically owned by scientific publishers
- Github, Stack Exchange, HN (yes); these are freely available, sometimes by license, so hard to argue for royalties
- Wikipedia, Project Gutenberg; these are also free by license
So the actual consequence of what you're proposing (or at least the realistically-enactable version of it) is the big AI firms paying scientific publishers a lot of money. Is this actually good? Is Elsevier, a basically pure rent-seeker, really more worthy than AI labs, which maybe you don't like but at least do something valuable?
If you're going to ignore the existence of copyright and licenses we should extend it to everything that's ever been posted on the internet, not just "web pages". Why shouldn't all books and films count as free too?
I'm actually open to the idea of just abolishing copyright but it's kind of silly to act like it's only about Elsevier. Lots of creatives depend on copyright in order to earn a living, similarly to how patents fund a lot of important research despite how noxious the patent system has become.
If we fixate on examples like Elsevier or Martin Shkreli in order to argue for completely abolishing the copyright or patent systems we risk destroying the framework that enables valuable creative works or new technologies to be developed in the first place. This is part of why people are so upset by AI companies arguing that they should just be able to ignore the whole framework in order to enrich themselves; once you allow the for-profit AI companies to do it, other groups are going to line up to also demand a free ride.
1) (minor nitpick) I don't see how HN is in the same category as GitHub or Stack Overflow.
2) "sometimes by license" or "free by license" imply you don't understand how copyright works. Code that is not accompanied by a license is proprietary. Period. [0] And if it has a license, then you have to follow it. If the license says that derivative works have to give credit and use the same license, then LLMs using that code for training and anything generated by those models is derivative and has to respect the license.
3) Arguably i didn't say this in the OP but the idea that the publisher owns copyright is absolutely nuts and only possible through extensive lobbying and basically extortion. It should be illegal. Copyright should always belong to the person doing the actual work. Fuck rent-seekers.
4) If western ML companies thought they can produce models of the same quality without for example stealing and laundering the entirety of Github, they would have. They don't so clearly GH is a very important input and its builders should either be compensated or the models and ML companies should only use the subset that they can without breaking the license.
Please don't get offended but I've seen this argument multiple times by proponents of A"I" and their entire argument hinged on the idea that the current ML models are so large that nobody can understand them and therefore a form of "intelligence" which they are clearly not (unless proven otherwise, at which point, they should get their own personhood but the fact no big ML company is arguing for that makes it obvious nobody really considers them intelligent).
[0]: https://opensource.stackexchange.com/questions/1720/what-can...