Only time will show if the next step will happen anyway. My gut feeling tells me that AI art will gain acceptance over time, and we will just think of it as "art" or "music", just as we did with recorded mysic and synthetic sounds.
I see little value to humanity in tools that are able to generate an endless amount of music derived from existing music, specifically designed to neatly slot into the place of human artists. We gain little in return from that.
Some people will make an argument like, this lets people generate lots of low-quality music for use in elevators or grocery stores. Well, there is already a massive oversupply of completely free music which can do that. Do people pretend to not know this?
The other weak argument is that it lets people express themselves who haven't studied or practiced music. But, it doesn't, because the interfaces (text prompts or "upload an existing file") are designed to take the place of a human being given instructions for criteria to fill, as if they were a worker, not an expression of the person giving the instructions. If the person giving the instructions were expressing themselves, most of the AI tool would not be redundant. It's as expressive as telling another person to write a song for you with some instructions. Hardly expressive at all.
For augmenting comoposers, sure, GenAI can be a tool like others. Musicians have been incorporating rhythms and melodies shipped with their electronic instruments for ages.
Entire genres have been defined by sounds and synth presets, too.
So I do see a bit of the similarities that you describe, but I think this is largely misleading.
Another side of it is that it will enable the creation of more music around more topics than before, by non-musicians. The accessibility bar is lower.
For a lot of people, music is a way to express their emotions, and not just by creating/playing it, but by listening to it. Now, you'll be make your own hyper-specific music with lyrics around topics specific to you, without learning any of the underlying skills yourself.
I've certainly wanted some kinds of music/representation in music of some of my experiences to exist, but not enough to go out and learn to make it myself. Now/soon I should be able to do that with AI tools, and I think that's actually neat!
Not a lot of individuals, to be honest. Only a handful of people make any kind of living from composing. Millions try it, but have to be content with performing for their friends. Which will continue unchanged.
So the actual economic impact of AI music will be different from the scenarios being described. What is true is, we will all have a lot more musical listening choices. Which is a net good for the rest of us?
Algorithmic music has already been around for decades and it never became popular. In the 90's it was of interest only to a small group of academic music nerds. The same is true today. Avant-garde shit for nerds. No one wants to listen to it.
> we got the synthesizer. Again we got the same complaints
I honestly don't believe this happened. Citations?
You make it sound like one day there was no recording and then bam! flac quality recordings of musicians, out of the blue. You do realize it started with exceedingly shitty wax cylinders that sounded absolutely atrocious by today standards (and by the past standards as well)
It is a test that selects for the ability to deceive.
1. We tell a human (test subject) that they will be a judge but they are a test subject. We will tell them that there will be two chats -- one will have a human and the other will have a computer and they need to decide which is which.
2. We will then give them access to two real time chats but the twist is both of them will be humans.
3. Our test subject needs to rebel against the experiment and say they are both humans.
What percentage of the population will be able to say both chats are humans? Is this a humane experiment? Will any ethics board clear it? Does it have any scientific value?
Maybe there will be a change of feeling, it's starting to come to me, instead of seeing this AI generated content as "soulless" etc I'm starting to see it as an extension of OUR human generated work. It's more like an endless remix of HUMAN talent.
All of that is boring though. The exciting stuff is all that will be displaced, and how we will solve the myth or meritocracy.
IMHO this and things like it are basically a sub-class of comedy or satire, so I have time for this sort of thing. It's a joke, and should (can?) only be appreciated as artistic as someone saying "Wouldn't it be funny if ... ?" because now that casual thought can be turned into a pretty instant "well here it is! LOL!".
I don't think you're supposed to appreciate it as music. Maybe I'm calling it wrong though.
(edit - I will admit that upon further thought, I'm not sure how I feel about this when compared to, for example, Nina Gordon recording "Straight Outta Compton" as an accoustic, mildly lamenting singer-songwriter style number 20-some years ago. It's clearly in the same satrical arena but one took a lot more effort and imagination. Kinda, because there was a lot of effort and imagination that went into both the training data and the model, even if this specific output was only a passing joke. It's quite hard to reason about this stuff...)
A lot of satirical songs are absolute bangers, because (for example) to do a send-up of the tropes of X music, you must know all the tropes of X music, and be able to perform them. So a lot of satirical music is actually done by people with a lot of skill and passion for the thing being satirised.
And because satirists don't have to worry about being predictable or unoriginal, they can put in more crowd-pleasing cliches per minute than 'serious' artists, giving them the most intense X of all X artists.
(Not saying the MIT license is a banger though - just that some satirical songs are)
I think it's a bit like everyone who said "black cabs in London have nothing to fear; X years of 'The Knowledge' will always be better than a guy with an App in an old Prius..." Except it was all wrong. Black cabs are probably still superior, but the bulk of people (and especially new customers) are all using Uber, because it's easier and they just don't care.
It sucks that production factory, literally built to make money, AI created junk music will be a thing, but it most likely will. Someone will exploit that they can make a ton of cash with low risk and budget. They'll have the connections to, despite initial (somewhat) faux outrage from the public/press, get radio play/playlisted online ("probably even 'ironic' addition from the hipster music crowd"). The song will be an earworm, and aside from the musicians that hear all the flaws, the passive every day mom and pop listeners will get the hook stuck in their head and it'll just be another great tune like any other.
I don't think we'll replace 'celebrity', I think that'll still happen, and I think maybe greater appreciation will happen to 'real musicians' (with a face!). But 'everyday' disposable pop (like all those one hit wonder tunes that are still played in nightclubs, pubs, throwback radio, wedding discos) - that's going to be disrupted massively.
I guess my point was that I don't hate the linked tune because it's not someone asking me to listen to a song, it's a joke and one that does have some humour to it.
I can absolutely see this tech (if it's allowed to) replacing a lot of working musicians who do music for ads, jingles, tv, film etc. And yes, disposable pop is probably on the chopping block.
> disposable pop ... that's going to be disrupted massively.
I wonder how the disruption will play out - the world is already drowning in content created by humans. If we envision that AI can make disposable pop to the same standard (and I have no reason to doubt it) then that surely creates an absolute deluge, almost boundless in size, of stuff. Promotion will become more or less the only art, to make things stand out from the crowd, and that can probably only continue for situations in which people want a shared experience. For an awful lot of situations the streaming of entirely ephemeral audio would probably do. It could as easily kill 'pop' as a business on the audio side, as it could steal it.
I'm just sorta daydreaming about possible outcomes here. All sorts could happen.
There's a craving for a deeper connection but that's usually the smaller part of our everyday consumption.
AI music will disrupt this market.
Spotify for example will try to produce their own AI music (like they already do with regular music). The Christmas playlist will then mainly contain their music. That saves a lot of money for the.
But have also had the non-surprising "Hedonistic Fatigue" that comes with excess access to something originally valued. I have now been able to generate 4 and 5 digit numbers of pictures of awesome colourful steam locomotives and epic dungeon vistas, but now find myself fatigued by "what on earth am I going to use 5000 dungeon pictures for?", coupled with the dread of being forced to CHOOSE from 5000 options. And I learn the known principle, that when you can choose from 4 options, you are happy you picked the best of 4, but when you can pick from 5000 options, you are left feeling inadequate with "I almost certainly was not able to pick the best of those 5000 options, and trying to do so would exhaust me". So suddenly, picking something from your menu of options, feels dreadful and fatiguing.. (I get the same feeling sometimes, when trying to pick a movie to watch out of 16.000 options).
So yeah, no doubt the AI sketch/refinement tool will be merged into our creative process, but for the time being, I feel a second generation of alienation-estrangement with my "available options".
Yeah, the last two live mega tours (Taylor Swift and Beyoncé) have a tad more personality than the average artist, but the usual stuff that you would hear on the radio might as well be AI generated and live-performed by animatronics and a significant chunk of the audience wouldn't care or even notice.
I mean people watch Love Island on TV for chrisssssakes. Humans can definitely be mindless consumers, where the sugar salt and fat in fast-food is the same as false drama, outrage, sex and violence in media. We all got buttons and they're so easy to push. Just look how popular TT is versus the type of content on there; mostly short-lived mindless stuff.
There will be a TON of these 60-second tracks, and no one wants to listen for anywhere near that long while skimming a feed, so these will now be ignored.
I wouldn't be surprised is all caps gets interpreted by it completely differently over multiple iterations.
Experiment: https://suno.com/song/3338f5e5-b36d-4596-815d-cd2804c9a344 (generated lyrics)
Same lyrics all caps (chorus): https://suno.com/song/b176c658-9e9c-4a4b-b05e-6db291f415c9 didn't really do anything.
After loads of different attempts trying to get something different: https://suno.com/song/398ef310-94b8-493a-ae61-22b790875689 not really that great.
Tbf I think it's because it's been trained on genres and maybe lyrics and lacks info on vocal styles and other stuff present in tracks during the training...
What I would like to see is what happens if you ask it to do in 5/8, 5/4 or any of the 7/* variants and see how that goes (I have no idea, might actually work). There are some examples in common music but not a lot especially for the more obscure ones. Unless they also trained it on a lot of international folk songs.
And yes it can do odd time signatures. It can do all sorts of genres like art pop, cinematic scores, ambient, math rock, avant garde, baroque, etc.
TBH that's a pretty low bar; "radio" songs have been engineered and polished for a very long time now. I have no hard numbers but gut feeling says radio music only represents 1% of the music industry.
I almost feel compelled to take action against the corruption and crookedness.