Intro0:00
Good morning, everyone. Good morning. Um, I'm excited to be here. I— I really have one main goal in mind, and that's for this talk to be— or this workshop— to be interactive. I want it to be fun, and I want you guys to walk away with something tangible.
Um, so I— I— we can actually switch to the slides. Uh,
slide show. It should be live now. If you guys want to go ahead and go to this link that I have up on the screen, or use your phone to go to the QR code, I have some notes for you guys at that URL.
Um, but— but those are my stated goals for this workshop. I want to really probably start by, uh, asking you guys some questions so I can get some interactions going. But I'm glad you guys came. Thank you for— for coming.
I really do appreciate it. Um, the music stuff is something I have been into for a long time, like lifelong musician, artist. Uh, but the AI music stuff is something I just started paying attention to, um, with— with regard to the tools I'm going to show you today, maybe about a month or so ago.
So basically, what this workshop is going to be about is I'm going to start with the insights that I've gained by, like, doing all the research over the last 30 days, and then we'll get straight into the tools so you guys can, again, walk away with something tangible.
Um, let's start with— okay, by— by— I know it's early and, and, uh, I'm still groggy myself, so, uh, you know, you don't have to necessarily go crazier here. But by show of hands, uh, how many of you have ever used a music generation model, period?
Okay, we got— okay, nice. About half? Half of you guys? That's dope. Um, by show of hands, how many of you guys use— I had someone in the crowd already say that they used Boomi. Have you guys ever used Boomi before?
No? No? No Boomi? Okay, that's— you're OG if you used Boomi, because it was from, like, 2 years ago. But it was a music generation model that, um, was actually pretty cool, but— but didn't get as complex as the— the music generation models we have today.
How many of you guys have— have— how many of you guys have ever used Suno or Udio, by show of hands? Okay, nice. Okay, so you guys will be familiar with— with what I'm going to do today. Uh, hopefully I can show you guys some, like, advanced prompting techniques, at the least.
But for the most part, that's— those are the tools that we're going to, uh, be focused on. Um, I was doing this earlier, asking some of the AV crew what you guys' favorite artists are. My man here—right—right here in the middle, if you could name, uh, your favorite artist of all time, what— what would it be?
Who would it be?
Me.
Yes.
Miles Davis.
Miles Davis, okay. And correct me if I'm wrong, because that's a little bit out of my league. Like, blues, R&B, jazz? Okay, okay. I— I think I have something special for you today, then. So, uh, let's— let's start.
I can get into, uh, a little bit of what we're going to do as well. And I— I actually would like to say, because I've already talked about my workshop goal, what I would actually like to say is that, um, I am not a classically trained musician.
I— I wrote my song— my first song when I was 13. But I welcome any discourse you guys have in terms of, like, uh, accuracy,right? So if you guys feel like, uh, you're actually wrong, like, that's cool. I— I welcome that.
And, uh, I actually have, uh, something for you on that link that I showed earlier, where you can submit questions or comments live, like, live while we're doing this, um, to either correct me or— or submit a question or whatever.
So that— that was one of the main things I wanted to say, is, like, I'm— I— I try and be as accurate as possible, but if I say something wrong, you know, feel free to call me out. I— I'm— I'm comfortable with that.
I'm cool with that. Um, and actually, let me make sure I have my stuff pulled up. If you guys can go to this link, there's a link on there that says "Poll." If you guys can go, uh, fill out that poll.
I think there's like 2 or 3 questions on there. Uh, if you guys could go to that link now and start submitting your answers, I'd really, really appreciate that, because it— it— I want— again, I want this to be really interactive, but I also, uh, want to tailor it to, uh, specifically what you guys want to know about, uh, AI music generation.
So if you guys could make your way over to that link and s— and, uh, submit, I would really, really appreciate— uh, appreciate it. And now that I've done some housekeeping, I can kind of get into, um, my journey here.
I would say I'm way more of a— my— first of all, my name's Phlo, as you guys may or may not know. Um, I'm way more of a songwriter than, like, a software engineer. Uh, but I have a lot of respect for what you guys do.
Background4:46
I— I— I know enough Python to be dangerous, but beyond that, I'm not really that good at, uh, software design or— or engineering. Um, but I— I, over the pandemic, uh, took a lot of time out to learn how to write scripts and aut— um, I'm really, really big into automation, so I use Python for automation for a lot of my stuff, da— daily, personal life stuff, uh, and— and, uh, stuff for work as well.
Um, the other thing I would say is I'm really, really big into experimentation. Um, Phlo is actually my artist name, and if you go on Google or YouTube or Spotify, Apple Music, you can actually look me up and find my music.
Uh, but the truth is that, like, I have a very eclectic taste, and I, like, have, uh, experimented in— in a lot of different genres. So I, um,
a couple years ago, I decided that I wanted to do some split testing, which is pretty odd for an artist. Uh, but I— I did want to do some split testing. And what— what I really found out was that the same song under 2 different artists can behave wildly diff— or can perform wildly differently.
So I literally would take the same recording, like I'm a self-taught engineer, would record at my home or other studios, and upload it under 2 different artist names. And depending on the cover art, the marketing, uh, who you try and target when you're— when you're posting online in terms of the daily content, I would see that songs would perform differently.
Um, so what I basically did was start, uh, creating these quote-unquote "pin names" and putting out music that way. And I actually have a pin name, uh, that has way more streams than my actual artist name. So if you look up me as an artist, I don't have that many streams, but I have a couple pin names that, uh, have been, like, viral.
I have one that's got, like, 5 million streams over the last couple months. Uh, and it's just instrumental music. So I— I'm an artist that does, like, R&B and hip-hop, but I've experimented with, like, EDM. I actually have a song with some, like, country influence, believe it or not.
Uh, so I've just experimented a lot, um, and— and that's kind of how we got here today. I was— I like to say I was experimenting with, uh, multiple genres way before AI music got here, but I— I love AI music for— for that reason.
Um, and this workshop will cover, uh, AI music generation and demystifying it. I kind of like distilling better, um, but, you know, ChatGPT gave me this word, so I put it in the slides. Um, and we're going to get hands-on with Suno and Udio, uh, and talk about turning ideas into songs.
And— and then, uh, actually, I— I have a question about this later on, but, um, I really want to talk about opportunities for using AI music. I'm super passionate about, like, distribution, um, and— and for, like, regular, everyday artists to be able to make, um, an income from— from their music.
So if any of you guys are actually artists, like, I would love to get into, uh, those type of, uh, discussions. Um, let's see. Have you guys— did anybody actually fill out the thing that I was asking about?
Let's see.
Oh, I'm disconnected. Break, uh, leg. Because I really can start throwing in little tidbits here and there about, um, what you guys want to learn about. I put one— like, one of the options that I put for the poll is, like, an advanced option where we can, like, actually pull down some of the music that Suno is generating, uh, get stems for that— that music, and start to, like, mix and master.
Um, so I don't know if you— if— if any of you guys would actually be interested in that, but I put that as.
That'd be great.
Yeah? Okay, awesome. Yeah, that— that was the hardest thing about, uh, preparing for this, is, like, I just don't know exactly who's going to be in the crowd and what— like, what level they're at. So I tried to, like, uh, come up with a couple different tracks just for this— this, uh, presentation alone, but, um, yeah, I just needed you guys' feedback.
So, oh, nice. We did get a bunch of submissions. And it looks like, um, prompt techniques. Okay, tweaking. A lot of prompt techniques. Okay, cool. I have this guy that I call, like, the prompt master. How— how many of you guys— by show of hands, how many of you guys are in the Suno or Udio Discord server?
Nobody. Oh, yeah. Oh, you're in there? Okay, nice. Yeah, even for you, I think I'll have a treat, because, like, uh, there's a couple people in there that have posted, uh, since the beginning of, uh, Suno and Udio.
And then there's this one guy who's, like, been able to do some really, really cool, like, banjo music and bagpipe music. That's insane. But hit the prompts are very unintuitive, in— in my opinion. So you kind of have to see it to, you know, understand what it's doing.
Um, so we'll definitely get into the— I'll— I'll make sure that I include a lot of prompting stuff. Thank you for everyone that just came in. I appreciate you guys joining.
Um,right, so this actually gets into.
You might notice that. You might notice that.
Okay. It did actually work, but I— I still might play it from my computer. So
this is where we start with the insights that I've learned by studying AI music and how it got to this point over the last, uh, couple months. And then I'll get into it— it— it's, uh, kind of painful, but I have to play some videos for you— for you guys to catch up to where I'm at in terms of, like, understanding and how this whole thing developed.
Music Types9:35
Um, so we'll have some videos to watch, but maybe 5 or 6 minutes of— of— of videos. Um,
when people say "AI music," I think that can mean a lot of different things. And when I first started doing my research, uh, it was really frustrating that someone— or everyone— everyone meant something slightly different. So what I've tried to do is compartmentalize the definitions of AI music and tidy up, like, like, I guess really tidy up the definitions, because there was a lot of, like, blurred lines when I started doing the research.
Um, and— and the way that I— I, uh, compartmentalized this was to say that we have, uh, text-to-music, uh, where you're generating short— short samples, um, and full songs using text— text prompts. You have, um, audio-to-music, and this is where you're taking an existing sound and transforming it into, into, uh, actual music.
And then you have, like, style transfer, uh, which is basically voice conversion. Um, and voice conversion is actually the first thing that came across, uh, when I found AI music, uh, last year, actually. So it's been, like, a long time coming, but I didn't really start doing the research until recently.
But I think what most people call AI music is really actually just voice conversion. What— what— it's hard to say that now, but maybe when I started doing research, it was. But— but I would say that, um, the— the general consumer, what they call AI music is— is mostly just style transfer, where someone, like, took their voice and converted it to another voice.
And that actually gets into the first video, and I'll go ahead and play it now. You guys may have seen this video. It's actually pretty cool. Text-to-music.
Oh.
You might notice that I sound like Kanye West. No.
Could we, uh, turn it up just a little bit in the house? Because it seems a little bit low. He gets a little louder later, but you'll— you'll catch it.
You might notice that I sound like Kanye West. No, Yeezy didn't record a voiceover for me for this video. I didn't learn how to do impressions. This is AI.
So let me come back to my original voice for a second, because this is crazy. Today on Metaverse, I posted AI Kanye covering popular songs. Here's an example of him singing "Day and Night."
Day and Night. Whoa. I toss and turn. I keep stressing my mind.
So that's clearly crazy. And I started thinking, you know, what are the implications of this for the music industry? Now all you have to do is record reference vocals and replace it with a trained model of any musician you like, which is exactly what I did.
I found this Kanye-style beat on YouTube. I wrote 8 bars, and I'm going to record them now, and then I'm going to have AI Kanye replace me. I got a fantasy that's beautiful, that's dark and twisted, but I attacked a whole religion all because of my ignorance.
What was I thinking? That was some bitch shit. I lost Adidas, but I'm still Yeezy. Back in the kitchen. Man, I'm a genius. Boys in the hood, just like I'm Eazy. Kanye Wheezy. Southside of Chicago, life ain't easy.
All praise be to Lord Jesus Donda. Please rest easy. Allright, let me cut it there. Let me cut it there. So let's hear those vocals I just recorded now with Kanye over there.
I got a fantasy that's beautiful, that's dark and twisted, but I attacked a whole religion all because of my ignorance. What was I thinking? That was some bitch shit. I lost Adidas, but I'm still Yeezy. Back in the kitchen.
Man, I'm a genius. Boys in the hood, just like I'm Eazy. Kanye Wheezy. Southside of Chicago, life ain't easy. All praise be to Lord Jesus Donda. Please rest easy. Allright, let me cut it there. Let me cut it there.
Yo, that new. So you guys kind of get the, uh, the gist of what he was doing there. But I think, um, a lot of times when we hear AI music, AI music, uh, that's what it is. It's like someone taking, uh, the— the regular songwriting process, where you go get production or do the production yourself, uh, you write the lyrics, and then record your own vocals.
And then what people are calling AI music is really just style transfer or voice conversion, where they're taking that voice and turning it in— in— into someone else's. So this is actually the first shout-out to this guy's name's, uh, Roberto Nixon, if I'm not mistaken.
And I saw this video, and it kind of— I think it just definitely changed the trajectory of— of what I was doing at the time. I dropped everything. I was like, I have to figure out how the heck he just did this Kanye West, uh, voice thing.
Uh, so I— I pretty much, uh, followed, uh, the steps in this video. I dove into some Discord servers and some subreddits, and that's how I really figured out how to, uh, do this for myself. Um, so— so I think, uh, that's— that's one compartment that I've kind of, like, created for AI music and put it, um, uh, to the side.
The other one— well, actually, I— I have a— a couple more examples of— of this as well. Let me—
I think the— the next one is, how many of you guys were familiar— by show of hands, how many of you guys heard that Drake AI song that went viral last year? Okay, we got a couple people. Nice.
Um, so I'll— I'll play it. I don't want to play the whole thing, but, uh, again, this is another example of, uh, voice conversion. And a lot of people were saying, "Oh my God, he pressed a button and generated a whole Drake song," and that's not really the truth.
Uh, this kid is— his name is Ghostwriter, the— the person behind the song. And he did an interview where he talked about the process of, uh, creating the song, and it was all him, except for the actual vocals.
He, like, changed his voice to sound like Drake in— in the weekend. I'll play a little bit of it.
I came in with my ex, like Selena the flex, and bumpin' just some beer, but her favor ain't left, and she know what she need, or her need, or she blessed, and givin' me my best, and I got my heart on my sleeve with a knife in my back.
What's with that, and 21, I love him, that my brother, that's my shot, and Metro made the beat, so you know that it's gon' slap, and yeah, it's gon' slap, and tell him run it back. Talking to a diva, yeah, she on my nerves.
That's actually my favorite part, but I'm going to cut it off. Um, so the Drake song was interesting to me because I kind of saw it unravel. Like, I— I think I found the video or heard the song when it was, like, under 1,000 views, uh, being posted in some of those sub— subreddits that— that I was talking about.
Uh, but it went super viral. I think he got, like, 9 million streams in the first 24 hours. He was on his way— this guy, uh, Ghostwriter, was on his way to charting, of course, before the RIA got a hold of it and was like, "We're shutting this down."
So he got DMCA'd, and the song got wiped from everywhere. Uh, you can still find it on the internet, but it— it pretty much got wiped from everywhere. And, uh, I just thought, I said, "Man, this is really going to change everything."
But, again, everybody was calling it and saying it was AI music without, uh, being specific about what ky— kind of, uh, AI music it was. Um, so I have a couple more videos to show you guys, but I just wanted to go over the Drake one because I thought that was pretty cool.
This next one is actually my favorite example of voice conversion that I've seen to date. Let— let me backtrack a little bit. These two songs that I just played or videos that I just played are from, uh, rough— roughly a year ago.
I think March of last year or April of last year. Since then, the voice conversion tools have gotten 10 times better. And, like, because when I listened to the Kanye West one back then, it was amazing. It literally blew my mind.
No— no— no lie. And today, I kind of listen to it and I'm like, "Wow, it sounds— it's really artifacty. It's— it's bad. It doesn't exactly sound like a natural human Kanye West." But today, they have stuff that sounds almost exactly, you know, like Kanye West.
I think— and I say the kids because when I joined the Discord server, it was literally a bunch of, like, 14, 15, 16-year-olds doing this, uh, these— these conversions. Um, but, uh, the kids have gotten a lot better at, like, uh, training— training, uh, these, um, models, these voice models.
Uh, so the next one is actually my favorite, uh, example of voice conversion. How many— by show of hands, how many of you guys in here are familiar with Randy Travis? Yeah, I knew I was going to get y'all back there.
Yes, sir. Yes, sir. Okay, so Randy Travis, uh, and forgive me because I'm not exactly sure how this happened, but he lost his voice some time ago. And, uh, Randy Travis and his team kind of saw this AI thing unraveling or un— unfolding and decided to sit down, write a song, uh, produce a song, and then use this voice conversion technique for, uh, Randy to sing the song without him actually singing the song.
So this is one of my favorite examples. I think also in the last year or so, the— the— not maybe temperament is not the best word, but
the general public opinion about AI music has shifted a little bit. It hasn't gotten too much better. It's really taboo, honestly. Like, I was a little bit, uh, nervous about doing this conference because, like, people just do not like AI music at all.
Don't want it in their lives. But since then, this— this Randy Travis song came out a couple of months ago or a couple weeks ago, I think it was. And when I looked at the YouTube comments of this video, the feedback was overwhelmingly positive.
People were, like, in the comments saying, "I'm— I'm a 75-year-old man. I'm bawling tearsright now listening to this Randy Travis song." Um, I— like, they were like, "We don't care how we get Randy. As long as we get him back, we'll— we'll take AI or not."
You know what I mean? So it— it just was, uh, a really cool example of, I think, how voice conversion can be used in the— in theright way, you know, by— by the original artist. So I'll go ahead and play some of this as well.
Hello, hello. Okay. Sorry, I apologize to the Randy Travis fans in the back. I'm— I'm going to cut them off a little bit because I— I do want to keep it going. But, um, even though I'm not super into country music, I— I thought it was really touching that he was able to, like, get with his team and— and use, uh, AI to, uh, to create, um, new original music.
I thought that was really cool. Uh, so that— that pretty much, uh, takes care of the style transfer stuff. I've showed you guys some examples. The next thing I want to get into is, uh, text-to-music. Um, and I'll go over full songs as opposed to, uh, um, uh, the— the short samples that I was talking about.
If you, uh, go— again, just want to remind you guys, if you go to the link aietalk.com/music and submit a question on there, I can see it pop up on my screen. Um, and so— so I can, like, answer it in the— in the— in motion.
Text-to-Music20:28
I also want to repeat it out loud for the transcription and all that good stuff. Um, okay, so
text-to-music, text-to-music, text-to-music. I'm going to show you guys a couple examples. The— I— I broke text-to-music down even a— a little bit further, uh, by
having— I have a couple folders here. And I think the way I did text-to-music was, yeah, meme songs and samples. Okay, by show of hands, how many of you guys have heard about this Drake and, uh, Kendrick Lamar beef going on?
Okay, dang. Okay, y'all— y'all are up on game. I appreciate y'all. Okay. Um, so
for anyone who hasn't been paying attention, just know that one man said very disparaging words about another man, and then the internet ran with it. So I don't have to catch you up on 18 million songs that these men have recorded now.
I couldn't even keep up with myself, to be honest with you. Um, but— but yeah, so— so basically, the internet ran with this whole concept of BBL Drizzy. And Kendrick Lamar said a lyric in one of these diss songs that they had back and forth over the last couple weeks, where he called Drake BBL Drizzy.
If I— I don't want to explain BBL to y'all. I ain't going to lie to you. Um, so hopefully you understand what a BBL is, and they call him BBL Drizzy. And it's funny because, like, uh, it can be taken a— a couple different ways.
But I'll just play the music so I'm not rambling too much.
BBL Drizzy. BBL Drizzy.
I'm going in. No diddy.
I'm going in. No diddy.
My grandma would probably cuss me out for, uh, turning that off early, but I'm going to cut it there. And, um, basically say that— that this song was genera— to me, this is really what gets to what AI music is.
Uh, because, uh, there's a comedian named King Wallonius who logged into Udio, one of the tools I'm going to tell you about, typed in a prompt, I think typed in some lyrics, if I'm not mistaken, and generated this song.
He didn't get in a studio. He didn't record his vocals. There was no voice conversion. He just typed in some words and got this BBL Drizzy song. Um, and I— I think, uh, this is, like, a monumental moment in music as well because it was one of the first times I saw something like this go viral.
And I think we'll get into— we— I mean, we can, if you guys are interested, we can get into, like, some of the legality of AI-generated music and the copyrights around it. Um, but— but I think this— this moment where he— he created this BBL Drizzy song and went viral, it's been used in a bunch of different songs now.
It's like— it's like monumental. But— but this is text-to-music. There was no prior recording. He just generated, uh, this song, one— one shot, uh, prompt to a song.
I think I might have one more good example for y'all. I'll play this for a little bit.
11 Labs is another, uh, AI startup that has now entered the text-to-music space. And this is a song that they generated, uh, again, no prior recordings. This is just text. Uh, music came out the other end, and it went viral on Twitter.
We were having fun programming young, dreaming that one day we'd make it work. Lines of code we'd write all night, hoping that one day we'd get itright. Can we teach.
Oh, he was about to get into it. I— I cut off a little too early. But— but basically, this is a sort of meta song because it's an AI model singing about GPUs. Um, and I thought that— that was pretty cool.
Uh, but— but yeah, another text-to-music song that I would kind of put into the, uh, meme— the meme song category. Uh, with that being said, I think I had— oh, yeah, so, uh, audio to music. These are actually the most interesting samples— examples and also, uh, the shortest, if I'm not mistaken.
Audio-to-Music25:04
Um,
so I'll get into some of these. These are really cool. I— I think the— the people who are, like, uh, musicians or artists who have, uh, a lot of talent already, uh, like mu— musical talent already, are— are into, um, oh, let me play this, into this style of, uh, music generation because it takes an existing sound and transforms it.
You might notice that it sounds like.
That new Yeezy track goes. Into another song. Uh, an existing sound and— and turns it into music. Sorry.
So this is, uh, using Stable Audio. It's another tool. It's not Suno or Udio, but it is called Stable Audio. And I think this was the first tool I saw that had this feature publicly available where you could put audio in and get a— a song out or get music out, maybe not a— a full song.
Um, but I— I really like this because it— I think it unlocks a lot of different ways to create music. Um, so I— I think that's— I might have one more example to show you guys. This one was cool.
This is also from the— the Suno team as opposed to Stable Audio. This one's pretty good.
Nice. Okay. Um, so I thought that was pretty cool because literally all he recorded was his fingers drumming like that on the desk, and then out came this full song with a guitar and all these other instruments. Um, so I think that's— you guys kind of get the— the gist of, like, what, uh, audio— uh, audio to audio, uh, model generation or music generation can be like.
Live Demo27:12
Um, and I want to say with that, we can get directly into actually generating some songs ourselves. Okay, we could talk about how it works, but, uh, you guys didn't really seem so much interested in, like, the technical side when I was looking at the poll.
So I could briefly touch on it, but I think, um, how— by show of hands, how many of you guys saw that Suno and Udio have been, uh, are getting sued this week? You guys saw that?
Already?
Yeah. The RIAA just, uh, filed a complaint against both services. So, um, and— and what they're alleging is that these, uh, AI models are trained on massive music datasets that are copyrighted. And we are— well, not we. Don't— don't put me in no indictment.
So I do not want to be— but because I— I don't— that's another thing I— I didn't mention at the beginning, but I have no affiliation to Suno, Udio, Stable Audio, Boomi, no one. Like, I— I— I just liked this stuff and learned it and, uh, figured it out and wanted to do a presentation on it.
Um, but— but what they're alleging is that Suno and Udio are infringing on their copyrights by training on this— this music.
That's— that's last month. You played before. That's a cool new Tom Petty.
Yeah. Yeah. Yeah. We had a— a comment from the audience saying that the last, uh, generation sounded like Tom Petty. Um, if you are more interested in how these models actually work, the architecture, transformers versus diffusion, uh, I have a link on the page that I mentioned earlier.
If you click on the notes button and then scroll all the way to the bottom, I have compiled a list of all of the, uh, music model papers. So if you guys are interested, uh, that's some— you— you can find it there.
Um, okay. So what I want to do now as we get into, uh, Udio and Suno, there is a lot less people than I thought there was going to be, so this will work well, hopefully. But there's a business card on the table in front of you, and that business card has a secret code on it.
This is a little PvP, uh, survival of the fittest, like may— may the best man win. But if you take that secret code, go to the link that I showed earlier. Let me put it on the screen again.
And click on the gateway button. There should be a green gateway button that looks like this. If you put in that code, what will pop up is a username and a password. If you use that username and password to sign into any of these services— let me try and make it a little bigger than that.
If you want to get into Udio, you can use Google, Discord, Twitter, Apple. I've signed up for all the services. If you want to get into Suno, you can do Discord, Google, or Hotmail, Microsoft, whatever they're calling it nowadays.
And then you can sign into— so what I did for this presentation so you guys could generate music, uh, freely because I— I know in the show notes, uh, for this workshop, I put that you guys should have your own account.
But I also thought it would be cool if we could do, like, unlimited generations as opposed to what they give you on the free, uh, account. So what I did was I signed up and paid for a Premier account for Udio and Suno and gave you guys a login just now.
So if you go log in, you can generate music as we're talking about, uh, the next couple steps.
Just know that, uh, I— like, I think 2FA is on Apple, if I'm not mistaken. So there might be some difficulty getting in, but just try and get into one of these services, then log into Udio or Suno.
And— and the— the, uh, URL for that, I think I have it here, is udio.com. Once you've signed into one of the other authorization services or suno.com, we'll get you into, uh, one of these— one of these sites.
Worked? Nice. Awesome. We got somebody in. Oh, another person. Awesome. Awesome. Awesome. Hey, PvP. Some of you guys might not make it. I'm going to lie to you.
But yeah, I just wanted to do that so you guys had, um, and actually, what I'm going to do as well, like, when you guys are in your seats generating songs, there's another link on this, uh, uh, aietalk.com/music page where it says submissions.
What I want to do, uh, we might be tight on time, but what I want to do is have everyone who generates songs here in person to submit to this link. Uh, it's basically just a Google form. So you generate some music, take the link, the share link to that music on either service, uh, click on this link here, submissions, and submit your song.
And I think what I'm going to do at the end is just take, like, a random number generator, and, uh, two people will be selected to get access, like, keep access to these two accounts that I set up.
So one of you guys will walk away with, like, a Suno account. One of you guys will walk away with the Udio account. And like I said, I— I paid for the Premier. It's— I didn't pay for the whole year.
After they got sued, I was like, "Yeah, I might not want to invest."
So I just went ahead and did 30 days.
Okay. So we can get into some of the techniques. While— while we're on this topic, uh, and you guys are hopefully generating in your seats as well, does anyone here have a birthday this week? Nice. My man, Randy Travis.
Um,
what's your name?
Patrick.
Patrick. Okay. I think what we're going to do is generate a— a birthday song for Patrick. And I'm going to do that from my account.
Wait, what's— one more time?
I'll be immortalized.
Yes, absolutely. Patrick, my man.
Uh, Randy Travis. Oh, they're going to block that. Fan and, uh, music. I'm not even going to try.
There you go.
Yeah, they just— I don't know. Udio may not be as bad about this, but, uh, if you put artist names and obviously now because they're being sued, but if you put artist names, they, like, they have a content filter, you know what I mean?
They don't.
I'm trying to get into the system and use it to build identity and keep voice guys and it's always going to.
Absolutely. I'm— I'm going to talk about that as well, like ways to kind of get around the filters. Um, but yeah, that definitely works. Uh, we had a comment from the audience that basically said, "If you, like, change the spelling up a little bit and then put, like, the style of music that the artist is related to, you can, uh,
you can skate."
Attorney.
An attorney.
He's a past attorney, so put that in there.
Oh, shit. Oh, you're the one suing, uh, Suno and Udio.
Uh, birthday song for Patrick. Let's put in some. Well, yeah. Patrick, what do you like? What kind of music do you like?
Pretty much
the fan.
Which is what genre? Rock? One more time?
Rocker Billy.
Rocker Billy. I don't— I'm— I'm be honest with you. I'm lost.
Did you say the band? The band. The band. Yeah. Yeah. That's kind of like this one's kind of about rock.
Let's go ahead and generate a song for my boy Patrick. And then I'm going to take this same prompt. I'm kind of getting into some of the, um, prompting techniques here, but you guys are able to see this screen, I hope, and see what I'm doing.
And then we'll play a song. Uh, couldn't generate that song description name, Randy Travis. What about Tandy Ravis?
Ravis? Who is Ravis? Listen, brother, just give me my song.
Allright.
Allright. So I'm not monologuing the whole time. I'm going to go ahead and wrap it up. But I— I the way I think of these two different services is that, uh, Udio is more like a songwriting partner. Um, they have an experimental model that you guys now have access to if you log into that Premier account that I paid for.
Suno vs Udio35:49
Uh, but— but the default Udio experience is that you generate 30 sec— 30 seconds of music at a time. Um, and you could either extend that idea, remix that idea, or there's a third option I didn't write here, and you can scrap it because a lot of times, you know, you get weird stuff, uh, when you— when you're not great at prompting.
Um, but you can also use Inpainting, uh, to— to modify, uh, specific sections of a song, like if you want to change or tweak one, uh, or two little things. And then now they have this feature on, uh, within Udio where you can do the audio to audio, and it— it allows you to upload music.
Um, so I— I if you are an artist and you write songs or have written a song before or spoken word or anything, I would urge you to try to take one of your creations, especially if you've recorded it, it gets crazy.
If you've recorded some music before, you should take the tempo, uh, the key that that song is in and your lyrics and put it into Udio and see what comes out the other end. It's blown my mind a couple times.
I've taken a couple of my own songs and put them in, and it— and it gets crazy what— what comes out the other side. Like, Udio does me better than I do me sometimes, honestly. So, so it's a really cool experience if you have, like, a— a pre-existing music, you can, uh, really play with Udio a lot.
And then the way that I think about Suno, um, is your in-house music producer. Uh, I like to describe it in a way that I— I feel like
somewhere in the system prompt of Suno's model, they have, like, top 40 music in there because it always comes out really polished and clean. And maybe it's the way that the— the— the data that they put into the mu into the model in the first place, but I— but, uh, my experience with Suno has been— been that, uh, it just comes out really, really clean, um, top 40-ish.
One thing it's gotten really, really good at lately, this is like a, a side note, is that when I first— I— I personally have onboarded, like, I think a dozen or so people over the last couple weeks because I wanted to, um, understand what is the actual use case, uh, for these services.
So I just started onboarding friends and family, some of my music friends, some people that are not into music at all. And what I found when a— a couple weeks ago was that, like, when I would generate songs, uh, if you're not super specific in counting the syllables in each line of the song when you're getting into the— the custom mode of Suno, it'll the timing of the beat and the artist, quote unquote, the vocals, uh, lyrics gets super off, really, really, really off.
Um, but they've gotten really good about this lately. Like, the last couple songs I've generated from Suno aren't off. Um, they've the song structure is a lot better because that's another tip we'll get into. The— the song structure really, um, has improved a lot, um, um, in Suno over the last couple weeks.
But— but it's more or less the same. Just the outputs are a little bit different from, uh, Udio. And now we get into the tutorial.
Um, so have you guys— anyone submitted any songs yet? Let me see. Let me recheck.
I was just, uh, generating some music for my boy Patrick. We're going to play that in a second. But that was— that was the basic tutorial. You kind of it's about as simple as it gets. Like, it really, really lowers the barrier to entry for creating music.
And I think we— we can get into some— some other stuff about, like, more conversations about that, but I think it for now, I'll just say that it— it lowers the— the barrier and makes it easier to, uh, create stuff for, you know, specific situations.
Okay. Nice. We have a couple submissions. We do have a couple submissions. But the gist of it is you go to udio.com, you log in, you go to this, uh, oh, they used to have oh, I guess here at the top, it always says create.
Then you can type in, uh, what's another one? The AI Engineer World's Fair. Let me start over. A song about the AI Engineer World's Fair.
And I know, uh, the organizer Benjamin Duffy really likes old school hip hop. Boom bap. So I'll generate a song. So you come in, you type in your prompt, you hit create. Uh, there's a couple of different options down here.
Again, if you have a free account, you won't see this experimental longer song. Um, uh, you'll— you'll be defaulted to the Udio 32 model, but you can come in and write your own lyrics on top of this prompt that you just put in.
Or— or you can select to only have an instrumental generated as opposed to, uh, a full song with lyrics and everything. Um, so at— at— at its core, that's how it works for pretty much all these services as well.
There's not too much beyond that that you need to know to start generating music. I would argue that there's a lot more that you need to know if you want to get good at it, but that's— that's pretty much, uh, how we, um, generate music on both services.
And what I like to do just to get a nice even comparison is use the same exact prompt. I'll refresh the page because, like, for some reason, they're really buggy, uh, sometimes where if you generate song after song without refreshing the page, you get, like, extensions of songs as opposed to a brand new song.
But I will go ahead and create based on this prompt.
And then I'm going to come back because we've now gone over the basic tutorial. I'm going to come back and look at some questions because I know some questions came in. Nice. And in the meantime, while I'm getting my thoughts together, I'm going to, uh, play my song or the song that— that was created for Patrick.
Wishing him a happy birthday. Oh, moderation error. They didn't like your, um, or my— my Randy Travis.
Let's go ahead and play this one for a little bit.
It's your day, Patrick. We're singing loud for you. Oh, yes. Oh, yes. It's true. It's your time, Patrick. Celebrate the whole day through. Oh, yeah. Oh, yeah. You knew.
Happy birthday.
Oh, yeah.
Oh, we know, Patrick, you love that rock and roll. Oh, yeah. Yeah. Oh, yeah. And country hits. You feel it in your soul. Oh, yeah. Oh, yeah.
Patrick, what do you think?
It's great. Thank you.
Patrick said it's the— the theme song for his sitcom. I'm going to play these other two quick. I just want to hear what they sound like.
Uh, real quick, if you click on the actual song name, you can see the lyrics on the Suno side.
Birthday candles flamed so bright for a connoisseur of music's light.
Rocking through the day and night. Patrick's world is pure delight.
I'm going to try the other version. See if, uh, was that more your speed, Patrick?
Yeah, it was great.
Oh, awesome. Okay. Allright. Oh, we got five minutes left on the TV.
Birthday candles.
Okay. So either they generated the same song twice or
yeah, maybe. Yeah. Let's see what the AI Engineer World's Fair song. And actually, let's address these questions because we only have a couple minutes left.
And then I'll get into the notes real quick because, uh, I did want to show you guys some of the prompt stuff. The truth is that, like, I have everything in, um, under the notes section, like pretty much everything I talked about today.
Q&A44:03
Also, when the recording becomes available, you can come back to the same page, get access to the recording, and then I'll hopefully have the transcript up in, like, an hour, um, because I recorded it myself. But if you come here to this page, aialtalk.com/music, and click on the notes section, um, and you want to not listen to me talk about the prompts, you can find the prompt stuff yourself, uh, not quite all the way at the bottom, but here, uh, close to the bottom.
And, uh, this is a GPT prompt, meaning you literally can go to ChatGPT, use this prompt. It— it got a little bit weird. Like, for some reason, when I use these closing brackets in GitHub, um, it, like, basically disappears the whole and, like, you won't even see lyrics anymore.
But basically, what you can do is I— I think even this would work. Like, GPT-4 is now smart enough to do this. You can take and you can put in this, uh, template and then actually fill in the top here where it says song description, and GPT will actually produce some coherent lyrics.
And it allows for better music generation when you, uh, come in with the lyrics as opposed to giving it a general topic and hoping it comes up with good lyrics itself. You can kind of tweak the lyrics yourself before you start to generate.
So that's the GPT prompt. And then for the music model prompts, uh, these are pretty much just examples of what we, uh, just did with, uh, Patrick and, um, the AI Engineer World's Fair song. It's just like a general description of a song, um, that you want to hear, that you want to generate, and then it pops out the other end.
Uh, super pressed for time. Um, there are some official Udio and Suno tips, um, that are very useful as well. And then here I want I kind of want to get into the prompt master thing, actually. Philip asked, is there a way to transfer to do voice transfer of a more generic voice rather than a specific voice, referring to, uh, the voice conversion stuff that we talked about earlier?
And I think the end of that question got cut off because you said like.
Yeah. Like, like the copyright. I don't want to be Eminem. I want to be, like, genericright now.
Are you talking about doing voice conversion training, like training models yourself, like going to one of the services or downloading the tools? You literally can just take and blend multiple voices. So the way that it works is, like, you need, depending on what service or model you're using or what, uh, um, training toolset you're using, you need, like, minimum six seconds of someone's voice to start training.
Um, and— and you get, like, a kind of coherent model after that. But what you can essentially do is just six seconds of your voice, six seconds of his voice, six seconds of Eminem, and it kind of, like, blends it together.
Or you can just go to a TTS service and get, like, an actual, like, robot voice, like someone a voice that doesn't belong to a real person, and then use that to train on. So yeah, for copyright purposes, yeah, you can— you can do that.
Dang, they got to wrap it up. Okay. Um, what recommendations do you have to help advocate for the responsible and ethical use of
yeah, sorry about that. I don't know why the questions are getting cut off.
Oh, ethical use of AI music. Um,
that's a good one. I'm going to be honest, I'm going to plead the fifth because I'm not a lawyer. I'm going to be completely honest. I'm going to be completely honest.
You're going to do it as well.
Oh, yeah. Yeah. Good point. Um, what tool generates the best cloned voice that can match emotional tone like Kendrick Lamar's "Euphoria"? Three switches. Generates the best cloned voice. I would definitely look into RVC. If you can— you can use RVC in the cloud.
You don't have to like, when you Google RVC, there comes up a bunch of tutorials of how to run it yourself locally. You do not have to do that. You can run it in the cloud because you— you need, um, some decent hardware to do, like, computer hardware to do it.
But— but definitely look into RVC. It's the best, uh,right now. Can we have access to the slides? Yeah. And my time is up shortly. Oh. Oh, my— did my thing go to sleep? It did go to sleep.
Oh, to log into Udio or Suno?
No.
Really? Uh.
You said request access.
Oh, okay. Okay. Sorry. That's my blunder.
Um, that was pretty much it. Oh, rate your— rate your music is something you definitely should check out. Like, what someone basically reverse engineered is that these models were trained on labels from Rate Your Music. So you can almost reverse engineer exactly what sound you want by going to Rate Your Music, typing in an artist name, typing in a genre, um, or really just typing in what you want to hear, and then using whatever labels they put on that music.
You put it back into the model itself, Suno or Udio, Stable Audio sometimes, I think. Um, and you can get, like, very, very close to a specific, uh, sound. I don't want to say specific artists, but, like, yeah, you can get pretty close.
Um, yes. And I— I was going to get into, like, the law thing. I don't— you— you want me to differentiate between ethics and law, and I understand, but I was going to get into a little bit, but we— we kind of ran out of time, so I apologize for that.
If you guys would like, I'd love to talk about this stuff all day. If you— if you want to approach me after the workshop and we chop it up all day about this stuff, I— I love it. Um, yeah.
And— and then what I'll do for giving away the accounts, maybe I'll do, like, a live stream or something like that and do, like, a grab everybody's email who submitted because I think we had— nice, we had 12 submissions.
Let me refresh. I think we had 12 submissions. And I'll pick somebody and just, like, send you guys I'll change the username and password so everybody can't log into your stuff, and then I'll give you— uh, I'll give you the accounts.
But yeah, I— I appreciate you guys for coming. Uh, hope you guys got something out of this and walked away with some music, generated some stuff. And I think I'll play a song to go out.
Walking through the fair, copper wires in the air, coders in the screens, dreams turning into scenes, blueprints in my hand, robots take a stand, lines of code they shift, tech minds give the gift. AI dreams, they see this fair.
Dang, that's dope. I wonder what Udio generated.
Oh. Oh, first because I just saw the question that someone said, uh, stems, how to, like, pull the stems out and, like, edit them. There's a tool, again, not— not, um, like, I have no affiliation with them, but if you just go to actually, the easy way to do this is let's say we just take this, uh, happy birthday song by Patrick.
Stems50:54
We go to suno.com. All you have to do is, uh, suno.com or sunodaw.com, like, literally go to the link where you just generated the song, type in DAW, Digital Audio Workstation, at the end of the, uh, URL, not the end of the URL, but the end of the domain name, and hit enter.
And I think it takes anywhere from 30 seconds to let's go ahead and sign in real quick to 90 seconds. I've had to wait longer for some of the songs because it's more complicated to do. Uh, sorry guys, I— I forgot to mention this, but yeah, like, if you go and do that, type in DAW at the end of the, uh, Suno song, it'll pull that— that exact song that you just generated into this Digital Audio Workstation and give you stem by stem and then let you control the volume of each— each specific stem.
And by stem, I just mean each, uh, specific instrument. They do all the different instruments and then the vocal itself. So this is a really, really cool tool that I think is, like, a startup that hasn't been around for that long.
Um, but it's— it's a really, really cool tool for, like, uh, being able to— to get more control out of your generations. So I just wanted to
DAW, so Digital Audio Workstation, and you can, like.
Birthday candles flamed so bright.
Like, let's say I don't like the vocals, but I really like the music.
Connoisseur of.
Rocking.
Patrick's world is pure delight.
End of chord and end of show.
And this is just vocals.
As a patent when seats on go.
It's the bass.
Rock tunes flowing in his veins.
Patrick dances through the rain.
Happy birthday, Patrick, dear.
Let the music keep you near.
Rocking hard all through.
So yeah, that— that's, uh, pretty much how you use Wave. It's called Wave Tool, that, uh, that service. So that's how you can get the stems. There's also another tool called UVR5, but it gets really involved. You do have to have, like, decent hardware, but you can also use, like, uh, instead of using, like, an online service where I think all the accounts on Wave Toolright now are free, but you can use UVR5 locally forever.
It's like an open-source software project that, uh, allows you to pull stems just like this, literally. You can get and that— that's not just, uh, AI-generated music. Any song you can put into UVR5 and get, like, the stems from.
It's actually pretty cool. I know. Yeah. Appreciate it. I know they want you guys to go over to the, uh, what was it called again?
General session.
Yeah, the general session. So I'm— I'm going to shut the hello now. I appreciate you guys. Thank you. Thank you for coming.





