Can anybody explain why the researchers are attempting to generate the whole song as a single waveform, as opposed to wiring generated MIDI into some instruments and separately a singing algorithm (perhaps a bit easier than the whole bulk work)?
We did work last year on MIDI alone - https://openai.com/blog/musenet/ and some early work now on conditioning the raw audio based on MIDI (early results at the bottom of the Jukebox blog). Agreed though there should be interesting results from modeling different blends of MIDI, stem, and raw audio data. Raw audio alone gives us the most flexibility in terms of the kinds of sounds we can create, but it's also the mos…
Jukebox
131–134 of 134 posts
Re: Jukebox
#132From the GitHub repo: "On a V100, it takes about 3 hrs to fully sample 20 seconds of music." That might make building off this project out of reach of the average engineer (you certainly cannot build that into a Colab notebook), although that necessary amount of compute is not surprising.
Re: Jukebox
#133Earlier quoted context omitted.
The future will be AI lawyers battling for rights of AI generated music in the style of deceased artists on behalf of AI media corporations at the expense of robotic listeners. Before all of this, we'll probably see improvised bands of deceased artists playing together AI generated music in their own style, not to mention long dead actors appearing in new movies etc. AI technology is going to give law firms a lot of…
This subthread made me immediately think of: "If you want a vision of the future, imagine a human face booting on a stamp forever." (From the last story at https://slatestarcodex.com/2016/10/17/the-moral-of-the-story... )
Re: Jukebox
#134Earlier quoted context omitted.
Just know that: much of the stuff OpenAI and other research orgs put out (including mine) are heavily cherry-picked. Most of the time its pumps out gibberish, but in the off chance it doesn't it gets used as marketing material.
monkey writing shakespeare