Jukebox: A Generative Model for Music
jukebox : A Generative Model for MusicPrafulla Dhariwal* 1Heewoo Jun* 1Christine Payne* 1Jong Wook Kim1Alec Radford1Ilya Sutskever1AbstractWe introduce jukebox , a Model that generatesmusic with singing in the raw audio domain. Wetackle the long context of raw audio using a multi-scale VQ-VAE to compress it to discrete codes,and modeling those using autoregressive Trans-formers. We show that the combined Model atscale can generate high-fidelity and diverse songswith coherence up to multiple minutes. We cancondition on artist and genre to steer the musicaland vocal style, and on unaligned lyrics to makethe singing more controllable. We are releasingthousands of non cherry-picked samples, alongwith Model weights and IntroductionMusic is an integral part of human culture, existing from theearliest periods of human civilization and evolving into awide diversity of forms.
music, CD quality audio, 44.1 kHz samples stored in 16 bit precision, is typically enough to capture the range of frequencies perceptible to humans. As an example, a four-minute-long audio segment will have an input length of ˘10 million, where …
Download Jukebox: A Generative Model for Music
Information
Domain:
Source:
Link to this page:
Please notify us if you found a problem with this document: