The advancement in voice conversion and text-to-speech technology has resulted in the emergence of musical deepfakes and audio compositions featuring the voices of celebrity artists, often created without their participation. These deepfakes have gained widespread attention, with many going viral and causing disruption in the music industry. This paper examines how advancements in technology, such as Generative Adversarial Networks (GANs) and Text-to-Speech (TTS) algorithms, have blurred the lines between human creativity and machine-generated content in music. These developments raise serious concerns, including the potential for discriminatory content creation and financial and contractual complications for artists affected by deepfakes. The paper suggests that further research is needed, particularly focusing on public perception of this technology and developing strategies to prevent potential harm.
Babita and Amit Ahuja. "Melodic Deception: Exploring the Complexities of Deepfakes of Music Generated by Generative Adversarial Networks (GANs)." Sangeet Galaxy, vol. 13, no. 2, 2024, pp. 16-24.