Play Me Something Icy: Practical Challenges, Explainability and the Semantic Gap in Generative AI Music

Jesse Allison,Drew Farrar,Treya Nash,Carlos Román,Morgan Weeks,Fiona Xue Ju
2024-08-14
Abstract:This pictorial aims to critically consider the nature of text-to-audio and text-to-music generative tools in the context of explainable AI. As a group of experimental musicians and researchers, we are enthusiastic about the creative potential of these tools and have sought to understand and evaluate them from perspectives of prompt creation, control, usability, understandability, explainability of the AI process, and overall aesthetic effectiveness of the results. One of the challenges we have identified that is not explicitly addressed by these tools is the inherent semantic gap in using text-based tools to describe something as abstract as music. Other gaps include explainability vs. useability, and user control and input vs. the human creative process. The aim of this pictorial is to raise questions for discussion and make a few general suggestions on the kinds of improvements we would like to see in generative AI music tools.
Sound,Artificial Intelligence,Audio and Speech Processing
What problem does this paper attempt to address?