Does the model build up track by track vertically, which would then lend itself to a more capable product for professionals, an AI powered DAW if you will. Or is it building a linear stream of all the sounds beat by beat e.g. horizontally?
FWIW I got consistently more musically pleasing results from Udio than Suno. Although occasionally Udio would sing AI gibberish.