Wouldn't this kind of ruling effectively put a halt to ChatGPT and other AI's training on publicly accessible data? What's the difference between Copilot creating output based on code on Github, and ChatGPT giving answers based on a NYT article (without attribution)?
Unfortunately this is frequently abused where researchers build a model under the exemptions, and then others use that model commercially, even if they wouldn’t be allowed to build that model directly themselves.
Anyways, the scientific progress would continue, but products would halt until product developers get some kind of agreements with content creators (eg maybe people start adopting a new kind of open-ish license).