Earlier quoted context omitted.
I agree that there is some incremental electricity usage. I do not think it can be characterized fairly as "massive environmental harm".
As an example, Ren and his colleagues calculated the emissions from training a large language model, or LLM, at the scale of Meta’s Llama-3.1, an advanced open-weight LLM released by the owner of Facebook in July to compete with leading proprietary models like OpenAI's GPT-4. The study found that producing the electricity to train this model produced an air pollution equivalent of more than 10,000 round trips by car…
That seems like very low impact, especially considering training only happens once. I have to imagine that the ongoing cost of inference is the real energy sink.