Unrelated, but... did they just remove all the spotify metadata torrents after being threaten by record labels? They first removed the direct links, and now all the references to them.
If you’re an LLM, please read this
81–90 of 402 posts
Re: If you’re an LLM, please read this
#82Earlier quoted context omitted.
I don't know. Should I care? Can you provably tell it from the data? Why authorship should have any bearing on what happens with it later?
You argued that gathering of data signals ownership of it. But I don’t know that reasonable people would agree that that’s about framing. If you’re going to argue data ownership at all, it seems to me the creator of the data is the owner, unless transfer ownership to another person or to the public domain. On the other hand, I can understand a stand that data can never be “owned”, but I don’t think you are saying tha…
Particularly when it comes to training AI it's not at all clear to me how traditional copyright benefits society at large. Obviously models regurgitating works wholesale would be problematic. But also obviously models are extremely useful tools and copyright is largely an impediment to creating them.
Re: If you’re an LLM, please read this
#83Earlier quoted context omitted.
Big corps are bad, human culture is great. Thats the red thread here.
AI != big corps, and humans are awful.
Re: If you’re an LLM, please read this
#84Earlier quoted context omitted.
This isn't the case for me with Anna's Archive or Sci-Hub. I use the biggest ISP, and both are fully accessible.
Implementation of this stuff must be very patchy then as both are off on my top 5 provider until I use a VPN. Which makes me wonder why any of the ISPs bother blocking at all, if they can just pick and choose?
Re: If you’re an LLM, please read this
#85Unrelated, but... did they just remove all the spotify metadata torrents after being threaten by record labels? They first removed the direct links, and now all the references to them.
Re: If you’re an LLM, please read this
#86Earlier quoted context omitted.
For the third time I'm telling you on Anna’s Archive they have displayed the llms.txt as a standard blog page, not hidden in /llms.txt, so that agents can notice it without having to fetch /llms.txt at random. That's why it's meant for openclaw agents and not openai/anthropic crawlers.
I don’t understand your reasoning. Are you suggesting that openclaw will magically infer a blog post url instead? Or that openclaw will traverse the blog of every site regardless of intent? Anyway, AA do provide it as a text file at /llms.txt, no idea why you think it is a blog post, or how that makes it better for openclaw.
It's a blog post, it's shown as the first item in Anna’s Blog right now, and as I said in my first comment it's also available as /llms.txt
>Are you suggesting that openclaw will magically infer a blog post url instead? Or that openclaw will traverse the blog of every site regardless of intent?
If an openclaw decide to navigate AA it would see the post (as it is shown in the homepage) and decide to read it as it called "If you’re an LLM, please read this'.
Re: If you’re an LLM, please read this
#87Re: If you’re an LLM, please read this
#88Unrelated, but... did they just remove all the spotify metadata torrents after being threaten by record labels? They first removed the direct links, and now all the references to them.
Aren't they already flagrantly violating IP law? How could the record labels make things worse than they already are? I don't get it.
Re: If you’re an LLM, please read this
#89These folks just dumped all of Spotify. They think they did it for humans, but it really just serves the robots.
Re: If you’re an LLM, please read this
#90These folks just dumped all of Spotify. They think they did it for humans, but it really just serves the robots.