I'm more interested in what that content farm is for. It looks pointless, but I suspect there's a bizarre economic incentive. There are affiliate links, but how much could that possibly bring in?
It's for shits-and-giggles and it's doing its job really well right now. Not everything needs to have an economic purpose, 100 trackers, ads and backed by a company.
Anyone got a contact at OpenAI. They have a spider problem
331–340 of 400 posts
Re: Anyone got a contact at OpenAI. They have a spider problem
#332Earlier quoted context omitted.
I'm not sure any publisher means for their robots.txt to be read as: "You're disallowed, but go head and slurp the content anyway so you can look for external links or any indication that maybe you are allowed to digest this material anyway, and then interpret that how you'd like. I trust you to know what's best and I'm sure you kind of get the gist of what I mean here."
How would one know he is disallowed without reading each site?
In this case, as in many, the disallow rules are intentionally meant to protect the signal quality and efficiency of the crawler.
Re: Anyone got a contact at OpenAI. They have a spider problem
#333Earlier quoted context omitted.
Please can you give an example of what might not be a memory error. Not that I think "memory error" is the right phrase either.
I was thinking along the lines of answering with correct information but not following the prompts. Maybe this could be considered confabulation also.
I believe that you can also cause something a bit like a transient dysphasia by giving them bad inputs as well, so there is that on the language production side. However there's still nothing that pertains to the experience aspects central to what hallucinations actually are.
Re: Anyone got a contact at OpenAI. They have a spider problem
#334Earlier quoted context omitted.
Which is crazy because there's plenty of good content for kids on Youtube (if you really need a break!). Blippy, Meekah, Seasame Street, even that mind-numbing drivel Cocomelon (which at least got my girls talking/singing really early).
There's actually no such thing as good "content" for kids, sorry.
Re: Anyone got a contact at OpenAI. They have a spider problem
#335Earlier quoted context omitted.
I'm really cross that the word "hallucination" has taken off to describe this as it's clearly in incorrect word. The correct word to describe it is "confabulation", which is clinically more accurate and a much clearer descriptor of what's actually going on. https://en.wikipedia.org/wiki/Confabulation
I think hallucination is better and more accurate at it implies a bit of imagination and buffoonery deceit. I don’t think confabulate matches as well as it implies confusion or mixture of different ideas. ChatGPT isn’t confused, it’s making things up. It’s trying to bullshit as best it can in hope that what it makes up convinces its user.
Re: Anyone got a contact at OpenAI. They have a spider problem
#336Earlier quoted context omitted.
Reminds me of a joke Three logicians walk into a bar. The bartender says "what'll it be, three beers?" The first logician says "I don't know". The second logician says "I don't know". The third logician says "Yes".
That sounds like my old neighbour, a professor of logic from the university of science.
Re: Anyone got a contact at OpenAI. They have a spider problem
#337Re: Anyone got a contact at OpenAI. They have a spider problem
#338Re: Anyone got a contact at OpenAI. They have a spider problem
#339Earlier quoted context omitted.
Reminds me of a joke Three logicians walk into a bar. The bartender says "what'll it be, three beers?" The first logician says "I don't know". The second logician says "I don't know". The third logician says "Yes".
If, like me, you didn't get the joke at first: Both of the first two logicians wanted a beer; otherwise they would know the answer was "no". The third logician recognizes this, and therefore knows the answer.
Re: Anyone got a contact at OpenAI. They have a spider problem
#340Earlier quoted context omitted.
Is this like, the AI equivalent of “another layer will fix it” that crypto fans used? “It’s ok bro, another model will fix, just please, one more ~layer~ ~agent~ model” It’s all fun and games until you can’t reliably generate your base models anymore, because all your _base_ data is too polluted. Let’s not forget MS has a $10bn stake in the current crop of LLM’s turning out to be as magic as they claim, so I’m sure t…
I mean, you can use Phi now and it outclasses anything else in its size. This isn’t some “it could happen” situation.
My point is about the inevitable future when _those_ models start to struggle.
The phi approach doesn’t seem like breaking the ouroboros, it just feels like inserting another model/snake into the loop.