Earlier quoted context omitted.
> don’t you think OpenAI would’ve worked on something like this? Along this line of thought: was it a massive oversight for them to not train the model to say "math detected, let me pass that to a solver" instead of trying to guess what token should come next in a math problem?
There's a million categories of problem you could ask an LLM to try to solve. You'd need a million solvers…
SafeGPT: New tool to detect LLMs' hallucinations, biases and privacy issues
61–70 of 85 posts
Re: SafeGPT: New tool to detect LLMs' hallucinations, biases and privacy issues
#62Seems too good to be true, and I don't understand what it even means for an LLM to be unbiased.
"Unbiased" almost always means "has biases that are similar to mine". I can't think of very many exceptions to that, frankly.
Re: SafeGPT: New tool to detect LLMs' hallucinations, biases and privacy issues
#63Earlier quoted context omitted.
There's a million categories of problem you could ask an LLM to try to solve. You'd need a million solvers…
If you used some sort of plugin system, you could just make a solver for your specific task and drop it in. Doesn't ChatGPT Plus do this now?
Re: SafeGPT: New tool to detect LLMs' hallucinations, biases and privacy issues
#64Earlier quoted context omitted.
I can't help but notice your accounts only activity before this post was praising another giskard.ai submission a few months ago. Anything you'd like to disclose?
You should assume everything posted on the internet has an ulterior motive. Relying on disclosures simply allows actual bad actors to avoid scrutiny. (And no one cares that you used to work at Microsoft or whatever).
Re: SafeGPT: New tool to detect LLMs' hallucinations, biases and privacy issues
#65Earlier quoted context omitted.
But isn't that half the reason people are so excited about this stuff - that you can ask it to make up an episode and it does a plausible job.
But deliberately requesting and receiving content generation is altogether different from requesting a factual answer and receiving plausible-seeming nonsense. Or at least, it's different to the person asking; it's the same thing as far as the model is concerned.
Re: SafeGPT: New tool to detect LLMs' hallucinations, biases and privacy issues
#66Earlier quoted context omitted.
> don’t you think OpenAI would’ve worked on something like this? Along this line of thought: was it a massive oversight for them to not train the model to say "math detected, let me pass that to a solver" instead of trying to guess what token should come next in a math problem?
There's a million categories of problem you could ask an LLM to try to solve. You'd need a million solvers…
Re: SafeGPT: New tool to detect LLMs' hallucinations, biases and privacy issues
#67Looks a bit like snakeoil to me. A lot of companies now spinning up simple demos with opaque backends, making huge claims they’ve solved X hard problem for/with AI, then saying “trust us” and “join our waitlist” without hard details or facts to show for it. If you could detect hallucinations/biases etc that easily, don’t you think OpenAI would’ve worked on something like this?
Re: SafeGPT: New tool to detect LLMs' hallucinations, biases and privacy issues
#68This is like saying, "I've developed a new compass for a deep space probe to help it find North!" Our society is actively declaring that falsehoods are truth, and should be celebrated. We're hallucinating ourselves. All this software does is make sure LLMs hallucinate with us.
Before long, we could end up with left-leaning and right-leaning AIs autonomously fighting the 'culture war' over social media, much more advanced than simple bots spamming copy+paste comments. Combined with ever-improving ways to fake video and voices, things could get even uglier than they've been over the last few years.
Re: SafeGPT: New tool to detect LLMs' hallucinations, biases and privacy issues
#69Re: SafeGPT: New tool to detect LLMs' hallucinations, biases and privacy issues
#70Earlier quoted context omitted.
"Unbiased" almost always means "has biases that are similar to mine". I can't think of very many exceptions to that, frankly.
That's obviously untrue after more than three seconds of critical thought