Live data from Hacker News

Open Challenges in LLM Research

huyenchip.com

1–10 of 74 posts

Re: Open Challenges in LLM Research

#2
Looking back in 25 years, the "Hallucination Problem" will sound a lot like the "Frame Problem" of the 1970s.

Looking back, it's a bit absurd to say that GOFAI would've got to AGI if only the Frame Problem could be solved. But the important point is why that sounds so absurd.

It doesn't sound absurd because we found out that the frame problem can't be solved; that's beside the point.

It also doesn't sound absurd because we found out that solving the frame problem isn't the key to GOFAI-based AGI. That's also beside the point.

It sounds absurd because the conjecture itself is... just funny. It's almost goofy, looking back, how people thought about AGI.

Hallucination is the Frame Problem of the 2023 AI Summer. Looking back from the other side of the next Winter, the whole thing will seem a bit goofy.

Re: Open Challenges in LLM Research

#3

Looking back in 25 years, the "Hallucination Problem" will sound a lot like the "Frame Problem" of the 1970s. Looking back, it's a bit absurd to say that GOFAI would've got to AGI if only the Frame Problem could be solved. But the important point is why that sounds so absurd. It doesn't sound absurd because we found out that the frame problem can't be solved; that's beside the point. It also doesn't sound absurd beca…

Does anyone think we would have AGI if only we could solve the hallucination problem?

Re: Open Challenges in LLM Research

#4

Looking back in 25 years, the "Hallucination Problem" will sound a lot like the "Frame Problem" of the 1970s. Looking back, it's a bit absurd to say that GOFAI would've got to AGI if only the Frame Problem could be solved. But the important point is why that sounds so absurd. It doesn't sound absurd because we found out that the frame problem can't be solved; that's beside the point. It also doesn't sound absurd beca…

I don't know much about AI research but the idea of "measuring" hallucination definitely seems very loaded to me. Humans hallucinate too and I don't think we can measure that. It almost feels like "we need AGI in order to develop AGI".

Re: Open Challenges in LLM Research

#5
So far it's been ~8 months since ChatGPT started the (popular) LLM craze. I've found raw GPT to be useful for a lot of things, but have yet to see my most frequently used apps integrate it in a useful way. Maybe I'm using the wrong apps...

It'll be interesting to see what improvements (in a lab or at a company) need to happen before most people use purpose-built LLMs (or behind the scenes LLM prompts) in the apps they use every day. The answer might be "no improvements" and we're just in the lag time before useful features can be built

Re: Open Challenges in LLM Research

#6
post #5

So far it's been ~8 months since ChatGPT started the (popular) LLM craze. I've found raw GPT to be useful for a lot of things, but have yet to see my most frequently used apps integrate it in a useful way. Maybe I'm using the wrong apps... It'll be interesting to see what improvements (in a lab or at a company) need to happen before most people use purpose-built LLMs (or behind the scenes LLM prompts) in the apps the…

I have helped making behind sense cases, one was to classify emails and redirect them to intended sides, second was quality monitoring of call center.

Re: Open Challenges in LLM Research

#7

Looking back in 25 years, the "Hallucination Problem" will sound a lot like the "Frame Problem" of the 1970s. Looking back, it's a bit absurd to say that GOFAI would've got to AGI if only the Frame Problem could be solved. But the important point is why that sounds so absurd. It doesn't sound absurd because we found out that the frame problem can't be solved; that's beside the point. It also doesn't sound absurd beca…

I don't know much about AI research but the idea of "measuring" hallucination definitely seems very loaded to me. Humans hallucinate too and I don't think we can measure that. It almost feels like "we need AGI in order to develop AGI".

Comparing human hallucinations with model “hallucinations” does not make sense to me.

Model hallucinations seems to me like a fancy way to call the model results that make no sense (ie blatant errors). Plus it makes the model more humanoid.

Re: Open Challenges in LLM Research

#8
post #3

Looking back in 25 years, the "Hallucination Problem" will sound a lot like the "Frame Problem" of the 1970s. Looking back, it's a bit absurd to say that GOFAI would've got to AGI if only the Frame Problem could be solved. But the important point is why that sounds so absurd. It doesn't sound absurd because we found out that the frame problem can't be solved; that's beside the point. It also doesn't sound absurd beca…

Does anyone think we would have AGI if only we could solve the hallucination problem?

There’s people who thought we could just wire up ChatGPT to a bunch of API calls and have AGI by now. Or some similar version of bootstrapping an LLM.

Re: Open Challenges in LLM Research

#9

Earlier quoted context omitted.

I don't know much about AI research but the idea of "measuring" hallucination definitely seems very loaded to me. Humans hallucinate too and I don't think we can measure that. It almost feels like "we need AGI in order to develop AGI".

Comparing human hallucinations with model “hallucinations” does not make sense to me. Model hallucinations seems to me like a fancy way to call the model results that make no sense (ie blatant errors). Plus it makes the model more humanoid.

Humans also very often produce results that don't make sense.

Re: Open Challenges in LLM Research

#10

Earlier quoted context omitted.

Comparing human hallucinations with model “hallucinations” does not make sense to me. Model hallucinations seems to me like a fancy way to call the model results that make no sense (ie blatant errors). Plus it makes the model more humanoid.

Humans also very often produce results that don't make sense.

Humans that produce output like LLMs are most likely to be diagnosed as schizophrenic, which I don't believe is the goal.
Post reply on HN