Earlier quoted context omitted.
GPT4 solves this problem In any combination easily. What do so many posters seem to claim to have stumped it?
Try this: There's this person standing in a field, and with them is a balloon, a vacuum cleaner, and a magical creature of unknown origin. They need to get across to the woods at the end of the field, and do so safely. They can only go together: they get very, extremely lonely if they do not travel together, and they will not be safe because of this loneliness. If left together, the baloon would suck up the vacuum cl…
OpenAI researchers warned board of AI breakthrough ahead of CEO ouster
711–720 of 1001 posts
Re: OpenAI researchers warned board of AI breakthrough ahead of CEO ouster
#712Remember, about a month ago Sam posted a comment along the lines of "AI will be capable of superhuman persuasion well before it is superhuman at general intelligence, which may lead to very strange outcomes". The board was likely spooked by the recent breakthroughs (which were most likely achieved by combining transformers with another approach), and hit the panic button. Anything capable of "superhuman persuasion",…
Re: OpenAI researchers warned board of AI breakthrough ahead of CEO ouster
#713I feel very comfortable saying, as a mathematician, that the ability to solve grade school maths problems would not be at all a predictor of ability to solve real mathematical problems at a research level. The reason LLMs fail at solving mathematical problems is because: 1) they are terrible at arithmetic, 2) they are terrible at algebra, but most importantly, 3) they are terrible at complex reasoning (more specifica…
Re: OpenAI researchers warned board of AI breakthrough ahead of CEO ouster
#714This would have major implications for long-term planning. For instance, if you have a sequence of 10 steps, each with a 90% success rate, the overall success rate after all 10 steps falls to just 34%. This is one of the reasons why agents like AutoGPT often fail in complex tasks.
Re: OpenAI researchers warned board of AI breakthrough ahead of CEO ouster
#715Re: OpenAI researchers warned board of AI breakthrough ahead of CEO ouster
#716Earlier quoted context omitted.
I don't know for Q* of course, but all the tests I made with GPT4, and all what I've read and seen about it, show that it is unable to reason. It was trained with an unfathomable amount of data, so it can simulate reasoning very well, but it is unable to reason
What is the difference between simulating reasoning very well and "actual" reasoning?
As a simple example that you can replicate using chatgpt, ask it to solve some simple maths problem. Very frequently you will get a solution that looks like reasoning but is not, and reveals that it does not have an actual model of the underlying maths but is in fact doing text prediction based on a history of maths. For example see here[1]. I ask it for some quadratics in x with some specification on the number of roots. It gives me what looks at first glance like a decent answer. Then I ask the same exact question but asking for quadratics in x and y[2]. Again the answer looks plausible except that for the solution "with one real root" it says the solution has one real root when x + y =1. Well there are infinite real values for x and y such that x + y =1, not one real root. It looks like it has solved the problem but instead it has simulated the solving of the problem.
Likewise stacking problems, used to check for whether an AI has a model of the world. This is covered in "From task structures to world models: What do LLMs know?"[3] but for example here[4] I ask it whether it's easier to balance a barrel on a plank or a plank on a barrel. The model says it's easier to balance a plank on a barrel with an output text that simulates reasoning discussing center of mass and the difference between the flatness of the plank and the tendency of the barrel to roll because of its curvature. Actual reasoning would say to put the barrel on its end so it doesn't roll (whether you put the plank on top or not).
[1] https://chat.openai.com/share/64556be8-ad20-41aa-99af-ed5a42...
[2] https://chat.openai.com/share/2cd39197-dc09-4d07-a0d6-6cd800...
[3] https://arxiv.org/abs/2310.04276
[4] https://chat.openai.com/share/4b631a92-0d55-4ae5-8892-9be025...
Re: OpenAI researchers warned board of AI breakthrough ahead of CEO ouster
#717There will come a day when 50% of jobs are being done by AI, major decisions are being made by AI, we're all riding around in cars driven by AI, people are having romantic relationships with AI... and we'll STILL be debating whether what has been created is really AGI. AGI will forever be the next threshold, then the next, then the next until one day we'll realize that we passed the line years before.
> day when 50% of jobs are being done by AI By OpenAI definition 50% is not enough to qualify for AGI, it has to be "nearly any economically valuable work"
Re: OpenAI researchers warned board of AI breakthrough ahead of CEO ouster
#718Earlier quoted context omitted.
Teachers use websites to try and detect if AI wrote essays (and often it gets it wrong, and they believe it) we've defacto passed it.
Turing test is not do AI sound like humans some of the time, but is it possible to tell an AI is AI just by speaking with it. The answer is definitely yes, but it's not by casual conversation, but by asking weird logic problems it has tremendous problems solving and will give totally nonsensical inhuman answers to.
Re: OpenAI researchers warned board of AI breakthrough ahead of CEO ouster
#719Earlier quoted context omitted.
> day when 50% of jobs are being done by AI By OpenAI definition 50% is not enough to qualify for AGI, it has to be "nearly any economically valuable work"
Not sure it has to replace plumbers to be AGI
Re: OpenAI researchers warned board of AI breakthrough ahead of CEO ouster
#720Interestingly, I experience more anxiety from the thought of being made irrelevant than from the prospect of complete human extinction. I guess this can be interpreted as either vanity or stupidity, but I do think it illustrates how important it is for some humans to maintain their position in the social hierarchy.
If there's a referendum between two government policies, the first that every single person had to publicly speak in front of at least ten strangers once a year, that policy would be terrifying and bad to people who don't like public speaking. If the second policy was that every single person should be killed, that might be scary but it's not really as viscerally scary as the forced public speaking madman, at least to a lot of people, and it's also so bad that we have a natural impulse to just reject it as possible.
Nevertheless, if we recognise these impulses in ourselves we can attempt to adjust for them and tick the right box on the imaginary referendum, because even though public speaking is really bad and scary, it's still better than everyone dying.