Live data from Hacker News

GPT-4

openai.com

291–300 of 1001 posts

Re: GPT-4

#291
post #218

A class of problem that GPT-4 appears to still really struggle with is variants of common puzzles. For example: >Suppose I have a cabbage, a goat and a lion, and I need to get them across a river. I have a boat that can only carry myself and a single other item. I am not allowed to leave the cabbage and lion alone together, and I am not allowed to leave the lion and goat alone together. How can I safely get all three…

I think you may have misstated the puzzle. It's ok to leave the lion and the cabbage together, assuming it's not a vegetarian lion.

He didn’t misstate the puzzle, the whole point is to give an alternative version of the puzzle, and GPT 4 doesn’t notice that alternative. It’s exactly as difficult as the standard version as long as you are doing the logic instead of pattern-matching the puzzle form to text.

Re: GPT-4

#292

Imagine ingesting the contents of the internet as though it's a perfect reflection of humanity, and then building that into a general purpose recommendation system. That's what this is Is the content on the internet what we should be basing our systematic thinking around? No, I think this is the lazy way to do it - by using commoncrawl you've enshrined the biases and values of the people who are commenting and provid…

[dead]

Re: GPT-4

#293

OpenAI is located in the same building as Musk's Neuralink. Can't wait for this to be implanted in babies at birth! https://www.youtube.com/watch?v=O2RIvJ1U7RE

[deleted]

Re: GPT-4

#294
post #218

A class of problem that GPT-4 appears to still really struggle with is variants of common puzzles. For example: >Suppose I have a cabbage, a goat and a lion, and I need to get them across a river. I have a boat that can only carry myself and a single other item. I am not allowed to leave the cabbage and lion alone together, and I am not allowed to leave the lion and goat alone together. How can I safely get all three…

LLMs aren’t reasoning about the puzzle. They’re predicting the most likely text to print out, based on the input and the model/training data.

If the solution is logical but unlikely (i.e. unseen in the training set and not mapped to an existing puzzle), then the probability of the puzzle answer appearing is very low.

Re: GPT-4

#295
post #86

> What are the implications for society when general thinking, reading, and writing becomes like Chess? I think going from LSAT to general thinking is still a very, very big leap. Passing exams is a really fascinating benchmark but by their nature these exams are limited in scope, have very clear assessment criteria and a lot of associated and easily categorized data (like example tests). General thought (particularl…

Even just in the exam passing category, GPT4 showed no improvement over GPT3.5 on AP Language & Composition or AP English Literature, and scored quite poorly.

Now, granted, plenty of humans don't score above a 2 on those exams either. But I think it's indicative that there's still plenty of progress left to make before this technology is indistinguishable from magic.

Re: GPT-4

#296
post #86

> What are the implications for society when general thinking, reading, and writing becomes like Chess? I think going from LSAT to general thinking is still a very, very big leap. Passing exams is a really fascinating benchmark but by their nature these exams are limited in scope, have very clear assessment criteria and a lot of associated and easily categorized data (like example tests). General thought (particularl…

General thinking requires an AGI, which GPT-4 is not. But it can already have a major impact. Unlike self-driving cars which we require 99.999+% safety to be deployed widely, people already use the imperfect GPT-3 and ChatGPT for many productive tasks.

Driving as well as an attentive human in real time, in all conditions, probably requires AGI as well.

GPT-4 is not an AGI and GPT-5 might not be it yet. But the barriers toward it are getting thinner and thinner. Are we really ready for AGI in a plausibly-within-our-lifetime future?

Sam Altman wrote that AGI is a top potential explanation for the Fermi Paradox. If that were remotely true, we should be doing 10x-100x work on AI Alignment research.

Re: GPT-4

#297
GPT is a better scraper/parser. It’s interesting but I don’t understand why people are acting like this is the second coming.

Re: GPT-4

#298

> Yes, you can send me an image as long as it's in a supported format such as JPEG, PNG, or GIF. Please note that as an AI language model, I am not able to visually process images like a human would. However, I can still provide guidance or advice on the content of the image or answer any questions you might have related to it. Fair, but if it can analyze linked image, I would expect it to be able to tell me what tex…

> Image inputs are still a research preview and not publicly available.

Re: GPT-4

#300

Test taking will change. In the future I could see the student engaging in a conversation with an AI and the AI producing an evaluation. This conversation may be focused on a single subject, or more likely range over many fields and ideas. And may stretch out over months. Eventually teaching and scoring could also be integrated as the AI becomes a life-long tutor. Even in a future where human testing/learning is no l…

There was blog post on HN recently about the upbringings of great scientists, physicists, polymaths, etc. They almost invariably had access to near unlimited time with high quality tutors. He cited a source that claimed modern students who had access to significant tutoring resources were very likely to be at the top of their class.

Personalized learning is highly effective. I think your idea is an exciting one indeed.

Post reply on HN