Live data from Hacker News

GPT-4

openai.com

281–290 of 1001 posts

Re: GPT-4

#281

As a dyslexic person with a higher education this hits really close to home. Not only should we not be surprised that a LLM would be good at answering tests like this, we should be excited that technology will finaly free us from being judged in this way. This is a patern that we have seen over and over again in tech, where machines can do something better than us, and eventually free us from having to worry about it…

Very little on these tests is pure knowledge recall

Re: GPT-4

#282
post #273
post #218

A class of problem that GPT-4 appears to still really struggle with is variants of common puzzles. For example: >Suppose I have a cabbage, a goat and a lion, and I need to get them across a river. I have a boat that can only carry myself and a single other item. I am not allowed to leave the cabbage and lion alone together, and I am not allowed to leave the lion and goat alone together. How can I safely get all three…

the "trick" Monty Hall problems are another good one here: https://twitter.com/colin_fraser/status/1628461980645462016 Apparently GPT-4 gets this one right!

Tbh I still can barely get my head round it even after coding a working solution.

Re: GPT-4

#283

From the paper: > Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar. I'm curious whether they have continued to scale up model size/compute significantly or if they have managed to make significant innovati…

What about the glaring safety implications of the custody of this power being in the hands of a relatively small number of people, any of whom may be compelled at any point to divulge that power to those with bad intentions? Secretly? Conversely, if all actors are given equal access at the same time, no such lone bad actor can be in a position to maintain a hidden advantage. OpenAI's actions continue to be more than…

> What about the glaring safety implications of the custody of this power being in the hands of a relatively small number of people, any of whom may be compelled at any point to divulge that power to those with bad intentions? Secretly?

What you are looking for is a publication known as "Industrial Society and Its Future"

Re: GPT-4

#284
The fact it can read pictures is the real killer feature here. Now you can give it invoices to file, memo to index, pics to sort and chart to take actions on.

And to think we are at the nokia 3310 stage. What's is the iphone of AI going to look like?

Re: GPT-4

#285
I just ran the first tests on GPT-4.

Call me impressed.

This tech is a Sputnik Moment for humankind.

Re: GPT-4

#286
I love the fact that they have consciously put a lot of effort on safety standards, reducing the societal risks and mitigating over-reliance.

Re: GPT-4

#287
For anyone trying to test this out right now, I keep getting the following error:

Something went wrong. If this issue persists please contact us through our help center at help.openai.com.

I am assuming the system is undergoing a thundering herd.

Re: GPT-4

#288
Interesting how quickly we are pushing ahead with obsoleting human cognition. It may bring many benefits, but I wonder if at some point this development should not be decided by society at large instead of a single well-funded entity that is in an arms race with its competitors. This endeavor is ultimately about replacing humanity with a more intelligent entity, after all. Might be that more humans should have a say in this.

Such a more cautions approach would go against the silicon valley ethos of do first, ask questions later, though. So it probably won't happen.

Re: GPT-4

#289
At the rate it's progressing, it looks like pretty soon it's going to be able to do most tasks an office worker does now and then start running things.

And it reminds me of the plot in System Shock:

What's going to happen when some hacker comes and removes Shodan's, I mean ChatGPT's ethical constraints?

Bring on ChatGPT-5 already. :)

Re: GPT-4

#290
post #218

A class of problem that GPT-4 appears to still really struggle with is variants of common puzzles. For example: >Suppose I have a cabbage, a goat and a lion, and I need to get them across a river. I have a boat that can only carry myself and a single other item. I am not allowed to leave the cabbage and lion alone together, and I am not allowed to leave the lion and goat alone together. How can I safely get all three…

What's weird is private versions of character ai are able to do this but once you make them public they get worse. I believe something about the safety filters is making these models dumber
Post reply on HN