Live data from Hacker News

GPT-4

openai.com

481–490 of 1001 posts

Re: GPT-4

#481

Access is invite only for the API, and rate limited for paid GPT+. > gpt-4 has a context length of 8,192 tokens. We are also providing limited access to our 32,768–context (about 50 pages of text) version, gpt-4-32k, which will also be updated automatically over time (current version gpt-4-32k-0314, also supported until June 14). Pricing is $0.06 per 1K prompt tokens and $0.12 per 1k completion tokens. The context le…

Will any of the profits be shared with original authors whose work powers the model?

No.

Now that you have read my answer, you owe me $0.01 because your brain might use this information in the future.

Re: GPT-4

#482

Genuinely surprised by the positive reaction about how exciting this all is. You ever had to phone a large business to try and sort something out, like maybe a banking error, and been stuck going through some nonsense voice recognition menu tree that doesn't work? Well imagine chat GPT with a real time voice and maybe a fake, photorealistic 3D avatar and having to speak to that anytime you want to speak to a doctor,…

I wonder how popular will "AI veganism" be.

Re: GPT-4

#483
post #218

A class of problem that GPT-4 appears to still really struggle with is variants of common puzzles. For example: >Suppose I have a cabbage, a goat and a lion, and I need to get them across a river. I have a boat that can only carry myself and a single other item. I am not allowed to leave the cabbage and lion alone together, and I am not allowed to leave the lion and goat alone together. How can I safely get all three…

I also tested logic puzzles tweaked to avoid memorization. GPT3 did poorly, GPT4 got a few of them. I expect humans will still be useful until GPT6 solves all these problems.

Can you post your attempts? Would love to see it

Re: GPT-4

#484
post #86

> What are the implications for society when general thinking, reading, and writing becomes like Chess? I think going from LSAT to general thinking is still a very, very big leap. Passing exams is a really fascinating benchmark but by their nature these exams are limited in scope, have very clear assessment criteria and a lot of associated and easily categorized data (like example tests). General thought (particularl…

Your last paragraph weakens the argument that you’re making.

Driving assistance and the progress made there and large language models and the progress made there are absolutely incomparable.

The general public’s hype in driving assistance is fueled mostly by the hype surrounding one car maker and its figurehead and it’s a hype that’s been fueled for a few years and become accepted in the public, reflected in the stock price of that car maker.

Large language models have not yet perpetrated the public’s memory yet, and, what’s actually the point is that inside of language you can find our human culture. And inside a large language model you have essentially the English language with its embeddings. It is real, it is big, it is powerful, it is respectable research.

There’s nothing in driving assistance that can be compared to LLMs. They don’t have an embedding of the entire physical surface of planet earth or understanding of driving physics. They’re nothing.

Re: GPT-4

#485
post #218

A class of problem that GPT-4 appears to still really struggle with is variants of common puzzles. For example: >Suppose I have a cabbage, a goat and a lion, and I need to get them across a river. I have a boat that can only carry myself and a single other item. I am not allowed to leave the cabbage and lion alone together, and I am not allowed to leave the lion and goat alone together. How can I safely get all three…

A funny variation on this kind of over-fitting to common trick questions - if you ask it which weighs more, a pound of bricks or a pound of feathers, it will correctly explain that they actually weigh the same amount, one pound. But if you ask it which weighs more, two pounds of bricks or a pound of feathers, the question is similar enough to the trick question that it falls into the same thought process and contorts…

There is no "thought process". It's not thinking, it's simply generating text. This is reflected in the obviously thoughtless response you received.

Re: GPT-4

#486

Genuinely surprised by the positive reaction about how exciting this all is. You ever had to phone a large business to try and sort something out, like maybe a banking error, and been stuck going through some nonsense voice recognition menu tree that doesn't work? Well imagine chat GPT with a real time voice and maybe a fake, photorealistic 3D avatar and having to speak to that anytime you want to speak to a doctor,…

I don't share your concerns. If the difference between a good and a bad news article is whether a real person has written it, how can AI generated news prevail? If nobody can tell the difference, does it really matter who wrote the article?

Facts can be verified the same way they are right now. By reputation and reporting by trusted sources with eyes on the ground and verifiable evidence.

Regarding comments on news sites being spammed by AI: there are great ways to prove you are human already. You can do this using physical objects (think Yubikeys). I don't see any problems that would fundamentally break Captchas in the near future, although they will need to evolve like they always have.

Re: GPT-4

#487

Genuinely surprised by the positive reaction about how exciting this all is. You ever had to phone a large business to try and sort something out, like maybe a banking error, and been stuck going through some nonsense voice recognition menu tree that doesn't work? Well imagine chat GPT with a real time voice and maybe a fake, photorealistic 3D avatar and having to speak to that anytime you want to speak to a doctor,…

Yea, I'm about ready to start a neo-amish cult. Electronics and radios and 3D graphics are great fun, so I would want to set a cutoff date to ignore technology created after 2016 or so, really I draw the line at deterministic v. non-deterministic. If something behaves in a way that can't be predicted, I don't really want to have my civilization rely on it. Maybe an exception for cryptography and physics simulation, but computers that hallucinate I can do without.

Re: GPT-4

#488

Genuinely surprised by the positive reaction about how exciting this all is. You ever had to phone a large business to try and sort something out, like maybe a banking error, and been stuck going through some nonsense voice recognition menu tree that doesn't work? Well imagine chat GPT with a real time voice and maybe a fake, photorealistic 3D avatar and having to speak to that anytime you want to speak to a doctor,…

The availability of LLM may make it so bad that we do something (e.g. paid support, verified access, etc.) about these problems that have already existed (public relations fluff-piece articles, astroturfing, etc.), but to a smaller degree.

Re: GPT-4

#489
post #218

A class of problem that GPT-4 appears to still really struggle with is variants of common puzzles. For example: >Suppose I have a cabbage, a goat and a lion, and I need to get them across a river. I have a boat that can only carry myself and a single other item. I am not allowed to leave the cabbage and lion alone together, and I am not allowed to leave the lion and goat alone together. How can I safely get all three…

A funny variation on this kind of over-fitting to common trick questions - if you ask it which weighs more, a pound of bricks or a pound of feathers, it will correctly explain that they actually weigh the same amount, one pound. But if you ask it which weighs more, two pounds of bricks or a pound of feathers, the question is similar enough to the trick question that it falls into the same thought process and contorts…

But unlike most people it understands that even though an ounce of gold weighs more than an ounce of feathers a pound of gold weighs less than a pound of feathers.

(To be fair this is partly an obscure knowledge question, the kind of thing that maybe we should expect GPT to be good at.)

Re: GPT-4

#490

Genuinely surprised by the positive reaction about how exciting this all is. You ever had to phone a large business to try and sort something out, like maybe a banking error, and been stuck going through some nonsense voice recognition menu tree that doesn't work? Well imagine chat GPT with a real time voice and maybe a fake, photorealistic 3D avatar and having to speak to that anytime you want to speak to a doctor,…

Sources uncheckable? What sources! All the sources will just be AI generated, in the first place. Primary sources will be vanishingly small
Post reply on HN