Live data from Hacker News

GPT-4

openai.com

571–580 of 1001 posts

Re: GPT-4

#571
post #218

A class of problem that GPT-4 appears to still really struggle with is variants of common puzzles. For example: >Suppose I have a cabbage, a goat and a lion, and I need to get them across a river. I have a boat that can only carry myself and a single other item. I am not allowed to leave the cabbage and lion alone together, and I am not allowed to leave the lion and goat alone together. How can I safely get all three…

Silk silk silk silk silk silk.

What do cows drink?

Re: GPT-4

#572

https://cdn.openai.com/papers/gpt-4.pdf >Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar. At that point, why bother putting out a paper?

Given how humorous the name’s become, I wonder if they regret calling themselves OpenAI.

Re: GPT-4

#573

Access is invite only for the API, and rate limited for paid GPT+. > gpt-4 has a context length of 8,192 tokens. We are also providing limited access to our 32,768–context (about 50 pages of text) version, gpt-4-32k, which will also be updated automatically over time (current version gpt-4-32k-0314, also supported until June 14). Pricing is $0.06 per 1K prompt tokens and $0.12 per 1k completion tokens. The context le…

Will any of the profits be shared with original authors whose work powers the model?

The model is powered by math.

Re: GPT-4

#575

How many parameters does it have? Are there different versions like LLaMa?

We don't know, OpenAI refused to publish any details about the architecture in the technical report. We don't know parameters, we don't know depth, we don't know how exactly it's integrating image data (ViT-style maybe?), we don't even know anything about the training data. Right now it's a giant black box.

Re: GPT-4

#576

It fails on this one, a horse is 15 dollar, a chicken 1 dollar, a egg .25 dollar. I can spend a 100 and i want 100 items total, what is the solution

I spend already 30 minutes on it, and still no solution.

Re: GPT-4

#577
I just finished reading the 'paper' and I'm astonished that they aren't even publishing the # of parameters or even a vague outline of the architecture changes. It feels like such a slap in the face to all the academic AI researchers that their work is built off over the years, to just say 'yeah we're not telling you how any of this is possible because reasons'. Not even the damned parameter count. Christ.

Re: GPT-4

#578

The comments on this thread are proof of the AI effect: People will continually push the goal posts back as progress occurs. “Meh, it’s just a fancy word predictor. It’s not actually useful.” “Boring, it’s just memorizing answers. And it scored in the lowest percentile anyways”. “Sure, it’s in the top percentile now but honestly are those tests that hard? Besides, it can’t do anything with images.” “Ok, it takes imag…

I’m one of these skeptics, but it’s not moving the goalposts. These goalposts are already there, in some sort of serial order that we expect them to be reached. It is good that when tech like this satisfied one of the easier/earlier goalposts, that skeptics refine our criticism based on evidence.

You will see skepticism until it is ubiquitous; for example, Tesla tech - it’s iterative and there are still skeptics about its current implementation.

Re: GPT-4

#579

From the paper: > Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar. I'm curious whether they have continued to scale up model size/compute significantly or if they have managed to make significant innovati…

Without paper and architecture, GPT-4 (GPT-3+1) could be just a marketing gimmick to upsell it and in reality it is just microservices of existing A.I models working together as AIaaS (A.I. as a service)

At this point, if it goes from being in the bottom 10% on a simulated bar exam to top 10% on a simulated bar exam, then who cares if that's all they're doing???
Post reply on HN