Live data from Hacker News

What we still don’t know about how A.I. is trained

newyorker.com

1–10 of 211 posts

Re: What we still don’t know about how A.I. is trained

#2
The author is right we know almost nothing about the design and training of GPT-4.

From the technical report https://cdn.openai.com/papers/gpt-4.pdf : "Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar."

Re: What we still don’t know about how A.I. is trained

#4
Say I'm a conspiracy theorist but I'm calling it, that the pentagon isn't letting OpenAI tell the details of GPT-4 even if they wanted to (which they don't except for some of the researchers probably). National security, export restrictions, munitions classification, new Manhattan project etc.

EDIT: I know that most people think it's unlikely and I can't give any direct evidence for it. Does that mean it's not welcome to say on hacker news?

Re: What we still don’t know about how A.I. is trained

#5
post #4

Say I'm a conspiracy theorist but I'm calling it, that the pentagon isn't letting OpenAI tell the details of GPT-4 even if they wanted to (which they don't except for some of the researchers probably). National security, export restrictions, munitions classification, new Manhattan project etc. EDIT: I know that most people think it's unlikely and I can't give any direct evidence for it. Does that mean it's not welcom…

It's far more likely that openai have just been building hype by showing off the earlier models, and now they're shutting things down so they can monetize their IP.

Re: What we still don’t know about how A.I. is trained

#6
post #4

Say I'm a conspiracy theorist but I'm calling it, that the pentagon isn't letting OpenAI tell the details of GPT-4 even if they wanted to (which they don't except for some of the researchers probably). National security, export restrictions, munitions classification, new Manhattan project etc. EDIT: I know that most people think it's unlikely and I can't give any direct evidence for it. Does that mean it's not welcom…

It's far more likely that openai have just been building hype by showing off the earlier models, and now they're shutting things down so they can monetize their IP.

Why not both?

Re: What we still don’t know about how A.I. is trained

#7
post #3

[flagged]

You can oppose literally every technology by saying "oh it'll use energy - won't someone think of climate change?"

The copy+paste of "CO2 emissions" and "energy" has been one of the most successful petroleum company propaganda coups in history.

Re: What we still don’t know about how A.I. is trained

#8
I don’t expect them to understand how to use the tech as it isn’t their domain but any article describing the ChatGPT website and mistaking it for GPT and its APIs is just missing the point.

The eve date of its training isn’t too important. That whatever sources it was fed weren’t complete enough to write an accurate book report on an obscure, modern book us t too important. Etc.

The important part is the statistical model that has been built and that you can use by intelligently providing a context for it to work within. Feed it your book in context and see how it goes.

Re: What we still don’t know about how A.I. is trained

#9
post #4

Say I'm a conspiracy theorist but I'm calling it, that the pentagon isn't letting OpenAI tell the details of GPT-4 even if they wanted to (which they don't except for some of the researchers probably). National security, export restrictions, munitions classification, new Manhattan project etc. EDIT: I know that most people think it's unlikely and I can't give any direct evidence for it. Does that mean it's not welcom…

Try as we might, some people still use the downvote button as a "disagree" button...

Re: What we still don’t know about how A.I. is trained

#10
post #2

The author is right we know almost nothing about the design and training of GPT-4. From the technical report https://cdn.openai.com/papers/gpt-4.pdf : "Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar."

It’s amazing how one can found a nonprofit with a goal of conducting “open” research and then end up publishing something like this a couple of years later. Greed is good I guess.
Post reply on HN