Live data from Hacker News

OpenAI's GPT-3 may be the biggest thing since Bitcoin

maraoz.com

181–190 of 554 posts

Re: OpenAI's GPT-3 may be the biggest thing since Bitcoin

#181
post #53

I am deeply enjoying this comment thread - it's a bit of a Barium Meal [0] for determining how many people read (a) the headline, (b) the first paragraph, or (c) the whole thing before jumping straight into the compose box. Having read to the bottom, the quality of text generation there absolutely blew me away. GPT-2 texts have a somewhat disconnected quality - "it only makes sense if you're not really paying attenti…

GPT-3 is a neat party trick. But the things that'll be done with web archives* in the next 20y will make it look like the PDP-8. ~love, a web archivist * GPT-3 is trained on one

Hopefully in a way that secures some funding for those making archives of the web.

Re: OpenAI's GPT-3 may be the biggest thing since Bitcoin

#182
post #137

I am deeply enjoying this comment thread - it's a bit of a Barium Meal [0] for determining how many people read (a) the headline, (b) the first paragraph, or (c) the whole thing before jumping straight into the compose box. Having read to the bottom, the quality of text generation there absolutely blew me away. GPT-2 texts have a somewhat disconnected quality - "it only makes sense if you're not really paying attenti…

I read the beginning, went "what the fuck is this guy on about? Get to the point" and then came to check the comments to if this was anything interesting or worth reading, saw your comment, and skimmed the end bit. Overall I'm pleased with my process as its an efficient way to find out which articles are worth reading. But it was also clear to me that the author had difficulty making a clear point or had a goal in hi…

From your comment it's not clear to me if you realize the author of the article is GPT-3.

Re: OpenAI's GPT-3 may be the biggest thing since Bitcoin

#185
post #147

Earlier quoted context omitted.

I've started working on a version of GPT-2 which generates English text. The purpose of this is to improve its ability to predict the next character in a text, by having it learn 'grammatical rules' for English. It already works well for predicting the next character when it has seen only a small amount of text, but becomes less accurate as the amount of training text increases. I have managed to improve this by havi…

The thing that kills me is that to the vast majority of human beings the nonsensical technobabble above is probably indistinguishable from real, honest, logically consistent technobabble.[a] Soon enough, someone will replicate the Sokal hoax[b] with GPT-3 or another state-of-the-art language-generation model. It's not hard to imagine GPT-3 writing a fake paper that gets published in certain academic journals in the s…

If he didn't put in the last line ("plot twist ...") I'm pretty sure no one here on HN would have guessed it.

In fact, while reading that comment I started to wonder why no one has tried to use GPT to generate text one character at a time. Or if someone has, what are the advantages and disadvantages over the BPE approach.

Re: OpenAI's GPT-3 may be the biggest thing since Bitcoin

#186
post #169

Earlier quoted context omitted.

> no-one can train GPT-3 in any acceptable time except Gwern, who has spent probably as much time with GPT-3 and GPT-2 as any 'amateur' out there, is publicly out there saying that for most use cases, GPT-3 + creative use of prompts gets you better results than GPT-2 with finetuning . That's an amazing capability that Gwern elaborates on more here: > A new programming paradigm? The GPT-3 neural network is so large a…

I don't really understand your post (and don't know much about GPT3) are you suggesting that the model is stateful in that it can continue learning from successive prompts? Maybe you can elaborate on this: >instead, you interact with it, expressing any task in terms of natural language descriptions, requests, and examples, tweaking the prompt until it “understands” & it meta-learns the new task based on the high-leve…

I don't think the claim is that the model is "stateful" in that it continues to learn from prompts. I think it's that the model no longer requires retraining for different situations; instead it has "learned" a set of lower (higher?) level abstractions from which those same (and possibly new) situations can be constructed dynamically from the input prompt.

Re: OpenAI's GPT-3 may be the biggest thing since Bitcoin

#187
post #172

Earlier quoted context omitted.

> I think you are underestimating what an advance these models are over previous NLP models in terms of quality. Yeah I mean, I agree. But in my opinion, it's a case of "doing the wrong thing right" instead of a more useful "doing the right thing wrong." I grant that these automated models are useful for low-value classification/generation tasks at high-frequency scale. I don't think that in any way is related to int…

This kind of facile moving the goalposts is imho a cheap shot and (not imho, fact) is a recurring phenomenon as we make incremental progress toward AI. Progress is often made with steps that would have been astonishing a few years ago. And every time the bar is raised higher. Rightly so, but characterizing this as doing the wrong thing is missing the point of what we, and the system, are learning. Yes it's not intell…

> This kind of facile moving the goalposts is imho a cheap shot and (not imho, fact) is a recurring phenomenon as we make incremental progress toward AI.

I think you missed my point. I think we're going in the wrong direction for AI entirely, and these "advances" are fundamentally misguided. OpenAI is explicitly about "intelligence," and so we should question if this is in fact that.

It's clear that humans have fundamental intelligence much better than all of this stuff with 6 orders of magnitude less input (at least of the same data sort) on a problem.

Perhaps it would be better to say, "I think the ML winter is just around the corner" as opposed to "the AI winter is just around the corner." That said, this really is math, and these algos still don't actually do anything resembling true intelligence.

Re: OpenAI's GPT-3 may be the biggest thing since Bitcoin

#188
post #137

Earlier quoted context omitted.

I read the beginning, went "what the fuck is this guy on about? Get to the point" and then came to check the comments to if this was anything interesting or worth reading, saw your comment, and skimmed the end bit. Overall I'm pleased with my process as its an efficient way to find out which articles are worth reading. But it was also clear to me that the author had difficulty making a clear point or had a goal in hi…

From your comment it's not clear to me if you realize the author of the article is GPT-3.

It's very clear to me that he does not. But he does an excellent job of making GP's point.

Re: OpenAI's GPT-3 may be the biggest thing since Bitcoin

#189
post #2

> I further predict that this will spark a creative gold rush among talented amateurs to train similar models and adapt them to a variety of purposes, including: mock news, “researched journalism”, advertising, politics, and propaganda. The first mention of 'Elon Musk' (who left the board) and this sentence alone gave me the tip-off that GPT-3 had generated that (and the whole blog) and it's following prediction make…

> no-one can train GPT-3 in any acceptable time except for those with access to large GPU/ASIC compute power (OpenAI, Microsoft, Google, NVIDIA, etc.) Any state actor has access to large compute power

I thought it cost about $10 million to train. Honestly, seems fairly cheap all things considered.

If it could be 100x smarter for only 100x more (whatever handwavey thing that really means) it would be a steal considering how the same model could be reused by thousands of companies without retraining.

Re: OpenAI's GPT-3 may be the biggest thing since Bitcoin

#190
post #157

Earlier quoted context omitted.

Can you elaborate on how you are having GPT2 contribute to your comments? What is your process?

I write a prompt, often copying text from the articles or other comments, and have it generate a lot of completions. I skim over the completions and grab interesting parts. For example, complaining about "quoted text" in my above comment was GPT2's suggestion (and also an actual issue with GPT2 which it was exhibiting by producing that text). Actually, all the text above from "There are a few" and beyond were written…

> Some authors have taken psychedelic drugs to enhance their creative processor, with GPT2 it is your word-processor that takes the drugs.

I love this quote. I hope it wasn't written by GPT-2.

Post reply on HN