Live data from Hacker News

GPT-4 details leaked?

threadreaderapp.com

491–500 of 648 posts

Re: GPT-4 details leaked?

#491

Can anyone provide an alternative link to https://twitter.com/i/web/status/1678545170508267522 I haven't registered for Twitter since it started and I'd rather not now (though I probably will if it's the only way to get leaked gpt4 training details)

Wayback failed to load the subtweets but archive.is has a copy but it seems to stop after around 10 subtweets. The threader link that was posted has it all though.

https://archive.is/Y72Gu

Re: GPT-4 details leaked?

#492

Earlier quoted context omitted.

How do you know? do you have insider knowledge of this or is it just based on what they share publically?

From what I see in that GPT 3 and 4 was a bit of a rugpull for the industry, now we're all laughing at Google because seemingly they had their hands on the rug for nearly a decade and did nothing - but from the other perspective, maybe they saw the future openai has now brought us and decided against being the pioneers

Amazingly enough, I think this is a bit of it. Some powerful enough people at Google became concerned about implications, including around "hallucinations", "poisoning", etc., and decided to put this sort of research on something of a backburner - justified, in part, by a lack of some obvious easy interfacing of this with search (scaling, hallucinations, etc.).

Of course, the 'wonderful' thing about humans / "independent agents with survival drives in competitive game-theoretic type scenarios" is: if enough people / "agents" have access / opportunity, someone WILL "push the button".

It's just delicious ... the same kinds of patterns over and over - "oh, we should really do something about X / nobody should have power like X, ... but, there's no stopping it, ... oh well".

And, the "rules" really are subtly, many levels down, in place, to make it apparently impossible to not get trapped, one way or another.

(Anyway... [Cartman voice] Screw you guys, I'm going to my other planet...)

Re: GPT-4 details leaked?

#493

Earlier quoted context omitted.

Everybody is self-soothing with the idea that OpenAI's (frankly, half hearted) push for regulation is just mundane regulatory capture and profit seeking, and not the fact that it will, at best, absolutely destroy everything about the internet and technology that we've come to love and know. Should a 4chan torrent show up like LLaMA, with weights and code for a base GPT4-level model, modern society is done. Golden age…

From your perspective, how would modern society be "done" if GPT-4 was generally available? How would it be substantially different from LLaMA?

GPT-4 is far more capable than LLaMA. Just as one area of impact - captchas would become permanently ineffective. If you're experienced in developing captchas and everything they do for us, you know the implications of that alone lead to a very dystopian internet and world.

I like to answer a question with a question: if you sit and think about it, what both unintentional misuses and intentional abuses can you think of? It helps to write down a list of known abilities, then thinking up several "what if..." negative utilities or implications of each, then iterating further to see second, third, fourth order effects.

Re: GPT-4 details leaked?

#494

Earlier quoted context omitted.

To me, you're describing the differences between a cook and a chef. Just because you've built something with well known methods doesn't make you a scientist. You didn't come up with anything new. In fact, we're starting to sound a lot like Apple. Apple is (in)famous for taking ideas that someone else did all of the hard work of developing and proving to work, and then take various ideas like that to combine into an a…

semantics, I say. cooks are "chiefs" of certain things otherwise they'd have no value on a team. unless you're saying that chefs setting the menu means theyre more of a scientist. at worst you're saying e.g. scientists are like chefs in that their used their creativity to make up the standard model. at best you're saying chefs are theorists and cooks are experimental scientists which is , again, to my own point.

i'm saying that chefs are more likely to be aware of what is happening when doing the whipping, baking, cooling, etc and why that's important vs just doing it as a step. so when they experiment with the menus, there's a bit more than basic understanding in why they think the experiments might work. cooks are just the kids in chemistry class following the directions, and depending on how well they follow the directions (and to a point how well the instructions are written), they might not blow up the lab (or burn the m.f. soufflé)

Re: GPT-4 details leaked?

#495

Earlier quoted context omitted.

From your perspective, how would modern society be "done" if GPT-4 was generally available? How would it be substantially different from LLaMA?

GPT-4 is far more capable than LLaMA. Just as one area of impact - captchas would become permanently ineffective. If you're experienced in developing captchas and everything they do for us, you know the implications of that alone lead to a very dystopian internet and world. I like to answer a question with a question: if you sit and think about it, what both unintentional misuses and intentional abuses can you think…

> If you're experienced in developing captchas and everything they do for us

What they have done, fairly overtly for a long time, is train AI to defeat captchas.

That this was self-limiting was somewhat obvious.

Re: GPT-4 details leaked?

#496
post #123

Earlier quoted context omitted.

You can run that model (Wizard-30) on a computer with 64 gigabytes of RAM (or smaller, I don't know how tight you can cut it). You obviously want fast RAM and a good CPU, but you don't need a GPU.

You can also travel on a bike from NY to LA.

>You can also travel on a bike from NY to LA.

You can. In fact, my brother did so a bunch of years ago. He found it to be a wonderful experience that made his life better.

He's also flown on a commercial airplane from NY to LA (as have I, as well as millions of others) and while it got him to Los Angeles, it didn't provide the levels of sensory input, personal interactions and experience that riding his bicycle did.

That's not to say everyone should ride bicycles across the US every time they need/want to make such a trip, but doing so at least once can be a more positive experience than sitting next to some strangers for five hours.

The satisfaction of doing so, or the experiences in interacting with people and the landscape during such a trip aren't quantifiable, but reducing the value of doing so (if I'm missing your point here, my apologies) to the time required to make such a trip is reductive in the extreme IMHO.

Edit: Clarified my prose.

Re: GPT-4 details leaked?

#497

Earlier quoted context omitted.

> the fact that I cannot freely read an article published 53 years ago is beyond ludicrous And has as much to do with individual versus corporate power as the health of a single tree can speak for a forest. Nobody is saying the current situation is good or even sustainable. OP just made a big claim for which there is no evidence, despite many looking for it.

"there is no evidence" As quoted above, "Ithaka's total revenue was $105 million in 2019, most of it ($79 million) from JSTOR service fees". All those $79 million and more JSTOR has stolen from the authors. Although JSTOR should have been dissolved long before this for effectively killing Aaron Swartz. But yes, there is no justice in this world. I suppose that's my main contention, why pretend anymore, we live in a s…

This isn’t a cogent argument. You’re identifying troubling behaviour. But it’s not being stitched into anything cohesive. Hiding bad rhetoric behind post-modernist nihilism is in vogue, but unproductive.

Re: GPT-4 details leaked?

#498

Earlier quoted context omitted.

LLaMA 30B or 60B can be very impressive when correctly prompted. Deploying the 60B version is a challenge though and you might need to apply 4-bit quantization with something like https://github.com/PanQiWei/AutoGPTQ or https://github.com/qwopqwop200/GPTQ-for-LLaMa . Then you can improve the inference speed by using https://github.com/turboderp/exllama . If you prefer to use an "instruct" model à la ChatGPT (i.e. tha…

Just a reminder that LLaMA is not open—in order to use it legally you have to agree to Meta's terms, which currently means research use only. The versions circulating on torrents are essential pirated, and while I don't have an ethical problem with that at all you can't use it safely in a business. The open replacements for LLaMA have yet to reach 30B, let alone 65B.

Falcon is an open (Apache licensed) replacement for LLaMA, with a 40B version that's competitive with LLaMA 65B on benchmarks.

Re: GPT-4 details leaked?

#499

Earlier quoted context omitted.

Just add epicycles.

Add enough epicycles and you've got a Fourier transform... Epicycles were a great idea but applied for the wrong reason.

Oh, they worked really well; they were still used for numerical calculation even after Galileo. But they weren't a "demonstration of knowledge and understanding" of planetary motion.

Re: GPT-4 details leaked?

#500

Earlier quoted context omitted.

I'm tired of science as a religion. People treat it as some gospel, like if you check out some criterions you're suddenly "scientific" and instantly get a sense of validity and authority that you shouldn't logically get. I judge things as "what you can do", not "what can you predict". The only demonstration of knowledge and understanding is being able to do something. Not predict. Not "scientific method" and ridiculo…

> In the end of the day, you either manage to do something or you don't. The hard sciences would like a word.

Taleb wrote well about that, changed my view on science. Not that science should be discounted, but it’s not the single source of truth. https://twitter.com/nntaleb/status/1419843561286160397
Post reply on HN