Can anyone provide an alternative link to https://twitter.com/i/web/status/1678545170508267522 I haven't registered for Twitter since it started and I'd rather not now (though I probably will if it's the only way to get leaked gpt4 training details)
GPT-4 details leaked?
491–500 of 648 posts
Re: GPT-4 details leaked?
#492Earlier quoted context omitted.
How do you know? do you have insider knowledge of this or is it just based on what they share publically?
From what I see in that GPT 3 and 4 was a bit of a rugpull for the industry, now we're all laughing at Google because seemingly they had their hands on the rug for nearly a decade and did nothing - but from the other perspective, maybe they saw the future openai has now brought us and decided against being the pioneers
Of course, the 'wonderful' thing about humans / "independent agents with survival drives in competitive game-theoretic type scenarios" is: if enough people / "agents" have access / opportunity, someone WILL "push the button".
It's just delicious ... the same kinds of patterns over and over - "oh, we should really do something about X / nobody should have power like X, ... but, there's no stopping it, ... oh well".
And, the "rules" really are subtly, many levels down, in place, to make it apparently impossible to not get trapped, one way or another.
(Anyway... [Cartman voice] Screw you guys, I'm going to my other planet...)
Re: GPT-4 details leaked?
#493Earlier quoted context omitted.
Everybody is self-soothing with the idea that OpenAI's (frankly, half hearted) push for regulation is just mundane regulatory capture and profit seeking, and not the fact that it will, at best, absolutely destroy everything about the internet and technology that we've come to love and know. Should a 4chan torrent show up like LLaMA, with weights and code for a base GPT4-level model, modern society is done. Golden age…
From your perspective, how would modern society be "done" if GPT-4 was generally available? How would it be substantially different from LLaMA?
I like to answer a question with a question: if you sit and think about it, what both unintentional misuses and intentional abuses can you think of? It helps to write down a list of known abilities, then thinking up several "what if..." negative utilities or implications of each, then iterating further to see second, third, fourth order effects.
Re: GPT-4 details leaked?
#494Earlier quoted context omitted.
To me, you're describing the differences between a cook and a chef. Just because you've built something with well known methods doesn't make you a scientist. You didn't come up with anything new. In fact, we're starting to sound a lot like Apple. Apple is (in)famous for taking ideas that someone else did all of the hard work of developing and proving to work, and then take various ideas like that to combine into an a…
semantics, I say. cooks are "chiefs" of certain things otherwise they'd have no value on a team. unless you're saying that chefs setting the menu means theyre more of a scientist. at worst you're saying e.g. scientists are like chefs in that their used their creativity to make up the standard model. at best you're saying chefs are theorists and cooks are experimental scientists which is , again, to my own point.
Re: GPT-4 details leaked?
#495Earlier quoted context omitted.
From your perspective, how would modern society be "done" if GPT-4 was generally available? How would it be substantially different from LLaMA?
GPT-4 is far more capable than LLaMA. Just as one area of impact - captchas would become permanently ineffective. If you're experienced in developing captchas and everything they do for us, you know the implications of that alone lead to a very dystopian internet and world. I like to answer a question with a question: if you sit and think about it, what both unintentional misuses and intentional abuses can you think…
What they have done, fairly overtly for a long time, is train AI to defeat captchas.
That this was self-limiting was somewhat obvious.
Re: GPT-4 details leaked?
#496Earlier quoted context omitted.
You can run that model (Wizard-30) on a computer with 64 gigabytes of RAM (or smaller, I don't know how tight you can cut it). You obviously want fast RAM and a good CPU, but you don't need a GPU.
You can also travel on a bike from NY to LA.
You can. In fact, my brother did so a bunch of years ago. He found it to be a wonderful experience that made his life better.
He's also flown on a commercial airplane from NY to LA (as have I, as well as millions of others) and while it got him to Los Angeles, it didn't provide the levels of sensory input, personal interactions and experience that riding his bicycle did.
That's not to say everyone should ride bicycles across the US every time they need/want to make such a trip, but doing so at least once can be a more positive experience than sitting next to some strangers for five hours.
The satisfaction of doing so, or the experiences in interacting with people and the landscape during such a trip aren't quantifiable, but reducing the value of doing so (if I'm missing your point here, my apologies) to the time required to make such a trip is reductive in the extreme IMHO.
Edit: Clarified my prose.
Re: GPT-4 details leaked?
#497Earlier quoted context omitted.
> the fact that I cannot freely read an article published 53 years ago is beyond ludicrous And has as much to do with individual versus corporate power as the health of a single tree can speak for a forest. Nobody is saying the current situation is good or even sustainable. OP just made a big claim for which there is no evidence, despite many looking for it.
"there is no evidence" As quoted above, "Ithaka's total revenue was $105 million in 2019, most of it ($79 million) from JSTOR service fees". All those $79 million and more JSTOR has stolen from the authors. Although JSTOR should have been dissolved long before this for effectively killing Aaron Swartz. But yes, there is no justice in this world. I suppose that's my main contention, why pretend anymore, we live in a s…
Re: GPT-4 details leaked?
#498Earlier quoted context omitted.
LLaMA 30B or 60B can be very impressive when correctly prompted. Deploying the 60B version is a challenge though and you might need to apply 4-bit quantization with something like https://github.com/PanQiWei/AutoGPTQ or https://github.com/qwopqwop200/GPTQ-for-LLaMa . Then you can improve the inference speed by using https://github.com/turboderp/exllama . If you prefer to use an "instruct" model à la ChatGPT (i.e. tha…
Just a reminder that LLaMA is not open—in order to use it legally you have to agree to Meta's terms, which currently means research use only. The versions circulating on torrents are essential pirated, and while I don't have an ethical problem with that at all you can't use it safely in a business. The open replacements for LLaMA have yet to reach 30B, let alone 65B.
Re: GPT-4 details leaked?
#499Earlier quoted context omitted.
Just add epicycles.
Add enough epicycles and you've got a Fourier transform... Epicycles were a great idea but applied for the wrong reason.
Re: GPT-4 details leaked?
#500Earlier quoted context omitted.
I'm tired of science as a religion. People treat it as some gospel, like if you check out some criterions you're suddenly "scientific" and instantly get a sense of validity and authority that you shouldn't logically get. I judge things as "what you can do", not "what can you predict". The only demonstration of knowledge and understanding is being able to do something. Not predict. Not "scientific method" and ridiculo…
> In the end of the day, you either manage to do something or you don't. The hard sciences would like a word.