Earlier quoted context omitted.
Try asking a markov chain to think step by step
By definition a Markov chain does everything step by step already, no need to ask!
Sorry, but a new prompt for GPT-4 is not a paper
91–100 of 193 posts
Re: Sorry, but a new prompt for GPT-4 is not a paper
#92Earlier quoted context omitted.
If I book telescope time and capture a supernova then no one will ever be able to reproduce my raw results because it has already happened. I don't see why OpenAI pulling old model snapshots is any different.
> If I book telescope time and capture a supernova then no one will ever be able to reproduce my raw results because it has already happened. I don't see why OpenAI pulling old model snapshots is any different. That's why you capture multiple of them and verify your data statistically?
The problem is that what works on small LLMs does not necessarily scale to larger ones. See page 35 of [1] for example. A researcher only using the models of a few years ago (where the open models had [1] https://arxiv.org/pdf/2308.03296.pdf
Re: Sorry, but a new prompt for GPT-4 is not a paper
#93Excuse me? Step by step wasn't paper-worthy? Hard disagree. LLM research is currently in its infancy, because they are no older than a few years old. And a research field in its infancy is bound to have a few noteworthy "no sh*t, Sherlock" papers that would be obvious from hindsight. The fact is, LLMs are a higher-order construct in machine learning, much like a fish is higher-order than a simple cellular colony. Low…
Sorry, ignoramus here: Which paper is “Step by step”?
Re: Sorry, but a new prompt for GPT-4 is not a paper
#94Earlier quoted context omitted.
What is incorrect in this reference? You have not proposed any counterarguments. Also if you need just more fresh data - how do you propose to interpret the result of the Rosenhan's experiment?
Dont put people inside mental asylums when they are not ill?
Re: Sorry, but a new prompt for GPT-4 is not a paper
#95Excuse me? Step by step wasn't paper-worthy? Hard disagree. LLM research is currently in its infancy, because they are no older than a few years old. And a research field in its infancy is bound to have a few noteworthy "no sh*t, Sherlock" papers that would be obvious from hindsight. The fact is, LLMs are a higher-order construct in machine learning, much like a fish is higher-order than a simple cellular colony. Low…
Re: Sorry, but a new prompt for GPT-4 is not a paper
#96Excuse me? Step by step wasn't paper-worthy? Hard disagree. LLM research is currently in its infancy, because they are no older than a few years old. And a research field in its infancy is bound to have a few noteworthy "no sh*t, Sherlock" papers that would be obvious from hindsight. The fact is, LLMs are a higher-order construct in machine learning, much like a fish is higher-order than a simple cellular colony. Low…
I feel like the author of this tweet wasn’t saying step-by-step isn’t worthy, he was saying that non-reproducible results are not science. He emphasizes this twice in that tweet: > one experiment on one data set with seed picking is not worthy reporting > Additionally, we all need to understand this is just one good empirical result, now we need to make it useful…
And while I obviously value very much the engineering advances we have seen, the science is still lacking, because not enough people are trying to understand why these things are happening. Although engineering advances are important and valuable, I don't understand exactly why people try so hard to call themselves scientists if they are basically skipping the scientific process entirely.
Re: Sorry, but a new prompt for GPT-4 is not a paper
#97Earlier quoted context omitted.
> People overestimate the value of "grand developments", and underestimate the value of actually knowing - in this case actually knowing how well something works, even if it is as simple as a prompt. I think this depends a lot on the "culture" of the subject area. For example in mathematics, it is common that only new results that have been thoroughly worked through are typically "publish-worthy".
Wouldn’t the “thoroughly worked through” part be analogous to extensive measurements of a prompt?
- hypothesis building
- experimental design
- doing experiments
- analyzing the experimental results
- doing new experiments
- analyzing in which sense the collected data support the hypothesis or not
- ...
work.
Re: Sorry, but a new prompt for GPT-4 is not a paper
#98Re: Sorry, but a new prompt for GPT-4 is not a paper
#99Earlier quoted context omitted.
>> LLM research is currently in its infancy Everything that went in to creating GPT4 is AI/science or whatever. Probing GPT4 and trying to understand and characterize it is also a very worthy thing to do - else how can it be improved upon? But if making GPT is science, I'd say this stuff is more akin to psychology ;-)
Machine psychology
Re: Sorry, but a new prompt for GPT-4 is not a paper
#100If you do enough measurements on that new prompt then I don't see why this shouldn't be a paper. People overestimate the value of "grand developments", and underestimate the value of actually knowing - in this case actually knowing how well something works, even if it is as simple as a prompt. Compare with drug trials: Adderall only differs from regular amphetamine in the relative concentration of enantiomers, and th…
If you do the rigor on why something really is interesting, publish it.