What is the point of all these different models? Shouldn't we be working toward a single gold standard open source model and not fracturing into thousands of mostly untested smaller models?
What's the point of inventing all these different materials? Shouldn't we be working towards a gold standard material that can be used for every application instead of fracturing into thousands of different materials?
Asking 60 LLMs a set of 20 questions
191–200 of 352 posts
Re: Asking 60 LLMs a set of 20 questions
#192Re: Asking 60 LLMs a set of 20 questions
#193Earlier quoted context omitted.
Using a temp of zero usually returns garbage results from most models, so it would likely do so in case of GPT 4 as well. Any other great ideas?
Temp of 0 gives the least random and most predictable results
Re: Asking 60 LLMs a set of 20 questions
#194Re: Asking 60 LLMs a set of 20 questions
#195Only tried chatGPT 3.5, but my god does it waffle on. Everything I ask ends with a paragraph saying "It's important to remember that..." like an after-school special from a 90s show. It can never just give you code, it has to say "Sure!, to {paraphase your question}, open a terminal...". It's interesting to see 20th century sci-fi depictions of this kind of AI/Search is being short and to the point. I guess they can'…
That's not GPT 3.5, that's ChatGPT. How waffly it gets depends on the context that was given to it by the people running ChatGPT; they likely told it to act as a helpful assistant and to give lots of information. If you run an LLM on your own, it's entirely possible to instruct it to be succinct.
Re: Asking 60 LLMs a set of 20 questions
#196Also, this page content would seem absolutely ridiculous just a few years ago.
Re: Asking 60 LLMs a set of 20 questions
#197Re: Asking 60 LLMs a set of 20 questions
#198Earlier quoted context omitted.
Also, MPT 7B gets it right over half the time. I've been testing every new LLM with that question. Also, I tend to include mention in the question that all siblings are from the same two parents to preclude half-siblings because half my friends have half-siblings from both sides scattered across the country; so the wrong answers actually do tend to apply to them sometimes.
> I've been testing every new LLM with that question We should pay more attention to data contamination when using popular prompts for testing.
Re: Asking 60 LLMs a set of 20 questions
#199Only tried chatGPT 3.5, but my god does it waffle on. Everything I ask ends with a paragraph saying "It's important to remember that..." like an after-school special from a 90s show. It can never just give you code, it has to say "Sure!, to {paraphase your question}, open a terminal...". It's interesting to see 20th century sci-fi depictions of this kind of AI/Search is being short and to the point. I guess they can'…
Basically, the LLM will formulate a better answer to the question if it talks itself through its reasoning process.
Re: Asking 60 LLMs a set of 20 questions
#200is anyone else feeling completely depressed and demotivated by how quickly this is happening?
No. When we were kids, my generation was promised flying cars, unlimited fusion power, and sentient computers. There's a good chance I'll live to see one out of three of those things happen, and that's better than the zero out of three I thought we'd get.