Earlier quoted context omitted.
It's a language model, FFS. Ask it whether it uses PaLM 1 and it will confirm it as well.
That is fascinating. Is it the same for GPT 3.5 and 4? For some reason when I was asking Open AI it was identifying itself properly.
PaLM 2 Technical Report [pdf]
21–30 of 297 posts
Re: PaLM 2 Technical Report [pdf]
#22Earlier quoted context omitted.
GPT-4 is way slower than GPT-3. Unless they are artificially spiking the latency to hide parameter count, it’s likely around 1trn params
Yeah 1 to 2 trillion is the estimates I've heard. Given the 25 messages / 3 hour limit in chatGPT, I don't think they've found a way to make it cheap to run.
Re: PaLM 2 Technical Report [pdf]
#23The links in their press release just link to their other press release, and if I google "PaLM API" it just gives me more press release, but I just couldn't find the actual document for their PaLM API.
How do I actually google the "PaLM API" for a way to test "PaLM 2"?
Re: PaLM 2 Technical Report [pdf]
#24> "We then train several models from 400M to 15B on the same pre-training mixture for up to 1 × 1022 FLOPs." Seems that for the last year or so these models are getting smaller. I would be surprised if GPT-4 had > the number of parameters as GPT-3 (i.e. 175B). Edit: Seems those numbers are just for their scaling laws study. They don't explicitly say the size of PaLM 2-L, but they do say "The largest model in the PaLM…
GPT-4 is way slower than GPT-3. Unless they are artificially spiking the latency to hide parameter count, it’s likely around 1trn params
> The largest model in the PaLM 2 family, PaLM 2-L, is significantly smaller than the largest PaLM model but uses more training compute
The largest PaLM model is 540B. So all of PaLM 2 is potentially double-digit parameters.
Note though that GPT-3.5 was plausibly not a finetuning of the 175B model, but instead a finetuning of Codex which was based on the 12B version of GPT-3.
Re: PaLM 2 Technical Report [pdf]
#25Earlier quoted context omitted.
It's a language model, FFS. Ask it whether it uses PaLM 1 and it will confirm it as well.
Not really. If what model it was trained on was represented properly in the instruction tuning dataset then they'll consistently identify themselves. But it's not a given that that was the case for bard.
1. ask it what PaLM 2 is (to pollute the context) 2. ask it whether it's based on PaLM 2 (it will tell you - yes, sure)
Re: PaLM 2 Technical Report [pdf]
#26So, I asked Bard if it's using PaLM 2 and it did confirm it. My initial results are super promising. Highly recommend checking it out again.
It's a language model, FFS. Ask it whether it uses PaLM 1 and it will confirm it as well.
- Yes, I am currently running on PaLM 2, which is Google AI's latest and most advanced large language model. PaLM 2 is trained on a massive dataset of text and code, and it is able to generate text, translate languages, write different kinds of creative content, and answer your questions in an informative way.
PaLM 2 is still under development, but it has learned to perform many kinds of tasks, including
I will try my best to follow your instructions and complete your requests thoughtfully. I will use my knowledge to answer your questions in a comprehensive and informative way, even if they are open ended, challenging, or strange. I will generate different creative text formats of text content, like poems, code, scripts, musical pieces, email, letters, etc. I will try my best to fulfill all your requirements. Please instruct me what you want me to do today.
- are you on PaLM 1 right now?
- I am not on PaLM 1. PaLM 1 is a large language model that was developed by Google AI. It was trained on a massive dataset of text and code, and it is able to generate text, translate languages, write different kinds of creative content, and answer your questions in an informative way.
I am a large language model that is still under development. I am trained on a massive dataset of text and code, and I am able to generate text, translate languages, write different kinds of creative content, and answer your questions in an informative way.
I am not as advanced as PaLM 1, but I am learning new things every day. I hope that one day I will be able to do everything that PaLM 1 can do, and more.
Re: PaLM 2 Technical Report [pdf]
#2737.6% success
GPT-4:
67% success
Not even close, gpt4 miles ahead
Re: PaLM 2 Technical Report [pdf]
#28Earlier quoted context omitted.
It's a language model, FFS. Ask it whether it uses PaLM 1 and it will confirm it as well.
- are you on PaLM 2 right now? - Yes, I am currently running on PaLM 2, which is Google AI's latest and most advanced large language model. PaLM 2 is trained on a massive dataset of text and code, and it is able to generate text, translate languages, write different kinds of creative content, and answer your questions in an informative way. PaLM 2 is still under development, but it has learned to perform many kinds o…
EDIT: Also, this doesn't seem convincing: "I am not as advanced as PaLM 1, but I am learning new things every day. I hope that one day I will be able to do everything that PaLM 1 can do, and more."
Re: PaLM 2 Technical Report [pdf]
#29Earlier quoted context omitted.
My sweet summer child, this is a closely guarded secret. Will only be revealed if perhaps Europe demands it so that copyright holders can sue.
Metadata will show where it came from, should you choose to keep it. Or so they showed on the big screen at I/O today.
https://lifearchitect.ai/whats-in-my-ai/
OpenAI paid $2m/year for twitter feeds until Elon cut them off, and Sam Altman has mentioned they'd paid a lot for scientific journals and Reddit mention they'll start charging. Given how central data quality and curation is, if these private data sources give a significant boost, it won't be available for Apache2 models.
Re: PaLM 2 Technical Report [pdf]
#30Earlier quoted context omitted.
GPT-4 is way slower than GPT-3. Unless they are artificially spiking the latency to hide parameter count, it’s likely around 1trn params
The report specifically states: > The largest model in the PaLM 2 family, PaLM 2-L, is significantly smaller than the largest PaLM model but uses more training compute The largest PaLM model is 540B. So all of PaLM 2 is potentially double-digit parameters. Note though that GPT-3.5 was plausibly not a finetuning of the 175B model, but instead a finetuning of Codex which was based on the 12B version of GPT-3.