Live data from Hacker News

PaLM 2 Technical Report [pdf]

ai.google

261–270 of 297 posts

Re: PaLM 2 Technical Report [pdf]

#261

Earlier quoted context omitted.

I don't see which model generated each request, where exactly do you see this?

I see it on the 'My Activity' page: https://myactivity.google.com/u/1/product/bard?utm_source=ba... Here's a screenshot: https://imgur.com/a/sgtVt2O

Mine says "Bard" where yours says LaMDA. https://i.imgur.com/p8wIPHj.png

Re: PaLM 2 Technical Report [pdf]

#262

Earlier quoted context omitted.

'Open Systems' predate 'Open Source' by a decade. https://en.wikipedia.org/wiki/Open_system_(computing) https://en.wikipedia.org/wiki/The_Open_Group

That's true but also not relevant to current wide spread use of the term. Concepts and common understanding evolve with language and I'm not even sure what point you're trying to make by pointing this out. Your first link even includes the language: > "open source" marketed as trumping "open system". Common use and understanding of the use "open" evolved decades ago. Your comment also tries to side step the issue at…

We're discussing the use of the word 'Open'. Which was first applied to systems. Then "source", which actually did in fact argue that openness of source was more important that system open-ness. As to system open-ness, that is well understood as open access to blackbox via open (non-proprietary) APIs. Which is precisely what "OpenAI" is providing.

> Your comment also tries to side step the issue ..

We disagree. Narrowly directed and addressing the "issue", in fact.

Re: PaLM 2 Technical Report [pdf]

#263
post #5

> "We then train several models from 400M to 15B on the same pre-training mixture for up to 1 × 1022 FLOPs." Seems that for the last year or so these models are getting smaller. I would be surprised if GPT-4 had > the number of parameters as GPT-3 (i.e. 175B). Edit: Seems those numbers are just for their scaling laws study. They don't explicitly say the size of PaLM 2-L, but they do say "The largest model in the PaLM…

GPT-4 is way slower than GPT-3. Unless they are artificially spiking the latency to hide parameter count, it’s likely around 1trn params

Assuming that PaLM 2 was trained Chinchilla optimal, the Chinchilla scaling law allows us to calculate how much compute (and training tokens) they would have needed for 1 trillion parameters. I haven't done the calculations, but I'm pretty sure we would get an absurdly large number.

Re: PaLM 2 Technical Report [pdf]

#264

Earlier quoted context omitted.

You mean when they filled out a form to incorporate their non-profit. Which they later turned into a for-profit company after reaping all the goodwill. The “Open” used to mean something.

"Open" means: "Open for business." Not sure how anyone could confuse that.

They are not confused. "OpenAI is a non-profit artificial intelligence research company. Our goal is to advance digital intelligence in the way that is most likely to benefit humanity as a whole, unconstrained by a need to generate financial return. Since our research is free from financial obligations, we can better focus on a positive human impact."[1]

[1] https://openai.com/blog/introducing-openai

Re: PaLM 2 Technical Report [pdf]

#265

I found an exciting feature—a way to submit a large amount of text—larger than you can paste in the Bard dialog window. (It's possible this isn't a new feature. Bard explained it to me this evening.) You can submit links to files in Google Drive. The links have to be publicly accessible. I just pasted the link to my file in Bard chat. Bard can access the contents of the 322K file I pasted the link to. It definitely k…

A similar thing actually works with Bing Chat. Just open a local text file with Edge:

https://twitter.com/MParakhin/status/1653164069032169473

Parakhin is the Bing manager.

Re: PaLM 2 Technical Report [pdf]

#266

Earlier quoted context omitted.

In their presentation, they talked about multiple sizes for the PaLM 2 model, named Gecko, Otter, Bison and Unicorn, with Gecko being small enough to run offline on mobile devices. I can't seem to find any info on what size model is being used with Bard at the moment.

Indeed, it's likely that they're running a fairly small model. But this is in and of itself a strange choice, given how ChatGPT became the gateway drug for OpenAI. Why would Google set Bard up for failure like that? Surely they can afford to run a more competent model as a promo, if OpenAI can?

This is just one task it fails at, hardly enough to generalize from.

Re: PaLM 2 Technical Report [pdf]

#267
post #183

Earlier quoted context omitted.

Is this sarcasm? I can’t tell.

It's not. The internet will be crazy once compute will be cheap enough to slightly modify all displayed content slightly to suit your personal user profile.

So you think Reddit is going to replace their actual content… with very believable generated text? And that’s going to fool people at scale? How does that help Reddit (or other org) combat bots? You can just put garbage text that seems real but has nothing to do with todays news (or politics or science).

I’m really struggling to understand how you think this is going to work and result in harm.

This assumes both the site and the reader are really dumb.

Re: PaLM 2 Technical Report [pdf]

#268
post #212

Here is their Chat Playground for PaLM 2 https://console.cloud.google.com/vertex-ai/generative/langua... (you have to be logged in to Google Cloud Console I think) Anyone know what parameters are best for code generation? I tried something simple for Node.js and it wasn't horrible but not working. Maybe I used the wron parameters. I tried using 0 for the temperature and turning everything else down like I do with the…

Isn't Chat-Bison-001 Palm 1? Edit: It seems I can't use my free credits on Vertex APIs... Not nice.

Bison is apparently the second largest PaLM 2 model:

> Even as PaLM 2 is more capable, it’s also faster and more efficient than previous models — and it comes in a variety of sizes, which makes it easy to deploy for a wide range of use cases. We’ll be making PaLM 2 available in four sizes from smallest to largest: Gecko, Otter, Bison and Unicorn. Gecko is so lightweight that it can work on mobile devices and is fast enough for great interactive applications on-device, even when offline.

https://blog.google/technology/ai/google-palm-2-ai-large-lan...

Re: PaLM 2 Technical Report [pdf]

#269

Earlier quoted context omitted.

Yeah 1 to 2 trillion is the estimates I've heard. Given the 25 messages / 3 hour limit in chatGPT, I don't think they've found a way to make it cheap to run.

1. there's no reason to think OpenAI wouldn't also be going the artificial scarcity route as have so many other companies in the past 2. Microsoft may not like them using too much azure compute and tell them to step off. Rumor has it they're trying to migrate github to it and it's seemingly not going ideal. And they're certainly nothing more than another microsoft purchase at this point.

Based on GPT3.5 supposedly using 8x A100's per query and the suspected magnitude size difference with GTP4 I really think they're struggling to run it.

At this stage I think they'd have more to benefit by making it more accessible, there's several use cases I have (or where I work) that only really make sense with GPT4, and it's way too expensive to even consider.

Also AFAIK Github Copilot is still not using GPT4 or even a bigger CODEX, and GPT4 still outperforms it especially in consistency (I'm in their copilot chat beta).

Re: PaLM 2 Technical Report [pdf]

#270
post #200

Earlier quoted context omitted.

All non-profit means is a different tax status. Don't assume they actually don't make money.

Nonprofit status makes it much harder to extract large profits. A charity founder can pay himself a million-dollar salary, but he can't sell his shares in the nonprofit and become a billionaire.

> Nonprofit status makes it much harder to extract large profits. A charity founder can pay himself a million-dollar salary, but he can't sell his shares in the nonprofit and become a billionaire.

What difference does it make for a non-public company? They can pay themselves more salary either way. The shares aren't really valuable until then.

As to a charity - if you really believe so. It doesn't even enter the books. Have you not seen an in-person donation site? Someone gives $100, the staff keeps the $100, takes out $50, records $50 and puts that in the donation box. After a few more layers the actual donation could be just $1. I've seen these at your regular big name charities - all the time.

And let's not get started on the sponsor a child that doesn't exist options...

Post reply on HN