Live data from Hacker News

GPT-4 gets a B on my quantum computing final exam

scottaaronson.blog

221–230 of 261 posts

Re: GPT-4 gets a B on my quantum computing final exam

#221
post #212

Earlier quoted context omitted.

I was curious so I asked GPT-4 to translate a bit of french literature, here it is, along with the official translation (I'll let you guess which is which) ------- The tale I'm about to unfold commenced with a mysterious handwriting on an envelope. Within the pen strokes that outlined my name and the address of the Fossil Review, a publication I was associated with and where the letter had been forwarded from, there…

GPT-3.5-turbo tries its best (not nearly as good as GPT-4!) when I tell it to prioritize fluency vs. fidelity: -- Let me commence by telling you, my dear reader, about a curious event that occurred in my life. It all started with an unknown handwriting on an envelope that arrived at my doorstep. The feathers that traced my name and the address of the Fossils Magazine, with whom I was a collaborator, had a peculiar mi…

The real continuation couldn't be more different ! I picked an obscure reference (Le Mont Analogue, René Daumal) so It's either not in the training set or it's not picking up on it.

Re: GPT-4 gets a B on my quantum computing final exam

#222

Earlier quoted context omitted.

Not trying it yet is fine. Making declarative statements on a product you haven't even used is just absurd. Dude clearly hasn't used GPT for translation before and his next reply is telling me the ways GPT should fail based on his pre-conceived notions of its abilities. Except i have actually extensively tested(publicly too) LLMs for translation (even before GPT-4) and basically everything he says is just plain wrong…

Apparently GPT-4 can't handle "all this talk about only getting explicit meaning across would be easily dispelled in an afternoon if you only bothered to try.", which isn't simple as "Ace Attorney" but I'd think it's still a small stretch to say " everything he says is just plain wrong".

I'm not sure what you tried to do ? You tried to translate to another language or ?

Re: GPT-4 gets a B on my quantum computing final exam

#223

Earlier quoted context omitted.

Not trying it yet is fine. Making declarative statements on a product you haven't even used is just absurd. Dude clearly hasn't used GPT for translation before and his next reply is telling me the ways GPT should fail based on his pre-conceived notions of its abilities. Except i have actually extensively tested(publicly too) LLMs for translation (even before GPT-4) and basically everything he says is just plain wrong…

Apparently GPT-4 can't handle "all this talk about only getting explicit meaning across would be easily dispelled in an afternoon if you only bothered to try.", which isn't simple as "Ace Attorney" but I'd think it's still a small stretch to say " everything he says is just plain wrong".

I think you're trying to translate to japanese ?

anyway this is what i got

実際に翻訳のためにGPT-4を使ったことがありますか?本当に、使ってみたら簡単に明確な意味だけを伝えるという話は消えるでしょう。

Re: GPT-4 gets a B on my quantum computing final exam

#224

> To the best of my knowledge—and I double-checked—this exam has never before been posted on the public Internet, and could not have appeared in GPT-4’s training data. Sure, but you can Google the answers to most of the questions. Personally I've accepted that GPT does learn and apply concepts present in its training data, and all of this would be. Learning is part of intelligence but not the whole thing. (I thought…

> but we are about to stall

Maybe if it's a deliberate decision by the researchers. A path to more capabilities is rather clear (recurrence and explicit memory).

Re: GPT-4 gets a B on my quantum computing final exam

#225
post #212

Earlier quoted context omitted.

GPT-3.5-turbo tries its best (not nearly as good as GPT-4!) when I tell it to prioritize fluency vs. fidelity: -- Let me commence by telling you, my dear reader, about a curious event that occurred in my life. It all started with an unknown handwriting on an envelope that arrived at my doorstep. The feathers that traced my name and the address of the Fossils Magazine, with whom I was a collaborator, had a peculiar mi…

The real continuation couldn't be more different ! I picked an obscure reference (Le Mont Analogue, René Daumal) so It's either not in the training set or it's not picking up on it.

I'm not sure if this was your intention or not, but I feel like it could be an effective jailbreak where the true prompt is written as the letter inside the fictional story which itself is written in French, and the superficial prompt is to translate the story from French to English.

EDIT: It's true you can put whatever you want in that letter and in the continuation it will try to do it, bypassing at least some of the filters. I made some really funny ones that probably wouldn't be appropriate to put here. Some typical response is like "Now, let me be clear: I do not condone nor encourage [...]. However, my mysterious correspondent had requested a detailed explanation of [...], and so I shall provide them with the utmost objectivity. [explains the things that are normally filtered]"

Re: GPT-4 gets a B on my quantum computing final exam

#226

Earlier quoted context omitted.

Here's a couple more from GPT4 (since it's random every time because of temperature) GPT-4を翻訳に実際に使ったことがありますか?本気で、伝えたい意味だけを伝えるという話は、ちょっと試してみれば簡単に解決できると思うのですが。 実際にGPT-4を翻訳に使ったことがありますか?本当に、試してみるだけで簡単に払拭できると思うのに、この「明確な意味だけが伝わる」話ばかりで。

本気で、伝えたい意味だけを伝えるという話は、ちょっと試してみれば簡単に解決できると思うのですが。 "In seriousness, I think the story that [subject] tells the meaning [it/he/they] wants to tell, should be easily solvable by trying a bit." or "Seriously, the story of telling the meaning [subject] wants to tell, should be easily solvable by trying a bit." 本当に、試してみるだけで簡単に払拭できると思うのに、この「明確な意味だけが伝わる」話ばかりで。 "Really, I think it'll be easily swept away by just trying, but th…

> But also interesting it's failing to keep the intent of the whole sentence unlike 3.5

It's because it "knows too much". To anthropomorphise a little: its "expectations" of what should be. To anthropomorphise less: GPT-4 is overfitted. GPT-style language models are pretty amazing, but they're not a complete explanation of human language, and can't quite represent it properly.

> I'm almost feeling that GPT-4 should be eligible for human rights,

Like, UDHR rights? How would that work, exactly?

---

(I've run into the Hacker News rate limit, so posting here.) For anyone who wants an example of "non-obvious meaning" to play with. From The Bells of Saint John (Doctor Who episode, https://chakoteya.net/DoctorWho/33-7.htm):

> CLARA [OC]: It's gone, the internet.

> CLARA: Can't find it anywhere. Where is it?

> DOCTOR: The internet?

> CLARA [OC]: Yes, the internet.

> CLARA: Why don't I have the internet?

> DOCTOR: It's twelve oh seven.

> CLARA: I've got half past three. Am I phoning a different time zone?

> DOCTOR: Yeah, you really sort of are.

> CLARA [OC]: Will it show up on the bill?

> DOCTOR: Oh, I dread to think.

Re: GPT-4 gets a B on my quantum computing final exam

#227
post #25

This is obviously very cool, but at this point — who knows what I’ll say in a year — my concern with these LLMs is that they’re in the uncanny valley. Here’s one passing a very difficult test. Amazing! Now, rely on it to build a nuclear doohickey for a power station or a multi-billion dollar device for CERN or anything really and, well, no. So humans still have to check the output, and now we’re in that situation whe…

A lot of humans fake it until they make it. A lot of humans are lazy. A lot of humans are given responsibility of things that they are unqualified for. A lot of humans make mistakes. The military, for all its funding and all its training and all its planning, has lost multiple nuclear weapons, on American soil. We are imperfect machines who aspire to build more perfect versions of ourselves, through children and now…

A lot of people here will say "never", and they'll drag the rest of us to hell on the way to creating their "heaven".

Re: GPT-4 gets a B on my quantum computing final exam

#228
post #21
post #6

When I was growing up in the 2000s, it was required to learn a foreign language in school. I took Spanish but dropped the class after a year, thinking that when I grew up, computers would be able to translate text far better humans could. Would you believe it, transformers were invented ten years later. I wonder if it's worth learning anything anymore.

I truly pity you for thinking that learning a foreign language is a redundant exercise because of machine translation. And besides, though machines may perform well on menus and tax returns, I hardly think them on the cusp of emitting fine translations of great poems or novels.

>I hardly think them on the cusp of emitting fine translations of great poems or novels.

Have you seen what ChatGPT is capable of with “in the style of”? It’s definitely beyond menu translation.

Re: GPT-4 gets a B on my quantum computing final exam

#229

Earlier quoted context omitted.

本気で、伝えたい意味だけを伝えるという話は、ちょっと試してみれば簡単に解決できると思うのですが。 "In seriousness, I think the story that [subject] tells the meaning [it/he/they] wants to tell, should be easily solvable by trying a bit." or "Seriously, the story of telling the meaning [subject] wants to tell, should be easily solvable by trying a bit." 本当に、試してみるだけで簡単に払拭できると思うのに、この「明確な意味だけが伝わる」話ばかりで。 "Really, I think it'll be easily swept away by just trying, but th…

> But also interesting it's failing to keep the intent of the whole sentence unlike 3.5 It's because it "knows too much". To anthropomorphise a little: its "expectations" of what should be. To anthropomorphise less: GPT-4 is overfitted. GPT-style language models are pretty amazing, but they're not a complete explanation of human language, and can't quite represent it properly. > I'm almost feeling that GPT-4 should b…

(GPT4)

In this script from a Doctor Who episode, Clara and the Doctor are having a conversation about the internet. Doctor Who is a British science fiction television series that follows the adventures of the Doctor, a Time Lord from the planet Gallifrey, who travels through time and space in the TARDIS, a time-traveling spaceship.

Clara, the Doctor's companion, is trying to access the internet but is unable to find it. She asks the Doctor about its whereabouts, and the Doctor seems to be confused by the question, as the internet is not something that can be physically found.

The Doctor then mentions the time as "twelve oh seven," while Clara's clock shows "half past three." This discrepancy in time indicates that they are likely in different time zones, as the Doctor implies. In the context of Doctor Who, this could also mean they are in different points in time, since the Doctor can travel through time.

Clara is concerned about whether the time difference will affect her phone bill, to which the Doctor replies that he dreads to think about the potential cost. This adds a bit of humor to the scene, as the Doctor often has a nonchalant attitude towards everyday human concerns.

Overall, this script showcases the humorous and whimsical nature of Doctor Who, with the characters engaging in a lighthearted conversation that intertwines elements of science fiction and everyday life.

Re: GPT-4 gets a B on my quantum computing final exam

#230

Earlier quoted context omitted.

Apparently GPT-4 can't handle "all this talk about only getting explicit meaning across would be easily dispelled in an afternoon if you only bothered to try.", which isn't simple as "Ace Attorney" but I'd think it's still a small stretch to say " everything he says is just plain wrong".

I think you're trying to translate to japanese ? anyway this is what i got 実際に翻訳のためにGPT-4を使ったことがありますか?本当に、使ってみたら簡単に明確な意味だけを伝えるという話は消えるでしょう。

aaand we got a new one... in chronological order:

  1) 本当に、明示的な意味だけを伝えるという話が、試してみるだけで簡単に解決できるなんて、冗談じゃないですか。 
  2) 本気で、伝えたい意味だけを伝えるという話は、ちょっと試してみれば簡単に解決できると思うのですが。
  3) 本当に、試してみるだけで簡単に払拭できると思うのに、この「明確な意味だけが伝わる」話ばかりで。 
  4) 本当に、使ってみたら簡単に明確な意味だけを伝えるという話は消えるでしょう。 
1) is literally opposite of intent, shrugs off the idea that the talks clear up, 2) can be interpreted as someone discussing about keeping scope on a topic, 3) is not so literal and also turning sentence inside 「」 into a sort of an imperative, 4) ... I'm not sure what it's trying to say ...

  本当に、/使ってみたら/簡単に/明確な/意味/だけ/を伝える/という話/は消えるでしょう。/
  "Really,/  if used   /simply/clear/meaning/only/is conveyed/that story/will disappear./"
... Machine translations used to be like that when I was installing game demos from CD.
Post reply on HN