Live data from Hacker News

GPT-4

openai.com

441–450 of 1001 posts

Re: GPT-4

#441
post #86

> What are the implications for society when general thinking, reading, and writing becomes like Chess? I think going from LSAT to general thinking is still a very, very big leap. Passing exams is a really fascinating benchmark but by their nature these exams are limited in scope, have very clear assessment criteria and a lot of associated and easily categorized data (like example tests). General thought (particularl…

Define: "general thinking".

Re: GPT-4

#442

How big is this model? (i.e., how many parameters?) I can't find this anywhere.

welp,

This report focuses on the capabilities, limitations, and safety properties of GPT-4. GPT-4 is a Transformer-style model [33 ] pre-trained to predict the next token in a document, using both publicly available data (such as internet data) and data licensed from third-party providers. The model was then fine-tuned using Reinforcement Learning from Human Feedback (RLHF) [34 ]. Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar.

Re: GPT-4

#443
post #218

A class of problem that GPT-4 appears to still really struggle with is variants of common puzzles. For example: >Suppose I have a cabbage, a goat and a lion, and I need to get them across a river. I have a boat that can only carry myself and a single other item. I am not allowed to leave the cabbage and lion alone together, and I am not allowed to leave the lion and goat alone together. How can I safely get all three…

I think you may have misstated the puzzle. It's ok to leave the lion and the cabbage together, assuming it's not a vegetarian lion.

this here is why it's not fair to criticize GPT-4 so quickly on this question.

for the record, I made the same mistake as nonfamous at first, i almost commented "but it's correct" before going back to double check what i was missing.

i simply skimmed the problem, recognized it as a common word problem and totally missed the unusual constraints from the question. i just didn't pay attention to the whole question.

Re: GPT-4

#444

This technology has been a true blessing to me. I have always wished to have a personal PhD in a particular subject whom I could ask endless questions until I grasped the topic. Thanks to recent advancements, I feel like I have my very own personal PhDs in multiple subjects, whom I can bombard with questions all day long. Although I acknowledge that the technology may occasionally produce inaccurate information, the…

If you don't know the subject, how can you be sure what it's telling you is true? Do you vet what ChatGPT tells you with other sources? I don't really know Typescript, so I've been using it a lot to supplement my learning, but I find it really hard to accept any of its answers that aren't straight code examples I can test.

> If you don't know the subject, how can you be sure what it's telling you is true?

That applies to any article, book, or a verbal communication with any human being, not only to LLMs

Re: GPT-4

#445
GTP is a cult, like any language upstart. Except, it's not a programming language, and it's not exactly natural language either. It's some hybrid without a manual or reference.

I'll continue to pass, thanks.

Re: GPT-4

#446
post #86

> What are the implications for society when general thinking, reading, and writing becomes like Chess? I think going from LSAT to general thinking is still a very, very big leap. Passing exams is a really fascinating benchmark but by their nature these exams are limited in scope, have very clear assessment criteria and a lot of associated and easily categorized data (like example tests). General thought (particularl…

What might be interesting is to feed in the transcripts & filings from actual court cases and ask the LLM to write the judgement, then compare notes vs the actual judge.

Re: GPT-4

#448
The measure of intelligence is language - specifically language evolved by the subject organisms themselves to co-operate together.

Wake me up when GPT-X decides to start talking to other GPT-Xs - until then you just have a very sophisticated statistics package (which may be quite useful, but not AI).

Re: GPT-4

#449

The measure of intelligence is language - specifically language evolved by the subject organisms themselves to co-operate together. Wake me up when GPT-X decides to start talking to other GPT-Xs - until then you just have a very sophisticated statistics package (which may be quite useful, but not AI).

It can already talk to other agents. It also can already use “language” better than almost all humans (multiple languages, more vocab, etc)

I guess what you’re talking about is it just going and doing something by itself with no prompt? Not sure why that should be a goal, and I also don’t see why it couldn’t do that right now? “Whenever the sky is blue, reach out to ChatGPT and talk about the weather”

Re: GPT-4

#450

The measure of intelligence is language - specifically language evolved by the subject organisms themselves to co-operate together. Wake me up when GPT-X decides to start talking to other GPT-Xs - until then you just have a very sophisticated statistics package (which may be quite useful, but not AI).

It can already talk to other agents. It also can already use “language” better than almost all humans (multiple languages, more vocab, etc)

I guess what you’re talking about is it just going and doing something by itself with no prompt? Not sure why that should be a goal, and I also don’t see why it couldn’t do that right now? “Develop a language with this other ChatBot”

Post reply on HN