Live data from Hacker News

GPT-4

openai.com

651–660 of 1001 posts

Re: GPT-4

#651
post #380

Leetcode (hard) from 0/45 (GPT-3.5) to 3/45 (GPT-4). The lack of progress here, says a lot more about is NOT happening as an AI paradigm change. Still a glorified pattern matching and pattern creation engine, even if a very impressive one.

Hmm, can the average developer get even 1 out of 45 right, without practice? (zero shot)

Re: GPT-4

#652

This is all cute and entertaining, but my digital assistant still remains as dumb as ever and can’t process the simplest of ordinary tasks. I still can’t ask my phone to “add a stop at cvs if it doesn’t add more than 5 minutes to my trip” while driving and using maps/navigation. Is that too much to ask from a superhuman-performing AI that’s mastering all tasks and will disrupt everything? Or maybe the hype is more th…

What are you on about? This is exactly what LLMs like GPT-3 or GPT-4 can and will solve. It just takes some time. But the capability to understand, reason about and execute via API calls such simple instructions has absolutely been demonstrated. Getting to a shipped product takes longer of course.

Re: GPT-4

#653

After watching the demos I'm convinced that the new context length will have the biggest impact. The ability to dump 32k tokens into a prompt (25,000 words) seems like it will drastically expand the reasoning capability and number of use cases. A doctor can put an entire patient's medical history in the prompt, a lawyer an entire case history, etc. As a professional...why not do this? There's a non-zero chance that i…

Do you think this will be enough context to allow the model to generate novel-length, coherent stories? I expect you could summarize the preceding, already generated story within that context, and then just prompt for the next chapter, until you reach a desired length. Just speculating here. The one thing I truly cannot wait for is LLM's reaching the ability to generate (prose) books.

[deleted]

Re: GPT-4

#654

What I don't understand is how GPT-4 is able to do reasonably well on tests like the AMC12: Many of the AMC12 questions require a number of logical/deductive steps. If GPT-4 is simply trained on a large corpus of text, how is it able to do this? Does this imply that there is some emergent deductive ability that you get simply by learning "language?" Or am I missing something? Obviously, I'm assuming that GPT-4 wasn't…

From the blog post: "A minority of the problems in the exams were seen by the model during training, but we believe the results to be representative—see our technical report for details." They have a chart where they broke out results for the model with versus without "vision" i.e. having trained on the exam questions before.

Re: GPT-4

#655

Test taking will change. In the future I could see the student engaging in a conversation with an AI and the AI producing an evaluation. This conversation may be focused on a single subject, or more likely range over many fields and ideas. And may stretch out over months. Eventually teaching and scoring could also be integrated as the AI becomes a life-long tutor. Even in a future where human testing/learning is no l…

There was blog post on HN recently about the upbringings of great scientists, physicists, polymaths, etc. They almost invariably had access to near unlimited time with high quality tutors. He cited a source that claimed modern students who had access to significant tutoring resources were very likely to be at the top of their class. Personalized learning is highly effective. I think your idea is an exciting one indee…

""AI"" conversations count for very little in the way of getting genuine understanding. The last two decades have made the intelligentsia of the planet brittle and myopic. The economy's been a dumpster fire, running on fumes with everyone addicted to glowing rectangles. If we put an entire generation in front of an """AI""" as pupils, it'll lead to even worse outcomes in the future.

I doubt the 2 Sigma effect applies to ""AI"".

The panic about this new tech is from how people that leveraged their intelligence now need to look at and understand the other side of the distribution.

Re: GPT-4

#656
post #609

After watching the demos I'm convinced that the new context length will have the biggest impact. The ability to dump 32k tokens into a prompt (25,000 words) seems like it will drastically expand the reasoning capability and number of use cases. A doctor can put an entire patient's medical history in the prompt, a lawyer an entire case history, etc. As a professional...why not do this? There's a non-zero chance that i…

Reading the press release, my jaw dropped when I saw 32k. The workaround using a vector database and embeddings will soon be obsolete.

I don't see how. Can you elaborate?

Re: GPT-4

#657
Violate this reasoning:

If we didn't have a use case for GPT 3, 3.5, and chatGPT that was sufficiently commercial to become a product, it will never happen. This technology is a feature, not a product. The only companies that successfully monetize features can be considered IP licensing houses; of which, their business success is not comparable to companies that make products and platforms.

Re: GPT-4

#658

Dude said something like "you could hook this up to a calculator". Anyone know if that is implying this generation of model could interface with some kind of symbol processor? Or is he just saying, "in theory", there could be a model that did that? The math seems much improved and it would be a cool trick if it were emulating a symbol processor under the hood. But humans can do that and we opt for calculators and com…

Why can't calculators or WolframAlpha serve as a computational oracle for ChatGPT?

It would seem as simple as assigning probably 1 to certain recognizable queries. Maybe the difficulty is that the very problem of choosing to use a calculator entails a meta-cognitive rational decision, and it's not clear how to organize that in neural networks, which are what Turing himself called an unorganized model of computation.

Re: GPT-4

#659

Can someone point me to where I can find information on API pricing for GPT-4? Or, have they not released pricing info yet?

It’s available on the linked page. “Pricing is $0.03 per 1k prompt tokens and $0.06 per 1k completion tokens. Default rate limits are 40k tokens per minute”

Re: GPT-4

#660
I'm equally excited and terrified. Excited for the possibilities of a new technological revolution, but terrified for all potential abuses of technology the said revolution would bring. What is stoping our adversaries from developing malicious AI models and unleashing them on us?
Post reply on HN