Live data from Hacker News

GPT-4

openai.com

261–270 of 1001 posts

Re: GPT-4

#261

I think it's interesting that they've benchmarked it against an array of standardized tests. Seems like LLMs would be particularly well suited to this kind of test by virtue of it being simple prompt:response, but I have to say...those results are terrifying. Especially when considering the rate of improvement. bottom 10% to top 10% of LSAT in What are the implications for society when general thinking, reading, and…

There is a fundamental disconnect between the answer on paper and the understanding which produces that answer.

Edit: feel free to respond and prove me wrong

Re: GPT-4

#262
I wonder how it scored on the individual sections in the LSAT? Which section is it the best at answering?

Re: GPT-4

#264

I think it's interesting that they've benchmarked it against an array of standardized tests. Seems like LLMs would be particularly well suited to this kind of test by virtue of it being simple prompt:response, but I have to say...those results are terrifying. Especially when considering the rate of improvement. bottom 10% to top 10% of LSAT in What are the implications for society when general thinking, reading, and…

>What happens when ALL of our decisions can be assigned an accuracy score?

What happens is the emergence of the decision economy - an evolution of the attention economy - where decision-making becomes one of the most valuable resources.

Decision-making as a service is already here, mostly behind the scenes. But we are on the cusp of consumer-facing DaaS. Finance, healthcare, personal decisions such as diet and time expenditure are all up for grabs.

Re: GPT-4

#265
The comments on this thread are proof of the AI effect: People will continually push the goal posts back as progress occurs.

“Meh, it’s just a fancy word predictor. It’s not actually useful.”

“Boring, it’s just memorizing answers. And it scored in the lowest percentile anyways”.

“Sure, it’s in the top percentile now but honestly are those tests that hard? Besides, it can’t do anything with images.”

“Ok, it takes image input now but honestly, it’s not useful in any way.”

Re: GPT-4

#266
I wonder what the largest scale they can reach is. Because, if they can prove there’s not risk in taking on AI, and they can scale to serve international demand, it feels like GPT4 can do your job (probably) for <10k year. That means white collar work for under minimum wage. And that means business owners just become rent owners while you get fucked with nothing.

Re: GPT-4

#267

Test taking will change. In the future I could see the student engaging in a conversation with an AI and the AI producing an evaluation. This conversation may be focused on a single subject, or more likely range over many fields and ideas. And may stretch out over months. Eventually teaching and scoring could also be integrated as the AI becomes a life-long tutor. Even in a future where human testing/learning is no l…

I think a shift towards Oxford’s tutorial method [0] would be great overall and compliments your point.

“Oxford's core teaching is based around conversations, normally between two or three students and their tutor, who is an expert on that topic. We call these tutorials, and it's your chance to talk in-depth about your subject and to receive individual feedback on your work.”

[0] https://www.ox.ac.uk/admissions/undergraduate/student-life/e...

Re: GPT-4

#269

From the paper: > Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar. I'm curious whether they have continued to scale up model size/compute significantly or if they have managed to make significant innovati…

These are all good reasons, but it’s really a new level of openness from them.

Re: GPT-4

#270

> Yes, you can send me an image as long as it's in a supported format such as JPEG, PNG, or GIF. Please note that as an AI language model, I am not able to visually process images like a human would. However, I can still provide guidance or advice on the content of the image or answer any questions you might have related to it. Fair, but if it can analyze linked image, I would expect it to be able to tell me what tex…

Are you actually using Chat-GPT4 though? That would explain why it's not handling images.
Post reply on HN