GPT-5
821–830 of 1001 posts
Re: GPT-5
#822It is frequently suggested that once one of the AI companies reaches an AGI threshold, they will take off ahead of the rest. It's interesting to note that at least so far, the trend has been the opposite: as time goes on and the models get better, the performance of the different company's gets clustered closer together. Right now GPT-5, Claude Opus, Grok 4, Gemini 2.5 Pro all seem quite good across the board (ie the…
Re: GPT-5
#823gpt-5 is now #1 at LMArena: https://lmarena.ai/leaderboard/text
The feel is pretty much all that matters. Needs a blind taste test, but really this is a place that mood or vibe works.
Re: GPT-5
#824It is frequently suggested that once one of the AI companies reaches an AGI threshold, they will take off ahead of the rest. It's interesting to note that at least so far, the trend has been the opposite: as time goes on and the models get better, the performance of the different company's gets clustered closer together. Right now GPT-5, Claude Opus, Grok 4, Gemini 2.5 Pro all seem quite good across the board (ie the…
It's also worth considering that past some threshold, it may be very difficult for us as users to discern which model is better. I don't think thats what's going on here, but we should be ready for it. For example, if you are an ELO 1000 chess player would you yourself be able to tell if Magnus Carlson or another grandmaster were better by playing them individually? To the extent that our AGI/SI metrics are based on…
Re: GPT-5
#825It's only when he stumbled a bit that I could tell for sure (well, mostly) that it wasn't an AI generated video - the corporate speak, body language mannerisms of Sam Altman, camera angles, all seemed pretty plausibly AI-generated!
Re: GPT-5
#826no way i am letting my kids near this. they are going to learn from books not from screens.
I hope your kids learn as well from books as their peers learn from AI.
Possible, but not very likely.
Teachers, should be terrified. Homeschool kids can literally put themselves through school now with the right motivation.
Re: GPT-5
#827Something that's really hitting me is something brought up in this piece: https://www.interconnects.ai/p/gpt-5-and-bending-the-arc-of-... When a model comes out, I usually think about it in terms of my own use. This is largely agentic tooling, and I mostly us Claude Code. All the hallucination and eval talk doesn't really catch me because I feel like I'm getting value of these tools today. However, this model is not…
Re: GPT-5
#828This livestream is atrocious
Re: GPT-5
#829I'm not really convinced, the benchmark blunder was really strange but the demos were quite underwhelming, and it appears this was reflected by a huge market correction in the betting markets as to who will have the best AI by end of the year. What excites me now is that Gemini 3.0 or some answer from Google is coming soon and that will be the one I will actually end up using. It seems like the last mover in the LLM…
The real last mover is Apple, because boy are they not moving.
Re: GPT-5
#830Claude Opus 4 has changed my workflow; never going back.