Earlier quoted context omitted.
Downvoted by the AI Nazis. They are running a tight ship before the IPOs.
I downvoted it because it doesn't add anything useful to the conversation, and I don't own any AI stock.
GPT-5.5
121–130 of 1001 posts
Re: GPT-5.5
#122I recommend anybody in offensive/defensive cybersecurity to experiment with this. This is the real data point we needed - without the hype!
Never thought I'd say this but OpenAI is the 'open' option again.
Re: GPT-5.5
#123If GPT-5.5 Pro really was Spud, and two years of pretraining culminated in one release, WOW, you cannot feel it at all from this announcement. If OpenAI wants to know why they like they’ve fallen behind the vibes of Anthropic, they need to look no further than their marketing department. This makes everything feel like a completely linear upgrade in every way.
Re: GPT-5.5
#124> Across all three evals, GPT‑5.5 improves on GPT‑5.4’s scores while using fewer tokens. Yeah, this was the next step. Have RLVR make the model good. Next iteration start penalising long + correct and reward short + correct. > CyberGym 81.8% Mythos was self reported at 83.1% ... So not far. Also it seems they're going the same route with verification. We're entering the era where SotA will only be available after KYC…
Re: GPT-5.5
#125A playable 3D dungeon arena prototype built with Codex and GPT models. Codex handled the game architecture, TypeScript/Three.js implementation, combat systems, enemy encounters, HUD feedback, and GPT‑generated environment textures. Character models, character textures, and animations were created with third-party asset-generation tools The game that this prompt generated looks pretty decent visually. A big part of th…
The meshes look interesting, but the gameplay is very basic. The tank one seems more sophisticated with the flying ships and whatnot. What's strange is that this Pietro Schirano dude seems to write incredibly cargo cult prompts. Game created by Pietro Schirano, CEO of MagicPath Prompt: Create a 3D game using three.js. It should be a UFO shooter where I control a tank and shoot down UFOs flying overhead. - Think step…
What is this, 2023?
I feel like this was generated by a model tapping in to 2023 notions of prompt engineering.
Re: GPT-5.5
#126A playable 3D dungeon arena prototype built with Codex and GPT models. Codex handled the game architecture, TypeScript/Three.js implementation, combat systems, enemy encounters, HUD feedback, and GPT‑generated environment textures. Character models, character textures, and animations were created with third-party asset-generation tools The game that this prompt generated looks pretty decent visually. A big part of th…
The meshes look interesting, but the gameplay is very basic. The tank one seems more sophisticated with the flying ships and whatnot. What's strange is that this Pietro Schirano dude seems to write incredibly cargo cult prompts. Game created by Pietro Schirano, CEO of MagicPath Prompt: Create a 3D game using three.js. It should be a UFO shooter where I control a tank and shoot down UFOs flying overhead. - Think step…
*BELIEVE!* https://www.youtube.com/watch?v=D2CRtES2K3E
Re: GPT-5.5
#127Everyone talked about the marketing stunt that was Anthropic's gated Mythos model with an 83% result on CyberGym. OpenAI just dropped GPT 5.5, which scores 82% and is open for anybody to use. I recommend anybody in offensive/defensive cybersecurity to experiment with this. This is the real data point we needed - without the hype! Never thought I'd say this but OpenAI is the 'open' option again.
Re: GPT-5.5
#128It's possible that "smarter" AI won't lead to more productivity in the economy. Why? Because software and "information technology" generally didn't increase productivity over the past 30 years. This has been long known as Solow's productivity paradox. There's lots of theories as to why this is observed, one of them being "mismeasurement" of productivity data. But my favorite theory is that information technology is m…
Do you think it'd be viable to run most businesses on pen and paper? I'll give you email and being able to consume informational websites - rest is pen and paper.
Re: GPT-5.5
#129> Across all three evals, GPT‑5.5 improves on GPT‑5.4’s scores while using fewer tokens. Yeah, this was the next step. Have RLVR make the model good. Next iteration start penalising long + correct and reward short + correct. > CyberGym 81.8% Mythos was self reported at 83.1% ... So not far. Also it seems they're going the same route with verification. We're entering the era where SotA will only be available after KYC…
Re: GPT-5.5
#130Earlier quoted context omitted.
You are paying per token, but what you care about is token efficiency. If token efficiency has improved by as much as they claim it did (i.e. you need less tokens to complete a task successfully) all seems well.
Not for coding because it actually needs to read and write large files
Same principle applies when designing plans for complex tasks, etc. Token amount to grasp a concept is what matters.