Earlier quoted context omitted.
I routinely leave codex running for a few hours overnight to debug stuff If you have a deterministic unit test that can reproduce the bug through your app front door, but you have no idea how the bug is actually happening, having a coding agent just grind through the slog of sticking debug prints everywhere, testing hypotheses, etc — it's an ideal usecase
How can you afford that?
GPT‑5.3‑Codex‑Spark
51–60 of 415 posts
Re: GPT‑5.3‑Codex‑Spark
#52> Our latest frontier models have shown particular strengths in their ability to do long-running tasks, working autonomously for hours, days or weeks without intervention. I have yet to see this (produce anything actually useful).
How hard have you tried? I've been finding that the Opus 4.5/4.6 and GPT-5.2/5.3 models really have represented a step-change in how good they are at running long tasks. I can one-shot prompt all sorts of useful coding challenges now that previously I would have expected to need multiple follow-ups to fix mistakes the agents made. I got all of this from a single prompt, for example: https://github.com/simonw/research…
Re: GPT‑5.3‑Codex‑Spark
#53I love this! I use coding agents to generate web-based slide decks where “master slides” are just components, and we already have rules + assets to enforce corporate identity. With content + prompts, it’s straightforward to generate a clean, predefined presentation. What I’d really want on top is an “improv mode”: during the talk, I can branch off based on audience questions or small wording changes, and the system p…
Re: GPT‑5.3‑Codex‑Spark
#54Re: GPT‑5.3‑Codex‑Spark
#55Off topic but how is it always this HN user sharing model releases within a couple of minutes of their announcement?
Re: GPT‑5.3‑Codex‑Spark
#56Earlier quoted context omitted.
I built something similar using an MCP that allows claude to "outsource" development to GLM 4.7 on Cerebras (or a different model, but GLM is what I use). The tool allows Claude to set the system prompt, instructions, specify the output file to write to and crucially allows it to list which additional files (or subsections of files) should be included as context for the prompt. Ive had great success with it, and it r…
Why use MCP instead of an agent skill for something like this when MCP is typically context inefficient?
Re: GPT‑5.3‑Codex‑Spark
#57I love this! I use coding agents to generate web-based slide decks where “master slides” are just components, and we already have rules + assets to enforce corporate identity. With content + prompts, it’s straightforward to generate a clean, predefined presentation. What I’d really want on top is an “improv mode”: during the talk, I can branch off based on audience questions or small wording changes, and the system p…
Re: GPT‑5.3‑Codex‑Spark
#58Re: GPT‑5.3‑Codex‑Spark
#59Re: GPT‑5.3‑Codex‑Spark
#60Earlier quoted context omitted.
They were afraid for the untested tech but it looks like a leap in speed now
This is nonsense what do you mean? Mistral uses Cerebras for their LLMs as well. [0] It's certainly not "untested". [0] https://www.cerebras.ai/blog/mistral-le-chat