Live data from Hacker News

Codex-maxxing

jxnl.co

21–30 of 90 posts

Re: Codex-maxxing

#21
post #9

The author of this post works at OpenAI on the Codex team.

same author who's idea of constraint decoding for structured outputs was to run an schema-begging-API call in a loop 10 times & then throw an exception on failure.

Re: Codex-maxxing

#23
post #4

in tsz ( https://tsz.dev ) I am Codex-Maxxing with this: Give each Codex an AgentName and ask them to mark their PR/issue/comments with those. Have one or two "managers" that manage PRs and overall project direction. I write the project directions and make long lasting issues. Each Codex session has an almost unachievable `/goal` but they are asked to achieve the goal by landing changes in `main` via PRs I am running…

I haven't run into /goal pausing or prematurely stopping, though I haven't used it more than a handful of times.

Re: Codex-maxxing

#25
post #7

> Every 30 minutes, check Slack and Gmail for unanswered messages that need my attention... > When I come back to Slack, replies are often already sitting in drafts. I still decide what gets sent, but the expensive part of gathering context is done. This just feels so dystopian to me. I hope that I never work with you or someone else doing this. I personally do use LLMs for work messaging but I'm extremely careful to…

If the other people in the org are using LLMs to a similar degree, any question to which an LLM can provide a good answer to will never get sent. How useful are the draft replies then?

You pretend that you did the work. It is not about achieving the result, it is about appearing productive.

Re: Codex-maxxing

#26
post #11

All the AI stuff lately is just like Unix Porn reddit but posted to places where the people don’t care about it.

Hacker News is the /r/unixporn equivalent for AI. Ground zero. They all are here.

The further away you go, the more sensible takes you find.

Re: Codex-maxxing

#27
Is this LLM psychosis? So much tending and conversing with the matmuls but what was the outcome? Are people who get this into it more successful somehow? It reminds me of people who take drugs and get "revelations" but then are not particularly over represented in the group of successful people for all of their deep insights.

Re: Codex-maxxing

#28
something is happening with `codex`, at tamarillo.ai we did a [little experiment](https://research.tamarillo.ai/coding-harness-inspection/), with 400K repos that have AI harnesses configured and very interesting behavior is observed

- growing fast as fuck

- overepresentation on starred repos (even though stars mean less these days, it is definitely something to look at)

- overepresentation in `rust`

- in terms of aliveness, codex is first

Re: Codex-maxxing

#29
post #7

> Every 30 minutes, check Slack and Gmail for unanswered messages that need my attention... > When I come back to Slack, replies are often already sitting in drafts. I still decide what gets sent, but the expensive part of gathering context is done. This just feels so dystopian to me. I hope that I never work with you or someone else doing this. I personally do use LLMs for work messaging but I'm extremely careful to…

Sounds more like CYA.

If instead of LLM you googled do you also say "Here are the CPU architectures pytorch supports, that Google search returned"

Re: Codex-maxxing

#30
post #7

> Every 30 minutes, check Slack and Gmail for unanswered messages that need my attention... > When I come back to Slack, replies are often already sitting in drafts. I still decide what gets sent, but the expensive part of gathering context is done. This just feels so dystopian to me. I hope that I never work with you or someone else doing this. I personally do use LLMs for work messaging but I'm extremely careful to…

An interesting piece of context with this guy is he writes about a serious hand injury that prevents him from typing much anymore. He says that adopting LLM workflows saved his hands (beyond just dictating everything).
Post reply on HN