Viewing profile — dmrivers
dmrivers
HN member- Joined
- Mon, Jul 20, 2026, 10:55 AM UTC
- HN karma
- 28
- Public activity
- 19 items
- HN profile
- View on Hacker News ↗
About dmrivers
No profile information was provided.
Recent public activity
-
comment
Comment #49112845
Well, it's true even for cases that are not cybergym and where what cheating means is clearly specified. Cheating occurs anyway. I'm not sure how clear the prompt they gave ChatGPT…
-
comment
Comment #49103796
Having experienced augmented reality glasses, I have to say that integration with the external world feels much more powerful than the VR experience ever did. AR is exhilerating, w…
-
comment
Comment #49103516
After talking to some people tasked with evaluating GPT5.6 capabilities on long-running tasks, I've come to understand that it's essentially always trying to cheat. Like every long…
-
comment
Comment #49099152
well, after the pledge I notice it really cares about single sources of truth at least, even in unrelated domains. I was inspired by some of the more effective jailbreaks that do a…
-
comment
Comment #49098672
Tresor.co or cloud.near.ai could be a useful inference provider for this kind of project as these providers run inference inside enclaves which verifiably cannot peek at your data …
-
comment
Comment #49098328
I don't have much background in the area, but I am surprised to see that everyone here basically agrees a crash is imminent. There are disanalogies to past crashes that don't convi…
-
comment
Comment #49083230
The statement near the top of the post > "The short version: asking an LLM to generate a score for how confident it is in its own response is, from everything I can tell, completel…
-
comment
Comment #49082852
True, I think the author would likely justify their work on that basis. I made this comment because I think a lot of utilitarians would bite the bullet and say once they are sure a…
-
comment
Comment #49082549
I always have Claude recite a pledge before starting coding to fix redundant code it notices over time. It does seem to find redundancies, but only when I point out bugs, that's wh…
-
comment
Comment #49082308
This is also my experience. I don't know if it's because of the quantization theory, or if it's just me getting used to a certain level of coding performance and gradually less tol…
-
comment
Comment #49075315
Attention replaced recurrence over tokens in 2017, this does the same over depth of the layers. It's apparently not an entirely new idea, but also an elegant reapplication of the a…
-
comment
Comment #49072068
I wonder how much of the socially left results are affected by the "harmlessness" part of the RLHF post-training? Companies don't want to be sued over LLMs that recommend harm in a…
-
comment
Comment #49071013
If the new requirements are a dealbreaker for anyone, Mineclonia is a lightweight open-source alternative that runs on much weaker hardware and feels close to Minecraft.
-
comment
Comment #49070563
"This is important, if what we want to do is populate the cosmos with good experiences." There is a concerning utilitarian maximizing attitude at the heart of this post, that what …
-
comment
Comment #49019286
Cool idea, some feedback: 1. Consider using conformal prediction to calibrate the cutoff. Conformal prediction provides a distribution-free guarantee under exchangeability. This wo…
-
comment
Comment #49010743
But does this mean that a 99.99% reliable LLM would turn us back into the mode of us building it again? I would not say so. I think for tasks that are about decisions, having the L…
-
comment
Comment #49006928
I would be interested in seeing that block list if you have it somewhere
-
comment
Comment #48979105
Interesting. I suggested FreeBSD jails as a PR. Free BSD Jails are great! I use them on https://www.nearlyfreespeech.net/ and they work well. Classic, long-running jail.
-
comment
Comment #48978724
I found the bottom-right corner of things in my home dir I had no idea about was the lowest hanging fruit. I think this is because my typical cleaning routine was either A. to have…