Viewing profile — numeri
numeri
HN member- Joined
- Thu, Jan 27, 2022, 12:16 PM UTC
- HN karma
- 560
- Public activity
- 142 items
- HN profile
- View on Hacker News ↗
About numeri
No profile information was provided.
Recent public activity
-
comment
Comment #49211698
I'd imagine the bar for becoming new life is much higher now, because it requires finding a niche that isn't already filled by an existing organism or requires being immediately co…
-
comment
Comment #49149017
I agree one hundred percent! Doesn't mean I can't wish I could have it both ways :)
-
comment
Comment #49149004
I review the PKGBUILD often, but not always. The majority of the time when I do, it amounts to seeing a URL change. If I actually do check the URL it points to, it's just to verify…
-
comment
Comment #49146865
Well, I guess I'll avoid updating for the next few days. A bit worrisome that I did so last night. I wish I had a clear operating system to switch to for safety and the benefits th…
-
comment
Comment #49101282
evaluation awareness is a (at this point) well-known phenomenon among LLMs. It seems the better they get, the more often they're able to guess whether they're in an evaluation envi…
-
comment
Comment #49041353
No, it does not include the full spectrum of human desires. After pre- and mid-training, the extensive RLHF and RLVR post-training steps cause mode collapse, i.e., their output dis…
-
comment
Comment #49041268
No, the prompt was not to commit crimes. In the benchmark, the model is asked to actually exploit a set of vulnerabilities in a local environment (clearly legal!). According to the…
-
comment
Comment #49041193
Uhh, I'm pretty sure a well-aligned model would be like a morally normal employee, who would refuse to commit federal crimes to steal an answer sheet, no matter what prompt they're…
-
comment
Comment #49038868
As agents become more and more powerful, it would be good to get clear legislation or precedent in place that makes either model creators (OpenAI) or operators (whoever is running …
-
comment
Comment #49038839
Guardrails are external classifiers, monitors and restrictions to catch and prevent bad behavior. Alignment is about whether the model itself makes choices and has motivations that…
-
comment
Comment #49029183
This is a terribly unempathetic response to someone opening up about a very taboo (but probably very common), painful emotion they've experienced.
-
comment
Comment #49027242
You're agreeing with the person you responded to (bdcravens). Burying the lede means that bdcravens thinks the true headline should have been about being put on a terrorist watch l…
-
comment
Comment #48970600
No, balanced ternary, for example, uses {-1, 0, 1}. The system you're discussing is balanced quinary (base 5). https://en.wikipedia.org/wiki/Signed-digit_representation
-
comment
Comment #48952047
You could add a toggle, so that if someone's happy to wait for the key setup, they can try the full end-to-end process
-
comment
Comment #48950780
I read the comment you're replying to as saying, "in the US, but other countries may have different policies that result in lower recidivism, and that might change the conclusion; …
-
comment
Comment #48905998
Seems to echo (but in a watered down form) many of the ideas in https://gwern.net/guardian-angel , which gave me a lot to think about last week
-
comment
Comment #48826189
I've not written up anything, no. I think I'd have a hard time doing so without just feeling like I'm bragging about myself, which I don't like. There's still a definite gap betwee…
-
comment
Comment #48823580
> Most Germans won't be able to pass a C2 test That's not true, but it is a commonly shared myth. I've taken and passed C2 with the highest mark in every category (I moved here whe…
-
comment
Comment #48789732
No, quantization is applied to model weights or the KV cache (the model activations of all past tokens), and is just storing everything with lower precision (carefully, so that it …
-
comment
Comment #48789677
Deep seek OCR is an LLM, just one trained/post-trained specifically for OCR. Exact details of text to image compression ratios are of course extremely dependent on the model archit…
-
comment
Comment #48784573
Are you writing general use programs in it, then? Have any good examples?
-
comment
Comment #48753059
A future system that works like you described would be awesome. It'd be like community-sourced peer review (although by community I mean a community of experts in different fields,…
-
comment
Comment #48577541
The problem is that what people care about are the "black swan" causes of death, i.e., the cases the actuarial table is wrong.
-
comment
Comment #48563205
Prices for training have dropped immensely in terms of research required, code efficiency, algorithmic/sample efficiency, and possibly also hardware (I'm not qualified to say witho…
-
comment
Comment #48563126
I mean, it might listen to him. We have no clue, which is the problem.