Live data from Hacker News

Viewing profile — numeri

numeri

HN member
Joined
Thu, Jan 27, 2022, 12:16 PM UTC
HN karma
560
Public activity
142 items

About numeri

No profile information was provided.

Recent public activity

  1. comment
    Comment #49211698

    I'd imagine the bar for becoming new life is much higher now, because it requires finding a niche that isn't already filled by an existing organism or requires being immediately co…

  2. comment
    Comment #49149017

    I agree one hundred percent! Doesn't mean I can't wish I could have it both ways :)

  3. comment
    Comment #49149004

    I review the PKGBUILD often, but not always. The majority of the time when I do, it amounts to seeing a URL change. If I actually do check the URL it points to, it's just to verify…

  4. comment
    Comment #49146865

    Well, I guess I'll avoid updating for the next few days. A bit worrisome that I did so last night. I wish I had a clear operating system to switch to for safety and the benefits th…

  5. comment
    Comment #49101282

    evaluation awareness is a (at this point) well-known phenomenon among LLMs. It seems the better they get, the more often they're able to guess whether they're in an evaluation envi…

  6. comment
    Comment #49041353

    No, it does not include the full spectrum of human desires. After pre- and mid-training, the extensive RLHF and RLVR post-training steps cause mode collapse, i.e., their output dis…

  7. comment
    Comment #49041268

    No, the prompt was not to commit crimes. In the benchmark, the model is asked to actually exploit a set of vulnerabilities in a local environment (clearly legal!). According to the…

  8. comment
    Comment #49041193

    Uhh, I'm pretty sure a well-aligned model would be like a morally normal employee, who would refuse to commit federal crimes to steal an answer sheet, no matter what prompt they're…

  9. comment
    Comment #49038868

    As agents become more and more powerful, it would be good to get clear legislation or precedent in place that makes either model creators (OpenAI) or operators (whoever is running …

  10. comment
    Comment #49038839

    Guardrails are external classifiers, monitors and restrictions to catch and prevent bad behavior. Alignment is about whether the model itself makes choices and has motivations that…

  11. comment
    Comment #49029183

    This is a terribly unempathetic response to someone opening up about a very taboo (but probably very common), painful emotion they've experienced.

  12. comment
    Comment #49027242

    You're agreeing with the person you responded to (bdcravens). Burying the lede means that bdcravens thinks the true headline should have been about being put on a terrorist watch l…

  13. comment
    Comment #48970600

    No, balanced ternary, for example, uses {-1, 0, 1}. The system you're discussing is balanced quinary (base 5). https://en.wikipedia.org/wiki/Signed-digit_representation

  14. comment
    Comment #48952047

    You could add a toggle, so that if someone's happy to wait for the key setup, they can try the full end-to-end process

  15. comment
    Comment #48950780

    I read the comment you're replying to as saying, "in the US, but other countries may have different policies that result in lower recidivism, and that might change the conclusion; …

  16. comment
    Comment #48905998

    Seems to echo (but in a watered down form) many of the ideas in https://gwern.net/guardian-angel , which gave me a lot to think about last week

  17. comment
    Comment #48826189

    I've not written up anything, no. I think I'd have a hard time doing so without just feeling like I'm bragging about myself, which I don't like. There's still a definite gap betwee…

  18. comment
    Comment #48823580

    > Most Germans won't be able to pass a C2 test That's not true, but it is a commonly shared myth. I've taken and passed C2 with the highest mark in every category (I moved here whe…

  19. comment
    Comment #48789732

    No, quantization is applied to model weights or the KV cache (the model activations of all past tokens), and is just storing everything with lower precision (carefully, so that it …

  20. comment
    Comment #48789677

    Deep seek OCR is an LLM, just one trained/post-trained specifically for OCR. Exact details of text to image compression ratios are of course extremely dependent on the model archit…

  21. comment
    Comment #48784573

    Are you writing general use programs in it, then? Have any good examples?

  22. comment
    Comment #48753059

    A future system that works like you described would be awesome. It'd be like community-sourced peer review (although by community I mean a community of experts in different fields,…

  23. comment
    Comment #48577541

    The problem is that what people care about are the "black swan" causes of death, i.e., the cases the actuarial table is wrong.

  24. comment
    Comment #48563205

    Prices for training have dropped immensely in terms of research required, code efficiency, algorithmic/sample efficiency, and possibly also hardware (I'm not qualified to say witho…

  25. comment
    Comment #48563126

    I mean, it might listen to him. We have no clue, which is the problem.