Live data from Hacker News

Viewing profile — agucova

agucova

HN member
Joined
Sun, Apr 21, 2019, 12:59 AM UTC
HN karma
345
Public activity
171 items

About agucova

AI Safety, open source, infosec.

agucova.dev - email: hn@agucova.dev

[ my public key: https://keybase.io/agucova; my proof: https://keybase.io/agucova/sigs/UXfmBddPlca_aGeUF959N4GLfFuoFgfgXvzk6-n8Y1g ]

Recent public activity

  1. comment
    Comment #45950920

    FWIW I work on AI and I also trust Pangram quite a lot (though exclusively on long-form text spanning at least 4 or more paragraphs). I'm pretty sure the book is heavily AI written…

  2. comment
    Comment #45950887

    How long were the extracts you gave to Pangram? Pangram only has the stated very high accuracy for long-form text covering at least a handful of paragraphs. When I ran this book, I…

  3. comment
    Comment #45950871

    I ran the introduction chapter through Pangram [1], which is one of the most reliable AI-generated text classifiers out there [2] (with a benchmarked accuracy of 99.85% over long-f…

  4. comment
    Comment #42098503

    This benchmark’s questions and answers will be kept fully private, and the benchmark will only be run by Epoch. Short of the companies fishing out the questions from API logs (whic…

  5. comment
    Comment #42097240

    For some context on why this is important: this benchmark was designed to be extremely challenging for LLMs, with problems requiring several hours or days of work by expert mathema…

  6. comment
    Comment #41833860

    I’m guessing he’s probably talking about LessWrong, which nowadays also hosts a ton of serious safety research (and is often dismissed offhandedly because of its reputation as an i…

  7. comment
    Comment #41461090

    I mean, this is how the Reflection model works. It's just hiding that from you in an interface.

  8. comment
    Comment #41446629

    You can use Daggity.jl: https://docs.juliahub.com/Dagitty/kxRMH/0.0.1/

  9. comment
    Comment #41446046

    I agree. I’m really more concerned about bioweapons, for which it’s generally understood (in security studies) that access to technical expertise is the limiting factor for terrori…

  10. comment
    Comment #41441564

    I imagine you meant societal harms? I think this was mostly my fault. I edited the areas of work a bit to better reflect what the UK AISI is actually working on right now.

  11. comment
    Comment #41441528

    I recommend checking out the UK AISI's work on this: - https://www.gov.uk/government/publications/ai-safety-institu... - https://www.aisi.gov.uk/work/advanced-ai-evaluations-may-up…

  12. comment
    Comment #41441474

    > A government agency determining limits on, say, heavy metals in drinking water is materially different than the government making declarations of what ideas are safe and which ar…

  13. comment
    Comment #41441467

    > Because lobbying exists in this country, and because legislators receive financial support from corporations like OpenAI, any so-called concession by a major US-based company to …

  14. comment
    Comment #41441456

    > My issue with AI safety is that it's an overloaded term. It could mean anything from an llm giving you instructions on how to make an atomic bomb to writing spicy jokes if you pr…

  15. comment
    Comment #41441351

    > What exactly does the evaluation entail? I believe the US AISI has published less on their specific approach, but they’re largely expected to follow the general approach implemen…

  16. comment
    Comment #41025916

    > If training data contains multiple conflicting perspectives on a topic, the LLM has a limited ability to recognize that a disagreement is present and what types of entities are m…

  17. story
  18. comment
    Comment #41012303

    I messed up the second reference, it should be https://arxiv.org/abs/2212.03827

  19. comment
    Comment #41012296

    This isn't really true. LLMs are discriminating actual truth (though perhaps not perfectly). Other similar studies suggest that they can differentiate, say, between commonly held m…

  20. comment
    Comment #40998518

    > Now, we can see from this description that nothing about the modeling ensures that the outputs accurately depict anything in the world. There is not much reason to think that the…

  21. comment
    Comment #40402985

    Is your hypothesis that what, Jan Leike resigned as part of an elaborate conspiracy to boost OpenAI's prospect by... criticizing it? I find these theories to be extremely convolute…

  22. comment
    Comment #40402957

    For context, the point of the Superalignment team was to work on a problem known as scalable oversight: the problem of aligning models in a way that holds up as models become more …

  23. comment
    Comment #40402912

    I find this kind of dismissive attitude annoying. There are good arguments in the literature for why you might want to care about these risks [1, 2], and I think there's lots of ro…

  24. comment
    Comment #40402893

    Agreeing with circuit10's comments, I don't think many proponents of AI Safety are doing so through Pascal wagers. People differ a lot in their assessment of how likely certain ris…

  25. comment
    Comment #40402820

    This is true, but even engineers see the advantages of Julia. My engineering school has went from almost pure Matlab usage to many key engineering courses switching to Julia due to…