Viewing profile — agucova
agucova
HN member- Joined
- Sun, Apr 21, 2019, 12:59 AM UTC
- HN karma
- 345
- Public activity
- 171 items
- HN profile
- View on Hacker News ↗
About agucova
agucova.dev - email: hn@agucova.dev
[ my public key: https://keybase.io/agucova; my proof: https://keybase.io/agucova/sigs/UXfmBddPlca_aGeUF959N4GLfFuoFgfgXvzk6-n8Y1g ]
Recent public activity
-
comment
Comment #45950920
FWIW I work on AI and I also trust Pangram quite a lot (though exclusively on long-form text spanning at least 4 or more paragraphs). I'm pretty sure the book is heavily AI written…
-
comment
Comment #45950887
How long were the extracts you gave to Pangram? Pangram only has the stated very high accuracy for long-form text covering at least a handful of paragraphs. When I ran this book, I…
-
comment
Comment #45950871
I ran the introduction chapter through Pangram [1], which is one of the most reliable AI-generated text classifiers out there [2] (with a benchmarked accuracy of 99.85% over long-f…
-
comment
Comment #42098503
This benchmark’s questions and answers will be kept fully private, and the benchmark will only be run by Epoch. Short of the companies fishing out the questions from API logs (whic…
-
comment
Comment #42097240
For some context on why this is important: this benchmark was designed to be extremely challenging for LLMs, with problems requiring several hours or days of work by expert mathema…
-
comment
Comment #41833860
I’m guessing he’s probably talking about LessWrong, which nowadays also hosts a ton of serious safety research (and is often dismissed offhandedly because of its reputation as an i…
-
comment
Comment #41461090
I mean, this is how the Reflection model works. It's just hiding that from you in an interface.
-
comment
Comment #41446629
You can use Daggity.jl: https://docs.juliahub.com/Dagitty/kxRMH/0.0.1/
-
comment
Comment #41446046
I agree. I’m really more concerned about bioweapons, for which it’s generally understood (in security studies) that access to technical expertise is the limiting factor for terrori…
-
comment
Comment #41441564
I imagine you meant societal harms? I think this was mostly my fault. I edited the areas of work a bit to better reflect what the UK AISI is actually working on right now.
-
comment
Comment #41441528
I recommend checking out the UK AISI's work on this: - https://www.gov.uk/government/publications/ai-safety-institu... - https://www.aisi.gov.uk/work/advanced-ai-evaluations-may-up…
-
comment
Comment #41441474
> A government agency determining limits on, say, heavy metals in drinking water is materially different than the government making declarations of what ideas are safe and which ar…
-
comment
Comment #41441467
> Because lobbying exists in this country, and because legislators receive financial support from corporations like OpenAI, any so-called concession by a major US-based company to …
-
comment
Comment #41441456
> My issue with AI safety is that it's an overloaded term. It could mean anything from an llm giving you instructions on how to make an atomic bomb to writing spicy jokes if you pr…
-
comment
Comment #41441351
> What exactly does the evaluation entail? I believe the US AISI has published less on their specific approach, but they’re largely expected to follow the general approach implemen…
-
comment
Comment #41025916
> If training data contains multiple conflicting perspectives on a topic, the LLM has a limited ability to recognize that a disagreement is present and what types of entities are m…
- story
-
comment
Comment #41012303
I messed up the second reference, it should be https://arxiv.org/abs/2212.03827
-
comment
Comment #41012296
This isn't really true. LLMs are discriminating actual truth (though perhaps not perfectly). Other similar studies suggest that they can differentiate, say, between commonly held m…
-
comment
Comment #40998518
> Now, we can see from this description that nothing about the modeling ensures that the outputs accurately depict anything in the world. There is not much reason to think that the…
-
comment
Comment #40402985
Is your hypothesis that what, Jan Leike resigned as part of an elaborate conspiracy to boost OpenAI's prospect by... criticizing it? I find these theories to be extremely convolute…
-
comment
Comment #40402957
For context, the point of the Superalignment team was to work on a problem known as scalable oversight: the problem of aligning models in a way that holds up as models become more …
-
comment
Comment #40402912
I find this kind of dismissive attitude annoying. There are good arguments in the literature for why you might want to care about these risks [1, 2], and I think there's lots of ro…
-
comment
Comment #40402893
Agreeing with circuit10's comments, I don't think many proponents of AI Safety are doing so through Pascal wagers. People differ a lot in their assessment of how likely certain ris…
-
comment
Comment #40402820
This is true, but even engineers see the advantages of Julia. My engineering school has went from almost pure Matlab usage to many key engineering courses switching to Julia due to…