Viewing profile — data_maan
data_maan
HN member- Joined
- Sun, May 08, 2022, 4:15 PM UTC
- HN karma
- 651
- Public activity
- 433 items
- HN profile
- View on Hacker News ↗
About data_maan
Recent public activity
-
comment
Comment #48580801
Makes me wonder what the value of a human on a PhD course is
-
comment
Comment #48552189
Michal Valko's paper should be mentioned. It's discussed here https://youtu.be/UUq4ixTmye8?is=K75EsFIYgyrPKdCI
-
comment
Comment #48475692
Another famous dude dumping his thoughts on HN who is gulping it up like an addict. Add this to the long list of names like Terence Tao, and others who seem to be intellectually in…
-
comment
Comment #48472632
Aside from the sad life events, little information is shares about his "system", the thing HN is interested in. Is this more than a harness built on top of a SOTA commercial LLM?
-
comment
Comment #48433882
Was this in the GPT2 paper?
-
comment
Comment #47704793
If LLMs lie as much as the OP claims in the article, why can they then solve Olympiad math problems they never saw during training, consistently? There's the aimoprize.com on Kaggl…
-
comment
Comment #47610036
More "American minds": https://en.wikipedia.org/wiki/Hartmut_Esslinger Chief designer at Apple war German.
-
comment
Comment #47599585
> built with American capital and mostly American minds. I would say "built with American agency and commercial spirit", not minds. Most of the things that we have were first built…
-
comment
Comment #47599501
To be fair, Iran is not pretentious either, killing a few thousand people because they dared to protest. There are no good guys in this conflict.
-
comment
Comment #47599468
https://www.worldatlas.com/us-history/wars-the-united-states...
-
comment
Comment #47498660
A model to whose internals we don't have access solved a problem we didn't knew was in their datasets. Great, I'm impressed
-
comment
Comment #47386452
Strategic? Yes. Moral? Hm. From a moral POV this would be about who has the right to terrorize the Iranian population: the Iranian government or the US/Israel government.
-
comment
Comment #47286935
Opinions differ: hobby coders love it, but domain expert secretly despise it because it narrows the gap between the skills they spent years honing and the average Claude, I mean Jo…
-
comment
Comment #47286551
> The people getting pushed out are the intermediates and seniors who aren't high performers. Also the people that can't market themselves. There are very average programmers that …
-
comment
Comment #47150192
I love these posts that are so on the edge that I can't tell if it's sarcastic or for real :)
-
comment
Comment #46965410
> What do you mean ? These are top-notch mathematicians YeS. I didn't dispute that. I disputed that they are NOT top notch ML specialist and have made one of the worst benchmarks o…
-
comment
Comment #46965317
If it's the latter case (which it has to be), it seems that attention credit (via, e.g., articles in NY Times) is very unfairly distributed. None of the people that advanced the st…
-
comment
Comment #46965280
> We will learn if the magical capabilities attributed to these tools are really true or not. They're not. We already know that. FrontierMath. Yu Tsumura's 553th problem, RealMath …
-
comment
Comment #46965240
> These problems are representative of the types of subproblems research mathematicians have to solve to get a “research result”. They are finding that LLMs aren’t that useful for …
-
comment
Comment #46965221
But everything has been explored in other datasets already. If only a bunch of mathematicians learn something, why are so many people talking about this, why is the NY Times postin…
-
comment
Comment #46932460
It's not angst. It's intense frustration that they 1) are not doing the science correctly, and 2) that others (e.g. FrontierMath) already did everything they claim to be doing, so …
-
comment
Comment #46932443
If you want to do this rigorously, you should run it as a competition like the guys at the AI-MO Prize are doing on Kaggle. That way you get all the necessary data. I still think t…
-
comment
Comment #46932433
> There are some experiments which cannot be carried out more than once Yes, in which case a very detailed methodology is required: which hardware, runtimes, token counts etc. This…
-
comment
Comment #46932427
It wasn't like this in any way. CASP relies on a robust benchmark (not just 10 random proteins), and has clear participation criteria, objective metrics how the eval plays out, etc…
-
comment
Comment #46932412
Yes, but people at those labs may be running those problems because a Fields Medalist is in the paper, and it got hype. Not because of the problems, and not because this is new met…