Viewing profile — versteegen
versteegen
HN member- Joined
- Fri, May 08, 2015, 6:58 PM UTC
- HN karma
- 1,809
- Public activity
- 747 items
- HN profile
- View on Hacker News ↗
About versteegen
rver017@aucklanduni.ac.nz
Recent public activity
-
comment
Comment #49205272
> For me it’s the massive amount of resources it takes to produce and run one It's amazing that LLM pretraining is both extremely data inefficient at learning concepts and cognitiv…
-
comment
Comment #49205099
You misread. "Pain and suffering" not "death". Of all the pain and suffering in the world, a vast amount of it really is our own fault. Famines and wars shouldn't happen. And if yo…
-
comment
Comment #49150642
To save anyone else the trouble: discussion there is not really worth looking at (largely a flame war), except: the author of this disproof seems to be a crank, and the disproof's …
-
comment
Comment #49107159
...but we're talking about compaction, and opencode's compaction is (or was) terrible. I've seen so many horrible problems that I keep it disabled (with an envvar flag, because eve…
-
comment
Comment #49094508
Wow, remarkable. Clearly Aum is a different league from the lone-wolf "Fort Detrick guy", treat the risks separately. But I'll take these questions as rhetorical. (See my reply to …
-
comment
Comment #49094152
I didn't argue "must be regulated". I'm arguing AI is powerful (at achieving things, and hence has dual-use dangers). Many people won't even admit that, which is the part that real…
-
comment
Comment #49092689
Thank you for taking this seriously enough to write this, and anyone else likewise. But this is attacking a strawman, amateur bioterrorists. AI is a force multiplier in the hands o…
-
comment
Comment #49092452
The old observation that people tend to define AI as whatever computers can't do yet is as true as ever. It's getting a bit absurd, moving from demanding "general intelligence" to …
-
comment
Comment #49046763
Isn't it ~$3000 per week? Extrapolating from the current limit on Pro plans.
-
comment
Comment #49043844
Disagree. Actually, in API cost equivalent, the $20/mo ChatGPT Plus plan gives you ~$100 of usage, while $20/mo Claude Pro gives you >$250 of usage (I measure at ~$300 in my last w…
-
comment
Comment #49017808
> It looks like it wrote a python script to generate test cases in our file format for testing. Just... you know, as a side quest. On the one hand, agents have done this sort of th…
-
comment
Comment #49016296
Algebra is useful because graphs are algebraic objects, and a lot of CS is about graphs, in particular search/planning. But no, I never saw rings mentioned except for generating fu…
-
comment
Comment #49016059
Having studied CS and maths to post-grad, I think you exaggerate. Although a CS course might use these tools, they didn't in my experience go into explaining or defining them. The …
-
comment
Comment #49015701
That was a tendency of 5.4 and earlier, OpenAI specifically worked to avoid it in 5.5 and I find it happens rarely know. It really felt like 5.4 had been intentionally trained to s…
-
comment
Comment #49005373
"Thinking" seems to be a political term now, people have completely different definitions of it, based on how they wish the world to be organised, and find defining it differently …
-
comment
Comment #49005302
Being correct doesn't give you licence to use insults.
-
comment
Comment #48993316
Not half, only on the order of 10% of the cost of inference is electricity.
-
comment
Comment #48979501
Pretty sure there isn't. HN has lots of hellbanned bots lately (I have showdead on [1]) but I never notice ones pushing a political slant. But I suppose nation states conducting a …
-
comment
Comment #48979239
I always run opencode with envvar `OPENCODE_DISABLE_AUTOCOMPACT=1` because I discover a new horrible bug in its autocompact every time I don't... including the config .json option …
-
comment
Comment #48968558
BTW the quotas for Go have very recently changed, now only $15 for some models instead of $60. Which is not actually a difference for DS4 Pro, because they lowered the token pricin…
-
comment
Comment #48944152
To summarise the full results table further down the page (which doesn't render on the page for me!): Kimi K3 beats each model (out of 35 benchmarks, excluding missing): vs Fable 5…
-
comment
Comment #48944140
It's a custom 3 bit quant of Llama 3.1 8B and other shortcuts. The quant is not good. Their newer arch switches to standard 4 bit quants, should be far better!
-
comment
Comment #48942902
I think MiMo 2.5 seems to be better than DS4 Flash, at the same price. DS4F can write pretty advanced code but it way overthinks the simple stuff, its CoT is full of errors (immedi…
-
comment
Comment #48872067
I've always been amazed that any time (very rarely) that I look at the tex source for a paper it's full of commented-out things not cleaned up. Nobody thinks to look?
- comment