Viewing profile — practice9
practice9
HN member- Joined
- Wed, Mar 28, 2018, 7:14 PM UTC
- HN karma
- 539
- Public activity
- 271 items
- HN profile
- View on Hacker News ↗
About practice9
No profile information was provided.
Recent public activity
-
comment
Comment #48157335
It's because HN is in AI meta-psychosis :) Our experience is very similar except we didn't really have a review process before, and now LLMs find bugs before PRs get merged in main…
-
comment
Comment #45726248
LLMs are getting quite good at reviewing the results and implementations, though
-
comment
Comment #45020509
I find it hilarious/sad that the 0.5x cheaper Ergo M575 has much better design in that regard (just plastic that doesn’t degrade)
-
comment
Comment #44809417
They should have used Claude Code for reviews
-
comment
Comment #44656486
Humans cannot reason about code at scale. Unless you add scaffolding like diagrams and maps and … Things that most teams don’t do or half-ass
-
comment
Comment #44546130
A variation of “no taxation without representation”?
-
comment
Comment #44288093
The human is a bad co-author here really. I deployed lots of high performance, clean, well documented etc code generated by Claude or o3. I reviewed it wrt requirements, added test…
-
comment
Comment #43842510
Well the system prompt is still the same for both models, right? Kinda points to people at OpenAI using o1/o3/o4 almost exclusively. That's why nobody noticed how cringe 4o has bec…
-
comment
Comment #43692246
Kinda similar in a way to China or Russia “disappearances”
-
comment
Comment #43341043
But who is the target group? Last time only some groups of enthusiasts were willing to work through bugs to even run the buggy release of Gemma Surely nobody runs this in productio…
-
comment
Comment #42977763
I tried the square example from the paper mentioned with o1-pro and it had no problem counting 4 nested squares… And the 5 square variation as well. So perhaps it is just a questio…
-
comment
Comment #42930091
With LLMs that even might be automated
-
comment
Comment #42706547
Well none of the labs have good frontend or mobile engineers or even infra engineers Anthropic is ahead in this because they keep their UIs simplistic so the failure modes are also…
-
comment
Comment #42067214
One of those guys needs to be fined for the pump & dump scheme (with SPCE: Virgin Galactic), and the other one should be investigated if he was receiving money from the Russian gov…
-
comment
Comment #41843477
[flagged]
-
comment
Comment #41739800
More like reparations
-
comment
Comment #41686224
The interesting thing is that ship was damaged almost immediately after leaving the port, had a chance to stop in Russian ports along the way but instead is doing a tour near EU co…
-
comment
Comment #40347323
It is cringe overenthusiastic, but a proper instructions/system prompt will fix that mostly
-
comment
Comment #40347302
I always double-check even the most obscure facts returned by GPT-4 and have yet to see a hallucination (as opposed to Claude Opus that sometimes made up historical facts). I doubt…
-
comment
Comment #40259427
Wasn't it pointing to https://x.ai just a few months ago? Interesting
-
comment
Comment #40098500
[flagged]
-
comment
Comment #39952002
teknium / Nous released Mistral finetunes (Hermes) that are quite great, and even published the datasets used for training. But for the worldsim I think they are really using Claud…
-
comment
Comment #39608844
Interesting. I had a problem a few months ago with DNS not resolving Meta servers on my Starlink internet connection, but I was able to use the UI and the apps nonetheless, just co…
-
comment
Comment #39605479
Unless it's a recent change, it works perfectly fine offline (wifi turned off). As for alternatives, there is Pico, but Quest 3 may be superior in games selection. Or go wired whic…
- story