Viewing profile — batshit_beaver
batshit_beaver
HN member- Joined
- Tue, Nov 02, 2021, 3:27 PM UTC
- HN karma
- 288
- Public activity
- 119 items
- HN profile
- View on Hacker News ↗
About batshit_beaver
No profile information was provided.
Recent public activity
-
comment
Comment #49247946
It’s not just math. Anecdotally, LLMs struggle the same way with software engineering where the code they write is correct (compiles and passes tests), but reasoning is wrong often…
-
comment
Comment #49245262
They score 99% but then can’t explain _why_ their solution works with any accuracy. Not as useful for learning as one might think.
-
comment
Comment #49245044
It was this way well before vibe coding. Over a decade of zero interest rates combined with talent wars and other anticompetitive behaviors by large tech companies did the industry…
-
comment
Comment #49160037
> Now back to AI for software creation. "Review is the bottleneck because EVERYTHING must be judged by a human." No. It doesn't. We just need to build AI review systems, that will …
-
comment
Comment #49098760
The challenge with comparing these things to humans, is that humans learn. A newbie might not respect your organization’s set of policies on day one, but what about 3 months in? Or…
-
comment
Comment #48778218
I think the issue is that obtaining empirical proof of AI or manual coding being more efficient is very difficult, since true costs and outcomes aren’t known for months and often y…
-
comment
Comment #48437865
Hopefully everyone? Else your job could have been outsourced or replaced by a junior with access to Google and StackOverflow way before LLMs (it just wasn’t due to zero interest ra…
-
comment
Comment #48407289
Humans have goal seeking behavior. LLMs don’t. You could maybe call the combination of LLMs and the RL-based harnesses somewhat “intelligent” in aggregate, but the problem is that …
-
comment
Comment #48365636
> Significantly increased my productivity as a software engineer. You’re going to have to define productivity as it applies to software engineering. With LLMs we’ve primarily seen …
-
comment
Comment #48348414
Yes that’s the right source. There would be no recovery in SWE market after the higher interest rates killed it if LLMs had any major impact on SWE employability.
-
comment
Comment #48343331
> The real risk isn't that some 19 year-old vibe coder is going to replace you, it's that there's simply less need for more experienced engineers. The market is shrinking. That las…
-
comment
Comment #48296166
Always has been tbh
-
comment
Comment #48295752
Back in the day, you couldn’t ask stack overflow about your specific business or project. You were forced to build at least some level of understanding of what you were doing on th…
-
comment
Comment #48275678
> We can debate as to how successful we’ve been toward the two goals above, but I think it’s misguided to say that the majority of people think LLMs should produce lower quality co…
-
comment
Comment #48275601
Can someone explain these complaints about boilerplate to me? What are y’all doing where boilerplate is the majority of your code? Am I the only one mostly writing concise business…
-
comment
Comment #48173196
If only it was that simple. The reason these inefficient companies continue to exist is due to regulatory capture and monopolistic behavior. Competing with them doesn't just requir…
-
comment
Comment #48170240
The problem is that organizations are inefficient in such a way that extra output from white collar workers doesn't translate to improved org-wide performance in a positively corre…
-
comment
Comment #48145454
Now they're looking at your token consumption, which is even more gameable (and stupid).
-
comment
Comment #48109605
Oh no, we should create a fear mongering blog post and delay the latest IDE version until we have better security in place!
-
comment
Comment #48071595
10x the amount of code or features =/= 10x the speed of software development.
-
comment
Comment #48044568
1. It's unclear why creating more code faster is a good thing. Software engineering wisdom for decades has been that code is a cost, not a product. There are great reasons for that…
-
comment
Comment #48043894
If you read these further, researchers believe this effect does exist, but only insofar as priming the model for the answer it was likely to give anyway and only when queries are i…
-
comment
Comment #48037540
Examples: https://arxiv.org/html/2506.02878v1 https://arxiv.org/pdf/2508.01191 Anthropic themselves: https://www.anthropic.com/research/reasoning-models-dont-say... They were appro…
-
comment
Comment #48026616
Right, and then look at any number of research papers showing that CoT output has limited impact on the end result. We've trained these models to pretend to reason.
-
comment
Comment #48012723
GitHub doesn't pay top of market.