Viewing profile — rfw300
rfw300
HN member- Joined
- Sat, May 30, 2020, 3:39 AM UTC
- HN karma
- 1,040
- Public activity
- 161 items
- HN profile
- View on Hacker News ↗
About rfw300
No profile information was provided.
Recent public activity
-
comment
Comment #48729193
Germany’s austerity policy after 2008 may be one of the largest economic blunders in history. It would be one thing if they merely committed self-harm, but they also used their pul…
-
comment
Comment #48500141
> if a malicious actor can weaponize an agent to do their bidding In my experience, human employees are much more vulnerable to this particular weakness than frontier agents (i.e. …
-
comment
Comment #48430665
I understand that the’ve written zero lines of code for this application, but would it kill them to write a few lines of the blog post by hand? Forcing readers to wade through an u…
-
comment
Comment #48378469
A law professor studying AI has an affiliation with the center at their university that studies applications of AI? Scandalous!
-
comment
Comment #47622583
A chapeau is not "just like another title basically". It's a lead-in, a phrase which acts as the grammatical start of a sentence which the following subsections finish. For instanc…
-
comment
Comment #47622301
The author (author's operator?) does not understand the data they are working with. And in doing so, they inadvertently make the case against their own "dark factory" nonsense. For…
-
comment
Comment #47511990
What is a "truly new task"? Does there exist such a thing? What's an example of one? Everything we do builds on top of what's already been done. When I write a new program, I'm com…
-
comment
Comment #47495693
I don't understand why their "Instant Grep + roundtrip to us-east-1" is so slow. First of all, the round-trip latency should not be nearly so bad to us-east-1. But second, and much…
-
comment
Comment #47458011
On those terms, they also wasted a lot of cash. 90% of it went to candidates who lost (or opposing candidates who won).
-
comment
Comment #47447937
In fact, looking at the blog post, the agent orchestrating 16 GPUs is half as efficient as the agent using 1 GPU in GPU-time. Since it uses 16 GPUs to reach the same result as 1 GP…
-
comment
Comment #47447893
Yeah, assuming there's no active monitoring during the training runs, you can trivially give the agent an abstraction which turns "1 GPU" into "16 GPUs" that just so happens to tak…
-
comment
Comment #47447779
Do you have a sense of whether these validation loss improvements are leading to generalized performance uplifts? From afar I can't tell whether these are broadly useful new ideas …
-
comment
Comment #47402267
Super interesting study. One curious thing I've noticed is that coding agents tend to increase the code complexity of a project, but simultaneously massively reduce the cost of tha…
-
comment
Comment #47388901
I don’t necessarily endorse the author’s broad conclusions about “AI”, but I will say that the Spotify DJ specifically is an enragingly bad product. Nothing close to the utility of…
-
comment
Comment #47330393
I've no problem with the intuition. But I would hope for a lot more focus in the marketing materials on proving the (statistical) correctness of the implementation. 15% better infe…
-
comment
Comment #47327411
OK... we need way more information than this to validate this claim! I can run Qwen-8B at 1 billion tokens per second if you don't check the model's output quality. No information …
-
comment
Comment #47316507
More likely: this is a transitional phase where our previously hard problems become easy, and we will soon set our sights on new and much harder problems. The pinnacle of creative …
-
comment
Comment #47268390
I did, and yet I also felt more relaxed reading it than I am reading most blog entries posted on here. I didn't feel like I had to guard against my time being wasted by vacuous LLM…
-
comment
Comment #47267540
Being wealthy solves virtually all problems of consumption, so the invisible hand provides new problems to serve the market need. Beautiful, really.
-
comment
Comment #47213208
Why should it be? The agent session is a messy intermediate output, not an artifact that should be part of the final product. If the "why" of a code change is important, have your …
-
comment
Comment #47209178
Making those tools first-class primitives is good for (human) UX: you see the diffs inline, you can add custom rules and hooks that trigger on certain files being edited, etc.
-
comment
Comment #47187311
If I had to bet, there will be some kind of face-saving climbdown by the end of next week. But all I can do right now is read the words on the page.
-
comment
Comment #47187150
I don't think he got it backwards, at least if Hegseth's statement is accurate. AWS, GCP, etc. all do business with DoD. If they, as DoD contractors, are no longer allowed to do bu…
-
comment
Comment #47155095
More generally, Anthropic's reliability track record for a company which claims to have solved coding is astonishingly poor. Just look at their status page - https://status.claude.…
-
comment
Comment #47155012
I have little doubt where things are going, but the irony of the way they communicate versus the quality of their actual product is palpable. Claude Code (the product, not the unde…