Viewing profile — WASDx
WASDx
HN member- Joined
- Sun, May 29, 2022, 6:05 PM UTC
- HN karma
- 166
- Public activity
- 66 items
- HN profile
- View on Hacker News ↗
About WASDx
No profile information was provided.
Recent public activity
-
comment
Comment #49193427
And they are all TUI's installed via curl | bash.
-
comment
Comment #49166980
At 830tok/s * 1 hour that's almost 3M tokens which is just $0.54 worth of tokens at Deepseeks current output price.
-
comment
Comment #49056833
Do you know why they don't just cache the system prompt for everyone? It seems so wasteful not to.
-
comment
Comment #48994637
DeepSWE and FrontierCode are more realistic if you read up on what they actually measure. But the most realistic is to try it yourself. Benchmarks can only vaguely represent typica…
-
comment
Comment #48971290
Great explanation, thanks!
-
comment
Comment #48952628
> On top of that, doing research in the open amortizes the cost. Can you elaborate on this? I appreciate the open models but don't see the economics behind just giving them away li…
-
comment
Comment #48665260
I see only these two possibilities: 1. If LLMs keep improving, burning models onto silicon becomes obsolete too fast and is not worth doing. Outcome: We keep getting better LLMs. 2…
-
comment
Comment #48572627
Are you suggesting it should summarize the image in text or generate it in HTML or something else?
-
comment
Comment #48558759
Looking at some benchmarks, the latest ~30B Gemma/Qwen score similar as Claude or GPT versions that were released just one year earlier . That's crazy progress. I can't imagine how…
-
comment
Comment #48520735
I think this is inevitable. Sooner or later, model-specific ASIC's will make economical sense. We're already seeing it happening with Taalas/Cerebras so I think it's sooner than 5 …
-
comment
Comment #48520543
> distributed LLM inference This seems extremely inefficient considering data transfer between model layers if the model is distributed. I found this project called Petals that cla…
-
comment
Comment #48467049
I like this one, although its data seem to overlap with ECI. https://artificialanalysis.ai/trends
-
comment
Comment #48325702
https://chatjimmy.ai/ from Taalas also feels like that.
-
comment
Comment #48316137
I think their "code" ranking is biased towards visual aesthetics more than raw coding as the voters are just asked which generated website they prefer.
-
comment
Comment #48133868
I've had mostly problem-free experiences with intellij (ultimate-only feature I think). One click finds declarations both in business code and buried deep in libraries.
-
comment
Comment #48040247
gemma-4-31B-it-assistant is a 0.5B model. So it's performance would likely be comparable to other models of such size.
-
comment
Comment #48039683
I think this is the future. When models start converging at "really good" (which I think is already happening) then burning them into ASIC silicon is the natural next step. Harness…
-
comment
Comment #47938879
I was impressed enough by AI finding vulnerabilities in source code, but doing it in binary executables is just amazing. This has so much potential, good and bad. And yet another l…
-
comment
Comment #47658259
Creating a custom tuple class to use as key could be faster though. Nested map lookups have less efficient memory access patterns.
-
comment
Comment #47658168
Similar site with same features: https://xn--1-zfa.com/
-
comment
Comment #47357164
I think these limitations could be addressed by allowing trivial manual adjustments to the generated code before committing. And/or allowing for trivial code changes without a spec…
-
comment
Comment #46651179
I've managed a 100+ node cluster for years without seeing any corruption. Where are you getting this from?
-
comment
Comment #45685874
You can customize it to get rid of all that. I set it to the "Robot" personality and a custom instruction to "No fluff and politeness. Be short and get straight to the point. Don't…
-
comment
Comment #45302397
Same. I recall the "stable volume" setting also eating cpu.
-
comment
Comment #44855931
FYI here is a list of hundreds of engineering blogs: https://github.com/kilimchoi/engineering-blogs