Viewing profile — coder543
coder543
HN member- Joined
- Wed, Jul 18, 2012, 7:31 PM UTC
- HN karma
- 6,593
- Public activity
- 1,948 items
- HN profile
- View on Hacker News ↗
About coder543
Backend-focused full-stack engineer with 8+ years of experience, mostly focused on Go, Rust, and TypeScript, but open to other options.
Resume: https://drive.google.com/file/d/1VNC272B3n7ZEfppMHkm2wGgaINwYl4Av/view
Recent public activity
-
comment
Comment #49217939
I think glancing at a random snapshot from today misses all the context. Nemotron 3 is far more significant than you're giving it credit for. At this point, Nemotron 3 is really an…
-
comment
Comment #49191310
Yes, the article uses AI-assisted text, so if that offends you, no need to accuse: you can just stop reading. I have a good amount of background knowledge on all of this, and I've …
- story
-
comment
Comment #49161228
A year and a half behind Seedance 2.0? That is a bold claim that needs evidence. According to one user preference leaderboard, MiniMax H3 is already ahead of Seedance 2.0 based on …
-
comment
Comment #49122435
The weights were just released a few minutes ago: https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731
-
comment
Comment #49062216
But if you compare to Anthropic's models? The cost difference is huge. Anthropic is clearly concerned that people are realizing they are expensive, since the Opus 5 blog post dedic…
- comment
-
comment
Comment #49061820
This article seems premature to post. Right now, the price is arbitrarily set by a single provider. Why wouldn't Moonshot collect extra revenue during this exclusivity period when …
- comment
- comment
-
comment
Comment #48895258
At this point, I would not recommend ignoring Parakeet TDT 0.6b v2/v3 (english-only versus multilingual). Those models have been out for a year, give or take, and they are both acc…
-
comment
Comment #48746257
It’s not FUD. It is my actual, lived experience. FUD is false, which this is not. I use both vLLM and llama-server. vLLM is very painful, even with the Spark community docker image…
-
comment
Comment #48726910
Unsloth Studio is also very low effort, and a lot better than LM Studio in my opinion. (Performance, compatibility with Gemma 4, actually open source, etc.)
-
comment
Comment #48726593
Compared to a dynamic quant like Unsloth's UD-Q4_K_XL, which keeps some important parameters in higher precision, a basic NVFP4 quant seems to do a lot more damage to the model unl…
-
comment
Comment #48726463
> The systems but old but I’m seeing 11tks 27b, 15tks 35b MoE If that's accurate, then you must be doing something wrong/weird. On a single RTX 3090, I'm seeing substantially highe…
-
comment
Comment #48726429
> For a MBP I have 48 GB of RAM M5 Pro. It runs at about 12-14 t/s at Q4 Are you running with MTP enabled? I have seen some people on M5 hardware report 20+ t/s on Qwen3.6-27B usin…
-
comment
Comment #48690930
5.5 Pro is $30 in / $180 out: https://developers.openai.com/api/docs/pricing I think you meant 5.5. I agree it is probably the same size model. It's probably exactly built on top o…
-
comment
Comment #48674633
EDIT: It's just not even worth arguing this point, so deleting my original, much longer comment. Abstract taxonomies can claim that Taalas is CIM, but this entirely and utterly mis…
-
comment
Comment #48671785
CIM does not bake the weights into silicon. The level of optimization that you can do down to the last transistor when the weights are fixed is on an entirely different level than …
-
comment
Comment #48668251
Yes, I’m focused on the topic at hand that the person I replied to was also talking about. The person I replied to was acting as if Taalas was ancient history. I was pointing out i…
-
comment
Comment #48668039
> It's odd to me that I haven't heard anything about this approach since. It has only been four months since they unveiled their first prototype. I don't understand your confusion.…
-
comment
Comment #48668022
Taalas' first chip is for a Llama 3.1 8B quant, not a 3.1B parameter model, to clarify.
-
comment
Comment #48649350
Yeah... one of the relevant issues: https://github.com/openai/codex/issues/11940#issuecomment-45... You would think they would support their own GPT-OSS model, but, not really anym…
-
comment
Comment #48648655
Well, the reason is simple: over the past several months, it has become very difficult to use Codex with non-OpenAI models. They removed the old edit tool that didn't require OpenA…
-
comment
Comment #48648489
> I have come to consider Gemma 4 31b the best model I can self-host I'm confused. Your own results show that Gemma 4 26B A4B and Qwen3.6-27B did better in these tests? I really li…