Live data from Hacker News

Viewing profile — foobar10000

foobar10000

HN member
Joined
Sun, Mar 16, 2025, 2:04 AM UTC
HN karma
73
Public activity
73 items

About foobar10000

No profile information was provided.

Recent public activity

  1. comment
    Comment #49133960

    One - and I do not mean to be snarky - you can literally ask Gpt 5.6 Sol this - and if you want to see cool stuff - Fable running in their app (not website) has a view thinking but…

  2. comment
    Comment #49051729

    Glm 5.2 nvfp4 on 4 b300 with dpattn 4 and ram will get you about 20 users live at 400k context - and 60 easily if you give that server 2tb of ram and 4 nvme 8 tb drives. There are …

  3. comment
    Comment #49016414

    Or more colloquially : paperclip maximization . From OpenAI - you know, the guys who _really_ know this... Sigh... Did they finish the prompt with "And do whatever you can to get t…

  4. comment
    Comment #49016386

    This is an _amazing_ typo :) Thank you, thank you :)

  5. comment
    Comment #48964362

    3T at nxfp4 (which is most of it) is only 1.5TB of vram - so 8x288GB B300 or MI355 will do it if you are careful with context - maybe dp-attn? Certainly not TP. 2 of those together…

  6. comment
    Comment #48964329

    Don't forget that you are not really seeing the thinking tokens used - so non-trivial to count them.

  7. comment
    Comment #48940544

    Yeah, if you have a fixed llm topology, you can just effectively burns 2 top layers of the chip as Rom (model weights) - which has a per area density even better than dram - so it’…

  8. comment
    Comment #48789031

    Well, for a lot of agentic stuff nowadays, having 250k-500K context is where things live - and the benchmarks don't really show that unfortunately - but they could :)

  9. comment
    Comment #48644513

    Agreed - there was always a set of things I wanted to do that I knew the magic core for, but wanted a team of implementers for the curft, the 100k of actual testing harnesses, hype…

  10. comment
    Comment #48559691

    We are there already pretty much - if I understand your point (“How the models are wielded”) refers to the harness - which is part of model training already. Fable was trained to u…

  11. comment
  12. comment
    Comment #48248207

    Well - there is a giant push to allow non-qualified investors to invest their 401k (and roth and whatever) into the private equities market - pre-IPO companies and such. I can't sh…

  13. comment
    Comment #48235415

    A lot - and over the coming 2 years, even more. Utilization rates are under 50% across the board, and special and cheaper chips are coming out all the time for inference. And a tru…

  14. comment
    Comment #48235400

    Imagine an agent shadowing all your terminals, providing ideas and asking to run commands that will let it verify the hypotheses it comes up with, while at the same time doing rese…

  15. comment
    Comment #48159650

    Minor nit re[2]: for agentic workloads that are actually worth money - i.e., claude code and similar, things are either prefill-bound - which this does not help - or more important…

  16. comment
    Comment #48159510

    Kindof yeah - predictivity is a question though for larger layers - when trying to scale this up. But yeah, this is a "95% predictor in latent space is a 7x improvement in speed if…

  17. comment
    Comment #48158859

    Yeah, forgot about them - 100%.

  18. comment
    Comment #48158856

    I kindof agree that it is unattractive - but the regulators are perfectly happy with "EOD also introduces credit risk on the clearing house/bilateral." if it allows them to protect…

  19. comment
    Comment #48157038

    Note that _passenger aviation_ is commercially non-competitive. The big 4 US airlines make money on credit cards, not airfare : they lose money on airfare. So, most people who are …

  20. comment
    Comment #48157011

    "who do carry liability when things go wrong" -> unless one pierces the corporate veil, it's just money. Not even their money. HIPAA - unless basically stealing data - will not gen…

  21. comment
    Comment #48156260

    We are multiple orders of magnitude away from Landauer limits - so next big thing in matmul could be photonic multipliers - there’s a bunch of them coming up in the next 3? years. …

  22. comment
    Comment #48133832

    I think the one thing you are not taking into account is that the investors on average fundamentally don’t care. Scale arbitrage means that small companies are fundamentally about …

  23. comment
    Comment #48081050

    But it does allow these investors to participate in the markets without losing their shirts - and the lack of such liquidity would impact the market more so than the cost of the ri…

  24. comment
    Comment #48080397

    I mean - I'd say electricity, agriculture, steam power, metallurgy, silicon computing (cmos), atomic power, the scientific method - these are _all_ very impressive - all lead to dr…

  25. comment
    Comment #48079592

    The EOD reconciliation (and corresponding inability to settle a position in milliseconds) is a feature - it allows "obvious erroneous trade" roll-back mechanisms, etc. Very few peo…