Live data from Hacker News

Viewing profile — tempusalaria

tempusalaria

HN member
Joined
Thu, Sep 08, 2022, 1:52 PM UTC
HN karma
567
Public activity
131 items

About tempusalaria

No profile information was provided.

Recent public activity

  1. comment
    Comment #45608555

    Most of EA’s revenue comes from franchise games that are way below typical AAA standard. EA’s value is from IP not talent

  2. comment
    Comment #45608253

    Lots of situations, here are 2 I’ve faced recently (cannot give too much detail for privacy reasons, but should be clear enough) 1) low latency desired, long user prompt 2) functio…

  3. comment
    Comment #45608155

    All these things are designed to create lock in for companies. They don’t really fundamentally add to the functionality of LLMs. Devs should focus on working directly with model ge…

  4. comment
    Comment #45596090

    I vastly prefer the manual caching. There are several aspects of automatic caching that are suboptimal, with only moderately less developer burden. I don’t use Anthropic much but I…

  5. comment
    Comment #45280166

    A lot of the current code and science capabilities do not come from NTP training. Indeed in seems in most language model RL there is not even process supervision, so a long way fro…

  6. comment
    Comment #45179668

    Cerebras has very limited scale. Mistral has very few users so they can use cerebra’s in inference whereas OpenAI and Anthropic cannot. If mistral grows a lot they will stop using …

  7. comment
    Comment #45097394

    Fast tire changes only matter a very limited amount of the time (pretty much only if the extra time drops you a place, so there has to be 1 car/20 in a specific 1 second window on …

  8. comment
    Comment #44997392

    I imagine it runs civ 2 pretty well

  9. comment
    Comment #44971797

    WhatsApp is certainly worth less today than what they paid for it plus the extra funding it has required over time. Let alone producing anything close to ROI. Has lost them more mo…

  10. comment
    Comment #44931367

    SFT is part of the classic RLHF process though

  11. comment
    Comment #44921540

    Yes this write-up is not about agents. In fact it’s a great illustration of why the hype around agents is misplaced!

  12. comment
    Comment #44921521

    I understand that calling it ‘agentic’ is nice for marketing, but most of what is described in this blog post is not related to agents. The design patterns you describe are explici…

  13. comment
    Comment #44767010

    They may not be acting in good faith but there is extremely clear evidence that UCLA has engaged in illegal racial hiring and admissions practices and has supported antisemitism on…

  14. comment
    Comment #44692396

    if you are p testing this isn’t the case. A positive result is a much stronger assertion

  15. comment
    Comment #44652210

    The term agent is just way overloaded. This guy defines it completely differently the the big labs, and I’ve seen half a dozen different definitions in the last few months. In the …

  16. comment
    Comment #44485493

    Even as someone who is skeptical about LLMs, I’m not sure how anyone can look at what was achieved in AlphaGo and not at least consider the possibility that NNs could be superhuman…

  17. comment
    Comment #44193822

    I agree I find claude easily the best model, at least for programming which is the only thing I use LLMs for

  18. comment
    Comment #43126138

    SemiAnalysis has made up many things. They claim that a small Chinese hedge fund could acquire $1bln in GPUs, with no state support, including many sanctioned chips, then trained a…

  19. comment
    Comment #43092855

    SemiAnalysis is wrong. They just made their numbers up (among many other things they have invented - they are not to be trusted). I have observed many errors of understanding, anal…

  20. comment
    Comment #42867537

    DeepSeek v3 (where the training cost claims come from) was announced a month ago and it had no impact outside of a small circle

  21. comment
    Comment #42787352

    Texas is a world leader in renewable energy. Easy permitting, lots of space, lots of existing grid infrastructure from the o&g industry.

  22. comment
    Comment #42344032

    1) DPO did exclude some practical aspects of the RLHF method, e.g. pretraining gradients. 2) the theoretical arguments of DPO equivalence make some assumptions that don’t necessari…

  23. comment
    Comment #42266381

    The reality is that these are not culturally significant institutions and most people in London don’t care. Ordinary Londoners rarely use these markets, and they mostly sell to res…

  24. comment
    Comment #42260882

    many of these labs have more funding in theory than OpenAI. FAIR, GDM, Qwen all are subsidiaries of companies with $10s of billions in annual profits.

  25. comment
    Comment #42126758

    Airbus was a company setup by consolidating companies controlled by some of the most powerful countries in the world, which sold planes to captive state airlines and militaries con…