Live data from Hacker News

Viewing profile — BoorishBears

BoorishBears

HN member
Joined
Thu, Aug 18, 2016, 3:48 AM UTC
HN karma
7,245
Public activity
5,328 items

About BoorishBears

No profile information was provided.

Recent public activity

  1. comment
    Comment #49251147

    If no hand model is what's letting them sell me a high quality mouse for $45 in a two pack... I'll take whatever hands they can get

  2. comment
    Comment #49238248

    MX4 finally got me looking elsewhere until I stumbled upon this: https://ergodriven.com/products/the-vertical-handshake-mouse I got it specifically because they had a mini rant abo…

  3. comment
    Comment #49177920

    OpenAI's moderation API is multi-modal and free with no strings attached in a way that truly boggles the mind. I've put easily over a billion requests (>$100,000 by typical moderat…

  4. comment
    Comment #49165919

    For those curious, peak valuation was Dec 2021: $11.7B

  5. story
  6. comment
    Comment #49159620

    You can't beat current API pricing running Kimi K3 a single node, so as I said anyone serious is not doing that for Kimi K3 . Not sure how you're getting "no one serves models on a…

  7. comment
    Comment #49150600

    Most discourse around GPU prices is nonsense right now. Some people using Spot prices for providers who won't have Spot capacity, some people using hourly rates for instances that …

  8. comment
    Comment #49140346

    This is absolutely hilarious https://x.com/CuiMao/status/2049828401201246395

  9. comment
    Comment #49139748

    I often see this kind of passive-aggressive characterization of the process with AI that's meant to bait people who obviously are expressing their creativity with these models. I h…

  10. comment
    Comment #49129457

    They do once you're getting margin called.

  11. comment
    Comment #49129357

    There's something amusingly circular about these conversations, because clearly laypeople like me and the person you responding to are saying "that sounds like it shouldn't be allo…

  12. comment
    Comment #49117106

    > Testing whether censorship transmits through unrelated data requires that it never appear in the data. > There was zero China-sensitive content in 220 training prompts, in 176 on…

  13. comment
    Comment #49116249

    > Changing how the model thinks about the Holodomor is completely irrelevant. Your post title is literally "Distilling DeepSeek into GPT-OSS doesn't transfer censorship." Like I'm …

  14. comment
    Comment #49115633

    This seems like mildly interesting distillation work wrapped up in a nonsense attempt to drag censorship into the discussion. There's no way your It feels like you're expecting rub…

  15. comment
    Comment #49103105

    Not surprising it's a hard cutoff: they almost certainly have two infrastructure configurations for the two max sequence lengths Fewer nodes dedicated to prefill per instance, and …

  16. comment
    Comment #49027805

    I prefer its sibling library Kotlin, from the makers of the world famous Java IDE: Jetbrains Fleet

  17. comment
    Comment #48994407

    "the premium of in person" is a string of words that only means something to people in a very specific circle where everyone is constantly trying to tastemake and write in lower ca…

  18. comment
    Comment #48989177

    https://i.imgur.com/3zWdJpz.png vs https://i.imgur.com/ta6g1d2.png

  19. comment
    Comment #48989157

    Why do you think an inference provider competing for the same compute as OpenAI and Anthropic gives up those margins rather than giving a modest discount over the frontier for near…

  20. comment
    Comment #48984097

    > If they had pivoted hard to digital they could have been what today? Having the first portable-ish digital camera they could have seen the true value of Fairchild's CCD business,…

  21. comment
    Comment #48977724

    I've found if you pay close attention, models in Codex get briefly lost on where it is in the plan post-compaction. If it was in the middle of a test it often tries to resume the t…

  22. comment
    Comment #48966523

    https://qwen.readthedocs.io/en/latest/training/ms_swift.html Qwen cares enough about model identity that their training framework and docs include a preset for training on it compl…

  23. comment
    Comment #48966273

    a) Why and b) With what compute? "Why" as in, why take lower margins when Moonshot currently can't service all the demand for the model anyways. Based on past models no one is goin…

  24. comment
    Comment #48965216

    Everyone is raising the bottom. Kimi got 60% more expensive during the 2.x cycle despite staying the exact same size. Now K3 is almost 6x the cost of the original K2 checkpoint, an…

  25. comment
    Comment #48961080

    So we don't leave it all up to the parents: parents can give it, but minors also can't buy it regardless of parental views. Also give it to your kids too often and the state can st…