Live data from Hacker News

Viewing profile — XCSme

XCSme

HN member
Joined
Sun, Dec 28, 2014, 6:47 PM UTC
HN karma
2,940
Public activity
3,022 items

About XCSme

Building self-hosted web analytics.

Self-Hosted Analytics with Heatmaps and Session Recordings: https://www.uxwizz.com

WordPress Analytics: https://www.wplytic.com

X (Twitter): @XCSme

Recent public activity

  1. comment
    Comment #49217887

    As a solo dev, this gives me hope. I feel like I have an advantage over big companies, if I can use the best models on a subscription and not worry about costs much, when they can'…

  2. comment
    Comment #49208480

    My question was more about more complex problems, which no seem to be multi-turn somehow, or maybe just the harnesses make it look that way. I am curious what the drop in thoughput…

  3. comment
    Comment #49208440

    Yeah, makes sense, if it's good for very small models only, then there's no point, as those van already run on cheap consumer hardware. Yet, maybe it can work well enough, so that …

  4. comment
    Comment #49205779

    I asked a LLM after posting my comment, to see if I had a genius idea or not,just for it to tell me the same as you, that's now they work already...

  5. comment
    Comment #49205731

    My concern is that reasoning could involve some sequential steps that instant models don't. Not sure if modern models "think" only by outputting blocks, or there is a more complex …

  6. comment
    Comment #49204573

    It was just a random example, you could think of it as being a lot more complex (detect which type of food it is, what detergent to use, how much water, remember patterns, learn ov…

  7. comment
    Comment #49204472

    So local personalized ads? Not sure if that's better or worse than online personalizaed ads...

  8. comment
    Comment #49203940

    Why not have some a device/hardware that programs itself on-boot. Sort of a FPGA, that (electrically) arranges the connections on-boot, and then it's like a static inference chip.

  9. comment
    Comment #49203917

    I think this would make sense for consumer hardware, not for AI companies. AI companies constantly update/change stuff, new models come out, new requirements, etc. But if you ship …

  10. comment
    Comment #49203902

    Wait, is it even thinking? Or is it an instant model?

  11. comment
    Comment #49203887

    Wow, that's instant, crazy.

  12. comment
    Comment #49154084

    I am surprised that they keep going with it, seeing how fast it improves and basically soon running themselves too out of business. What's even their end goal? Open source models m…

  13. comment
    Comment #49151019

    If they trained it well, and can do computer use, it will be a new era. Companies can keep PCs, put Qwen 3.8 27b on it and get rid of the employees, lol...

  14. comment
    Comment #49121972

    Can't really use it now, without giving away your data: > Trains: this provider may use prompts for training and may retain prompt data.

  15. comment
    Comment #49121946

    Or maybe not: https://news.ycombinator.com/item?id=49119559

  16. comment
    Comment #49119888

    I think that with this change, DeepSeek v4 Flash has finally been dethroned.

  17. comment
    Comment #49119369

    Wasn't one of the main original points of LLMs to be creative? To create stories, creative writing?

  18. comment
    Comment #49080820

    Is this like Runescape free armour trimming? You send your gpu, and get it back with 2x memory?

  19. comment
    Comment #49069986

    Depends on whether the models report the correct amount of tokens. 5.5 Sol repors 10x fewer reasoning tokens than Kimi k3. If it is correct, than it unlikely has those doubt issues…

  20. comment
    Comment #49069194

    Yeah, they thing forever and doubt everything "wait but" for 200k tokens for almost any question.

  21. comment
    Comment #49069187

    So nowadays the hardware and hosting providers must be in an optimization race, whoever can make the model just a bit smaller or more efficient (to fit on fewer/less powerful cards…

  22. comment
    Comment #49044428

    One of the best hamsters [0]. Again, their "none" version costs more than "low", and says zero reasoning tokens, makes no sense[1]. As always, the "low" version seems to be the bes…

  23. comment
    Comment #49043813

    Twice the cost for 4% more intelligence, is it worth it?

  24. comment
    Comment #49031287

    I have a spare 3090 that I want to use to off-load some tasks from Claude to a local model (probably Qwen 3.6 27b), any success with that? Is it good enough to follow some tasks, c…

  25. comment
    Comment #49031282

    I thought login-protected apps are not allowed on Show HN.