Live data from Hacker News

Viewing profile — ycui7

ycui7

HN member
Joined
Wed, Sep 21, 2011, 11:44 PM UTC
HN karma
75
Public activity
45 items

About ycui7

No profile information was provided.

Recent public activity

  1. comment
    Comment #49223504

    on one side, deepmind makes a lot of advancement in science related application. but, on the commercial side, they struggle to compete with other major LLM providers.

  2. comment
    Comment #49217986

    it is funny when people say i am struggling to spend money.

  3. comment
    Comment #49202634

    they need a competent gov contractor. recognizing license plate is a fully solved problem many years ago.

  4. comment
    Comment #49202415

    so qwen3.x-27b on hardware? or better deepseek-v4-flash on hardware .

  5. comment
    Comment #49151096

    Can they still go public ? MiniMax M3 Pro is also coming, then DeepSeek-v4-Pro GA, then GLM5.5. There will only be bad news for them in the coming few weeks/months.

  6. comment
    Comment #49125080

    the rational in one’s mind is similar to buying expensive supercar but no driving it daily. owning a few GPUs is a lot cheaper than supercars.

  7. comment
    Comment #49124947

    Huawei Ascend NPU

  8. comment
    Comment #49124941

    my AC is noiser than my GPU server.

  9. comment
    Comment #49124914

    you don’t need new weight. try vllm-moet from github. it will autogenerate 2-bit plane.

  10. comment
    Comment #49124092

    K3 is natively trained to mxfp4, if they cannot get a hold of Blackwell chip, it is meaningless. Hopper does not do native 4-bit floating math. Either they have Blackwell with nati…

  11. comment
    Comment #49124059

    For people with single RTX PRO 6000 96GB or DGX Spark 128GB, vllm-moet is a very good engine, although lesser known. It auto generate a symmetric 2-bit plane for inference and also…

  12. comment
    Comment #49030867

    and it was created by Chinese born Professor and Student.

  13. comment
    Comment #49006065

    so GLM won?

  14. comment
    Comment #48987487

    discounted competitor could cheat. they can offer subpar model response and sell it as deepseek-v4. it is uneconomical to prove inference providers are cheating, so they get away w…

  15. comment
    Comment #48985138

    feels like the American domestic manufacturing is done, there is no hope to save it.

  16. comment
    Comment #48971545

    it takes extra effort to open source a model even if you had it running internally. even traditional software takes extra effort to get released as open source.

  17. comment
    Comment #48971505

    every cloud provider trains on your data, regardless of what they promise. real user interaction is the best reinforcement-learning trace.

  18. comment
    Comment #48971491

    OpenAntrophic and OpenOpenAI ?

  19. comment
    Comment #48855873

    China successfully recovers Long March 10B rocket following maiden flight, marking a breakthrough in rocket reusability

  20. story
  21. comment
    Comment #48847013

    this is not subsidizing. this is way too expensive for a no-name model.

  22. comment
    Comment #48840184

    there is an algorithm called quick-select. getting median of an array should not require full sort of the whole array, only a partial sort is needed to get the median. quick-select…

  23. comment
  24. comment
    Comment #48768743

    24-bit was created because microphone want to record large dynamic range without gain switching circuit. 96kHz was created to better reproduce 20kHz high frequency, so the digital …

  25. comment
    Comment #48669427

    Those resellers are simply just selling Kimi K2.5 or GLM5.1 as counterfeit Opus. We, Chinese, know how to play the counterfeit game for a long time in so many industry.