Live data from Hacker News

Viewing profile — PhilippGille

PhilippGille

HN member
Joined
Fri, Aug 04, 2017, 7:36 AM UTC
HN karma
1,293
Public activity
412 items

About PhilippGille

https://github.com/philippgille

meet.hn/city/de-Leipzig

Recent public activity

  1. comment
    Comment #49200256

    98.33 according to https://mrshu.github.io/github-statuses/

  2. comment
    Comment #49121231

    Yes that's my point. The old and the new version are different in capabilities, but now when someone talks about DeepSeek V4 Flash (in benchmarks, on inference providers), you don'…

  3. comment
    Comment #49121207

    That's what I mean. On DeepSeek it's now just `deepseek-v4-flash`, while OpenRouter calls it `deepseek/deepseek-v4-flash-0731`, so now when someone talks about DeepSeek V4 Flash, l…

  4. comment
    Comment #49120152

    The previous V4 version wasn't called “Preview” by most inference providers. For example, the OpenRouter model slug was `deepseek/deepseek-v4-flash`. So now there will be confusion…

  5. comment
    Comment #49113834

    Depends on the reasoning effort, see https://deepswe.datacurve.ai (add Luna via model selection drop down, if it's not shown by default)

  6. story
  7. story
  8. story
  9. comment
    Comment #49013421

    The project looks very interesting, thanks for sharing! You seem to have created a new GitHub account just for this project a week ago. Do you have any other GitHub accounts that e…

  10. comment
    Comment #48965522

    Handy already supports streaming transcription models, and you can see the words in the small Handy pop-up while you are talking. So in general this definitely works. Handy is just…

  11. comment
    Comment #48853510

    That was the case for early models (Llama etc), but they got much better since then. Not perfect, but good enough. This is from Ministral 3 14B, a 2025 model without reasoning, tha…

  12. comment
    Comment #48751806

    Currently this is for payments with stablecoins. For Bitcoin / Lightning these kind of pay-per-request API paywalls have existed for many years already (e.g. my own from 8 years ag…

  13. comment
    Comment #48670501

    > Kimi and GLM models have coined a new term: Thinkslop. > [...] > So for now I'm happy with just two models: GPT and DeepSeek. 1. DeepSeek V3.2, V4 Flash, V4 Pro, at high or max t…

  14. comment
    Comment #48457372

    The interesting bits on how they achieved it: > On the model side, we applied FP4 quantization > introduced DFlash, an efficient speculative decoding method based on block-level ma…

  15. comment
    Comment #48353654

    The blog post has more info: https://www.minimax.io/blog/minimax-m3

  16. comment
    Comment #48329983

    Do you mean MiMo V2 Flash? V2.5 doesn't have a Flash version.

  17. comment
    Comment #48114889

    It's in the article: > HTTP also allows the DuckDB-Wasm distribution to speak Quack natively! So DuckDB running in a browser can e.g., directly connect to a DuckDB instance running…

  18. comment
    Comment #48073497

    Both the original Markdown spec [1] as well as CommonMark [2] clearly specify support for inline HTML. With that you can kind of get the best of both words depending on your use ca…

  19. story
  20. comment
    Comment #48052926

    On max it uses more than twice as many tokens as on high when running the ArtificialAnalysis benchmark suite, and then it's indeed the model with the highest token usage (among the…

  21. comment
    Comment #48018649

    Benchmarks only paint part of the picture, but it's still a decent place to start looking into recent models: https://huggingface.co/spaces/mteb/leaderboard

  22. comment
    Comment #47890770

    When you say "Gemini", which exact model do you mean? You know there are several and they vary a lot in how capable they are? Pro 3.1 Preview, 2.5 Pro (their latest non-preview pro…

  23. comment
    Comment #47737331

    > C# [...] only really works properly in Windows What do you mean with this? Maybe you are thinking of the old ".NET Framework" runtime, which only runs on Windows? Nowadays there …

  24. comment
    Comment #47737064

    He specifically mentions that he is using GitHub Copilot because of how Microsoft bills per request instead of token.

  25. comment
    Comment #47714298

    > it is possible with some software to have everything massively cached, with the cloud doing that, with the origin server in my basement, only accessible from the allowed cache ar…