Viewing profile — PhilippGille
PhilippGille
HN member- Joined
- Fri, Aug 04, 2017, 7:36 AM UTC
- HN karma
- 1,293
- Public activity
- 412 items
- HN profile
- View on Hacker News ↗
About PhilippGille
meet.hn/city/de-Leipzig
Recent public activity
-
comment
Comment #49200256
98.33 according to https://mrshu.github.io/github-statuses/
-
comment
Comment #49121231
Yes that's my point. The old and the new version are different in capabilities, but now when someone talks about DeepSeek V4 Flash (in benchmarks, on inference providers), you don'…
-
comment
Comment #49121207
That's what I mean. On DeepSeek it's now just `deepseek-v4-flash`, while OpenRouter calls it `deepseek/deepseek-v4-flash-0731`, so now when someone talks about DeepSeek V4 Flash, l…
-
comment
Comment #49120152
The previous V4 version wasn't called “Preview” by most inference providers. For example, the OpenRouter model slug was `deepseek/deepseek-v4-flash`. So now there will be confusion…
-
comment
Comment #49113834
Depends on the reasoning effort, see https://deepswe.datacurve.ai (add Luna via model selection drop down, if it's not shown by default)
- story
- story
- story
-
comment
Comment #49013421
The project looks very interesting, thanks for sharing! You seem to have created a new GitHub account just for this project a week ago. Do you have any other GitHub accounts that e…
-
comment
Comment #48965522
Handy already supports streaming transcription models, and you can see the words in the small Handy pop-up while you are talking. So in general this definitely works. Handy is just…
-
comment
Comment #48853510
That was the case for early models (Llama etc), but they got much better since then. Not perfect, but good enough. This is from Ministral 3 14B, a 2025 model without reasoning, tha…
-
comment
Comment #48751806
Currently this is for payments with stablecoins. For Bitcoin / Lightning these kind of pay-per-request API paywalls have existed for many years already (e.g. my own from 8 years ag…
-
comment
Comment #48670501
> Kimi and GLM models have coined a new term: Thinkslop. > [...] > So for now I'm happy with just two models: GPT and DeepSeek. 1. DeepSeek V3.2, V4 Flash, V4 Pro, at high or max t…
-
comment
Comment #48457372
The interesting bits on how they achieved it: > On the model side, we applied FP4 quantization > introduced DFlash, an efficient speculative decoding method based on block-level ma…
-
comment
Comment #48353654
The blog post has more info: https://www.minimax.io/blog/minimax-m3
-
comment
Comment #48329983
Do you mean MiMo V2 Flash? V2.5 doesn't have a Flash version.
-
comment
Comment #48114889
It's in the article: > HTTP also allows the DuckDB-Wasm distribution to speak Quack natively! So DuckDB running in a browser can e.g., directly connect to a DuckDB instance running…
-
comment
Comment #48073497
Both the original Markdown spec [1] as well as CommonMark [2] clearly specify support for inline HTML. With that you can kind of get the best of both words depending on your use ca…
- story
-
comment
Comment #48052926
On max it uses more than twice as many tokens as on high when running the ArtificialAnalysis benchmark suite, and then it's indeed the model with the highest token usage (among the…
-
comment
Comment #48018649
Benchmarks only paint part of the picture, but it's still a decent place to start looking into recent models: https://huggingface.co/spaces/mteb/leaderboard
-
comment
Comment #47890770
When you say "Gemini", which exact model do you mean? You know there are several and they vary a lot in how capable they are? Pro 3.1 Preview, 2.5 Pro (their latest non-preview pro…
-
comment
Comment #47737331
> C# [...] only really works properly in Windows What do you mean with this? Maybe you are thinking of the old ".NET Framework" runtime, which only runs on Windows? Nowadays there …
-
comment
Comment #47737064
He specifically mentions that he is using GitHub Copilot because of how Microsoft bills per request instead of token.
-
comment
Comment #47714298
> it is possible with some software to have everything massively cached, with the cloud doing that, with the origin server in my basement, only accessible from the allowed cache ar…