Viewing profile — Palmik
Palmik
HN member- Joined
- Sat, Nov 27, 2010, 12:45 PM UTC
- HN karma
- 2,804
- Public activity
- 528 items
- HN profile
- View on Hacker News ↗
About Palmik
No profile information was provided.
Recent public activity
- story
-
comment
Comment #49220288
Low, High and Max, obviously, can't be compared across models. They only mean the model is likely to spend less reasoning effort (~output tokens) with Low than High on the same, *s…
- story
-
comment
Comment #49110549
It used to be $0.02 per screening, jumped to $0.07
- story
-
comment
Comment #49108430
Will these be compatible with the Digital Credentials API in Chrome ( https://developer.chrome.com/blog/digital-credentials-api-or... ) or will websites be essentially locked out o…
-
comment
Comment #49070990
Based on the best available information, DeepSeek is pricing the API such that they can repay their infra capex over 10 months, while deprecating/amortizing the cost of said infra …
-
comment
Comment #49068354
You're assuming inference providers are going to sell tokens at cost. You're also assuming that the inference providers have will optimized inference engine. I haven't seen that to…
-
comment
Comment #49056216
gpt-4o is still available on the API
- story
-
comment
Comment #49011882
So, this is 'open' as in 'OpenAI', not as in 'open source'. What's the benefit of this compared to something like NOW Payments for merchants (or the myriad of alternatives)? NOW Pa…
-
comment
Comment #48965164
Commodity providers aren't a good indicator. They have margins too. Remember they ~doubled the price going from GLM 5 to GLM 5.2, despite same [1] cost of inference. [1] GLM 5.2 is…
- story
- story
-
comment
Comment #48599233
By requiring various forms of identification to use social media, it will be harder to criticize your leaders anonymously without fear of retribution.
-
comment
Comment #48424913
The company representative said that they report all users that use Graphene OS, without any additional qualifiers . Presumably after they've already uploaded their personal detail…
-
comment
Comment #48245863
DeepSeek V4's KV cache is very efficient due to its heavily compressed and sparse attention architecture. DeepSeek V3.2 which uses DSA only (sparse attention, but without compressi…
-
comment
Comment #48245835
I really hope Huawei ramps up Ascend production and DeepSeek open sources their optimized inference engine (they already open source a lot of their kernels -- kudos to them). This …
-
comment
Comment #48245824
There are several things at play: Inference stack efficiency: Many of these providers take off the shelf sglang / vllm / trtllm and hope for the best. Meanwhile DeepSeek team is kn…
- story
-
comment
Comment #47993783
Why was the title changed from "DeepSeek V4—almost on the frontier, a fraction of the price" to "DeepSeek V4—almost on the frontier"?
-
comment
Comment #47939739
Surely art also exists in textual realm.
- story
-
comment
Comment #47907670
I don't think "friendly" and "publishing benchmarks" are at odds with each other. Model makers (both open and closed weight) typically publish benchmarks against other models and w…
-
comment
Comment #47907446
Similar article for vLLM: https://vllm-website-pdzeaspbm-inferact-inc.vercel.app/blog/... Bechmarks from InferenceX (they do not have apples-to-apples setups to compare the differe…