Viewing profile — dandaka
dandaka
HN member- Joined
- Fri, Jun 26, 2020, 10:21 AM UTC
- HN karma
- 132
- Public activity
- 103 items
- HN profile
- View on Hacker News ↗
About dandaka
Recent public activity
-
comment
Comment #49255958
This is exactly what happened in Moscow after same tech was launched. It is much 'safer' on streets, since every entrance door and traffic intersection is equipped with cameras now…
-
comment
Comment #49194755
1/ LLMs are doing increasingly bigger share of work of SWE with an increasing success rate 2/ SWE are doing way more than "writing code", and since those areas are less "computatio…
-
comment
Comment #49146335
Whenever I have this choice "destroy or sustain any and all life", I always pick "sustain". Problem is, I don't often meet such an option to choose. Maybe your life is different.
-
comment
Comment #49143872
There is not much we can do/build with nature. Even if everyone would take idea of "protection/inspiration", there are not many sustainable business models towards that direction. …
-
comment
Comment #49122279
Text classification, structured data extraction, rewrite
-
comment
Comment #49107824
check transcribe.cpp from Handy, I see Qwen3-ASR in the list
-
comment
Comment #49107812
I am not only talking about "long context" performance (context rot), but also about noise that is confusing model about its goal (from correctly extracting operator's intent). I t…
-
comment
Comment #49107795
Great summarization of current capabilities of models, thank you
-
comment
Comment #49103610
Models work best when they have short instructions and no noise. "Conversation [may] contain" also means "conversation has a lot of noise". It degrades performance and increases co…
-
comment
Comment #49100046
How does it compare to models from handy.computer and whisprflow?
-
comment
Comment #49081546
Can we use those metrics in a review pass for every PR? So review agent will have to pinpoint new complexity and propose refactoring or justify increase?
-
comment
Comment #49050748
100% with you on the taste and nutrition. I think you forget sustainability at scale. I love the idea of eating local food, but with current levels of personal productivity and hom…
-
comment
Comment #49047565
Why are people obsessed with idea of "produce their own X"? When we learned to specialize, we have achieved low cost, scale and sophistication.
- story
-
comment
Comment #48962828
Few cases I have found very useful myself 1/ Using GUI software. My agents are using headful Google Chrome and Figma. It helps a lot to have separate environment, which is not inte…
-
comment
Comment #48853149
Not at all, we love them all with Chinese labs. And wish them to continue competing and not winning. That is how we get best models, lower prices and better availability.
-
comment
Comment #48837898
do they have a website? I have found only paper PDF and it seems more general than SWE
-
comment
Comment #48837708
What is considered SOTA for SWE benchmarks now?
-
comment
Comment #48835192
> calculating the exact discount you get using a subscription could be difficult Why can't you derive this discount from experiments? Example of such research https://she-llac.com/…
-
comment
Comment #48834976
Can I connect it to my skills/tools? Example case, I have a knowledge base and event log in my company. I need a brainstorm companion, which will have full access to this knowledge…
-
comment
Comment #48810603
Another important benchmark would be — cost per benchmark task using subscription tokens. Since most of us are using subscriptions and cost per token there is quite different from …
-
comment
Comment #48760421
Naming is worrisome!
-
comment
Comment #48757758
Codex, Claude Code, ZAI — they continue work in headless mode, when you close your laptop, if you have connected to remote machine
-
comment
Comment #48716960
What is your top chart? I am using GUI harnesses from model providers and not happy with any of them.
-
comment
Comment #48696454
Every country is moving in that direction