Viewing profile — kimjune01
kimjune01
HN member- Joined
- Sat, May 16, 2015, 7:47 PM UTC
- HN karma
- 135
- Public activity
- 151 items
- HN profile
- View on Hacker News ↗
About kimjune01
Recent public activity
-
comment
Comment #49205292
for throughput, you can use a stopwatch and a big movie transfer from one end to another
-
comment
Comment #49205287
It only takes a small minority to ruin it for the rest of us. for example, bike theft
-
comment
Comment #49205282
Vancouver dream, not achievable with a median income
-
comment
Comment #49201694
would you let taste make withdrawls from your bank account
-
comment
Comment #49172488
Thinking that harnesses and agents can improve without human involvement, even if it works, will yield a much lower growth rate than if a human gets involved in the loop.
-
comment
Comment #49151218
it should be entirely acceptable to filter out factually incorrect or unverifiable submissions without human intervention
-
comment
Comment #49151169
paper submissions require LLM disclosure
-
comment
Comment #49151155
nowhere near plateau, but right at the inflection point of diminishing returns imo
-
comment
Comment #49151087
i think it's a rite of passage to have attempted encoding thinking and the scientific process for AI/ML researchers
-
comment
Comment #49073082
the issues were already validated by the maintainers, and the maintainers merged it into their repo voluntarily. When maintainers accepted the PRs, they were the ones who found it …
-
comment
Comment #49066221
actor paradigm is huge for clamping down on agents
-
comment
Comment #49064089
as the proof of verification decreases, the value of credentials that act as shortcut proofs of human competence will decrease, too.
-
comment
Comment #49062329
i learned that no matter how good i am at writing prose, it doesnt matter if nobody reads it. so i write for AI agents instead, hoping that it'll get picked up by an agent and find…
-
comment
Comment #49060517
for small to medium sized bugs I managed to get a bit more than half of my PRs to get merged. ~90 PRs since May https://june.kim/speedrunning-open-source Verify yourself: { merged:…
-
comment
Comment #49055101
You can actually run these benches yourself, as Frontier-Bench is open source. Also have a look at these other coding benchmarks I audited. Frontier-Bench v0.1: all 74 tasks grade …
-
comment
Comment #49055090
the bottleneck for these open source repos is maintainer attention. AI seemingly does not yet improve that throughput.
-
comment
Comment #49054850
one time i told a google voice to kill itself and i could never get it to work again
-
comment
Comment #49054816
someone's promotion depends on benching harder
-
comment
Comment #49054808
it's not just social science, even AI papers that would be trivial to replicate dont
-
comment
Comment #49054802
anyone else notice that the topline numbers are effort xhigh? anybody actually use the models at those levels?
-
comment
Comment #49051469
end the fed?
-
comment
Comment #49050059
it's magic!
-
comment
Comment #49050041
i like the tiny vision but it's yet to break out from its research project phase
-
comment
Comment #49039850
vibe coding becomes much more tolerable when you disambiguate your agents with goals, constraints and heuristics up front.
-
comment
Comment #49017243
this kind of benchmark is particularly susceptible to a narrow grader, where the input request is vague and therefore a significant range of SQL output would be valid, yet graded i…