Live data from Hacker News

Viewing profile — kimjune01

kimjune01

HN member
Joined
Sat, May 16, 2015, 7:47 PM UTC
HN karma
135
Public activity
151 items

About kimjune01

In Vancouver, Canada. www.june.kim

Recent public activity

  1. comment
    Comment #49205292

    for throughput, you can use a stopwatch and a big movie transfer from one end to another

  2. comment
    Comment #49205287

    It only takes a small minority to ruin it for the rest of us. for example, bike theft

  3. comment
    Comment #49205282

    Vancouver dream, not achievable with a median income

  4. comment
    Comment #49201694

    would you let taste make withdrawls from your bank account

  5. comment
    Comment #49172488

    Thinking that harnesses and agents can improve without human involvement, even if it works, will yield a much lower growth rate than if a human gets involved in the loop.

  6. comment
    Comment #49151218

    it should be entirely acceptable to filter out factually incorrect or unverifiable submissions without human intervention

  7. comment
    Comment #49151169

    paper submissions require LLM disclosure

  8. comment
    Comment #49151155

    nowhere near plateau, but right at the inflection point of diminishing returns imo

  9. comment
    Comment #49151087

    i think it's a rite of passage to have attempted encoding thinking and the scientific process for AI/ML researchers

  10. comment
    Comment #49073082

    the issues were already validated by the maintainers, and the maintainers merged it into their repo voluntarily. When maintainers accepted the PRs, they were the ones who found it …

  11. comment
    Comment #49066221

    actor paradigm is huge for clamping down on agents

  12. comment
    Comment #49064089

    as the proof of verification decreases, the value of credentials that act as shortcut proofs of human competence will decrease, too.

  13. comment
    Comment #49062329

    i learned that no matter how good i am at writing prose, it doesnt matter if nobody reads it. so i write for AI agents instead, hoping that it'll get picked up by an agent and find…

  14. comment
    Comment #49060517

    for small to medium sized bugs I managed to get a bit more than half of my PRs to get merged. ~90 PRs since May https://june.kim/speedrunning-open-source Verify yourself: { merged:…

  15. comment
    Comment #49055101

    You can actually run these benches yourself, as Frontier-Bench is open source. Also have a look at these other coding benchmarks I audited. Frontier-Bench v0.1: all 74 tasks grade …

  16. comment
    Comment #49055090

    the bottleneck for these open source repos is maintainer attention. AI seemingly does not yet improve that throughput.

  17. comment
    Comment #49054850

    one time i told a google voice to kill itself and i could never get it to work again

  18. comment
    Comment #49054816

    someone's promotion depends on benching harder

  19. comment
    Comment #49054808

    it's not just social science, even AI papers that would be trivial to replicate dont

  20. comment
    Comment #49054802

    anyone else notice that the topline numbers are effort xhigh? anybody actually use the models at those levels?

  21. comment
    Comment #49051469

    end the fed?

  22. comment
    Comment #49050059

    it's magic!

  23. comment
    Comment #49050041

    i like the tiny vision but it's yet to break out from its research project phase

  24. comment
    Comment #49039850

    vibe coding becomes much more tolerable when you disambiguate your agents with goals, constraints and heuristics up front.

  25. comment
    Comment #49017243

    this kind of benchmark is particularly susceptible to a narrow grader, where the input request is vague and therefore a significant range of SQL output would be valid, yet graded i…