Live data from Hacker News

Viewing profile — Mkengin

Mkengin

HN member
Joined
Fri, Aug 08, 2025, 8:15 AM UTC
HN karma
12
Public activity
19 items

About Mkengin

No profile information was provided.

Recent public activity

  1. comment
    Comment #49219557

    Are you also talking about personal projects? That wouldn't be enough for me for work either, but for personal use, my $20 Codex subscription is perfectly fine. Then again, as a ne…

  2. comment
    Comment #48249724

    There Seen to be more and more harness benchmarks out there, pretty interesting read: https://neuralnoise.com/2026/harness-bench-wip/

  3. comment
    Comment #47278728

    I don't think so, I use GrapheneOS and I think I can't even use the USB-C port for anything other than charging (which should be configurable).

  4. comment
    Comment #46344815

    No, I am not affiliated with the website, I just want to see more discussions based on uncontaminated benchmarks and feel that people rely too much on benchmarks that companies can…

  5. comment
    Comment #46344698

    I would assume that if a tool is there and the alternative too costly that they would use the tool instead of buring their project. Just today I stumbled over this for example, whe…

  6. comment
    Comment #46344686

    Not for coding, but today I stumbled upon these two building their passion project using GenAI, which would otherwise perhaps not be possible: https://reddit.com/comments/1prqfsu

  7. comment
    Comment #46344670

    It doesn't have to be hyped to be used, for example today I found these two building their passion project using GenAI, which would otherwise maybe not possible, who knows: https:/…

  8. comment
    Comment #46344597

    This is just one example, but today I found this where two people build their passion project using GenAI for image generation (+ photoshop), maybe otherwise this project wouldn't …

  9. comment
    Comment #46319914

    Though this Codex version isnt on the leaderboard, GPT-5.2-Medium already seems to be a bit better than Opus 4.5: https://swe-rebench.com/

  10. comment
    Comment #46319880

    At least on swe-rebench it does pretty well: https://swe-rebench.com/

  11. comment
    Comment #46319840

    Your experience seems to match the recent results from swe-rebench: https://swe-rebench.com/

  12. comment
    Comment #46319821

    According to SWE-Rebench Anthropic and OpenAI are really close in performance, while GPT-5.2 costs less than half the cost of CC per problem. https://swe-rebench.com/

  13. comment
    Comment #46057003

    Interesting. So similar to the vision encoder + projector in VLMs?

  14. comment
    Comment #46040669

    I am eagerly awaiting swe-rebench results for November with all the new models: https://swe-rebench.com/

  15. comment
    Comment #46040641

    I like this one: https://swe-rebench.com/

  16. comment
    Comment #45725802

    Or use RL to beat any AI detectors: https://reddit.com/r/LocalLLaMA/comments/1lnrd1t/you_can_jus...

  17. comment
    Comment #45471991

    https://arxiv.org/abs/2311.13600 https://arxiv.org/abs/2410.22911 https://arxiv.org/abs/2409.16167

  18. comment
    Comment #44835536

    Thank you for testing, I will test GPT-OSS for my use case as well. If you're interested I have 8 GB VRAM, 32 GB RAM and get around 21 token/s with tensor offloading, I would assum…

  19. comment
    Comment #44834767

    Why Qwen2.5 and not Qwen3-30B-A3B-Thinking-2507 or Qwen3-Coder-30B-A3B-Instruct?