Live data from Hacker News

Viewing profile — zhyder

zhyder

HN member
Joined
Tue, Mar 20, 2007, 11:44 PM UTC
HN karma
2,769
Public activity
515 items

About zhyder

Cofounder of http://uphop.ai : one-on-one AI trainers for team training

Former Googler, though started my SW career by learning on HN!

https://in.linkedin.com/in/zohairhyder/

Email: same user id as HN @ my former company's consumer email domain

Recent public activity

  1. comment
    Comment #48835277

    A big part of this announcement does seem to be _delegation_ in the background; they give the example of web search but that could be any tool. I haven't tried it yet either but so…

  2. comment
    Comment #48287855

    Aside from the issue of platform owners (Apple, Google, Microsoft) offering storage sync as an integrated feature, which others have pointed out, the other reason growth is limited…

  3. story
  4. comment
    Comment #47250122

    Looks like the best display you can get in laptops at this price: 2408x1506 resolution, 500 nits, antireflective coating (!). And bonus points for no silly notch.

  5. comment
    Comment #47220221

    I guess it could warn about it but the VM sandbox is the best part of Cowork. The sandbox itself is necessary to balance the power you get with generating code (that's hidden-to-us…

  6. story
  7. comment
    Comment #47170915

    Model card: https://deepmind.google/models/model-cards/gemini-3-1-flash-... Pretty close to Gemini 3 Pro Image (aka Nano Banana Pro) in most benchmarks, even without thinking+searc…

  8. story
  9. comment
    Comment #47076596

    Agree, can't wait for updates to the diffusion model. Could be useful for planning too, given its tendency to think big picture first. Even if it's just an additional subagent to d…

  10. comment
    Comment #47075664

    Surprisingly big jump in ARC-AGI-2 from 31% to 77%, guess there's some RLHF focused on the benchmark given it was previously far behind the competition and is now ahead. Apart from…

  11. comment
    Comment #47063656

    "the value of a human eyeball" / attention is and always will be the limited resource. But I wish the way the economy worked wasn't that attention is sold for money, which makes mo…

  12. comment
    Comment #46970704

    Hmm the whole point of checkpoints seems to be to reduce token waste by saving repeat thinking work. But wouldn't trying to pull N checkpoints into context of the N+1 task be MUCH …

  13. comment
    Comment #46929408

    So 2.5x the speed at 6x the price [1]. Quite a premium for speed. Especially when Gemini 3 Pro is 1.8x the tokens/sec speed (of regular-speed Opus 4.6) at 0.45x the price [2]. Thou…

  14. comment
    Comment #46680352

    Love it. Wonder if it's viable for citizen journalism in warzones and areas of civil unrest, with the larger size of photos (and short videos), given the inherently slow transfer r…

  15. comment
    Comment #46581786

    Sounds like antirez, simonw, et al are still advocating reviewing the code output of these agents for now. But presumably soon (within months?) the agents will be good enough such …

  16. story
    Ask HN: How do you review the code from agents?

    Engineers are increasingly setting up coding agents to run continuously, some running multiple agents in parallel. I've struggled to do that because I've struggled to build confide…

  17. comment
    Comment #46515851

    Most car manufacturers made this mistake because they started mimicking the then leader for innovation (and customer satisfaction), Tesla, too much. General cautionary tale: just c…

  18. comment
    Comment #46317081

    "Almost anyone can prompt an LLM to generate a thousand-line patch and submit it for code review. That’s no longer valuable. What’s valuable is contributing code that is proven to …

  19. comment
    Comment #46304419

    Have you tried them with providing a grounding resource, e.g. attaching a file to ChatGPT or NotebookLM? Yes need some human expert to create (or curate) that grounding resource in…

  20. comment
  21. comment
    Comment #46302925

    End of an era: video (with broadband Internet penetration) was the best tool we had for 15+ years. But LLMs are now good enough, including in image+infographic generation and factu…

  22. comment
    Comment #46302453

    Glad to see big improvement in the SimpleQA Verified benchmark (28->69%), which is meant to measure factuality (built-in, i.e. without adding grounding resources). That's one bench…

  23. story
  24. story
  25. comment
    Comment #46235511

    Big knowledge cutoff jump from Sep 2024 to Aug 2025. How'd they pull that off for a small point release, which presumably hasn't done a fresh pre-training over the web? Did they fi…