Live data from Hacker News

Viewing profile — zzleeper

zzleeper

HN member
Joined
Fri, Jul 31, 2009, 3:33 AM UTC
HN karma
3,077
Public activity
978 items

About zzleeper

No profile information was provided.

Recent public activity

  1. comment
    Comment #49114018

    Had to ctrl+f for someone saying this. I typically do lots of mini calls for research (100s of millions or something in that ball park). Newer models made that absolutely impossibl…

  2. comment
    Comment #49111976

    yes you are!

  3. comment
    Comment #49089871

    It means only those vetted will be allowed to use frontier models (i.e. let's pace ourselves and not share the frontier broadly)

  4. comment
    Comment #49086784

    Does Chrome use URLs you click on to help its indexer? EG if someone sends you a link to www.example.com/mysecretpage and somehow it appears in Google later. That might be a case w…

  5. comment
    Comment #49045252

    How believable is this benchmark? EG maybe opus was training on this? (You can try to identify the IP of wherever previous ARC questions came from)

  6. comment
    Comment #49041574

    > But there’s a catch. Actually, several. > Laos isn’t doing this for climate headlines. The logic is economic. > The headline writes itself. A country went nearly 100% EV overnigh…

  7. comment
    Comment #48993383

    [flagged]

  8. comment
    Comment #48993073

    Random question, but has there been any improvement in OCR/document understanding in these newer models? Last time I checked (1mo ago) SOTA was still sadly Gemini, unless you wante…

  9. comment
    Comment #48965018

    > You don’t have access to this conversation. Make sure you’re logged in to the right account, or ask the conversation owner to send you a share link.

  10. comment
    Comment #48865725

    gemini 2 and 2.5 were great models for quick-and-dirty OCR It was fine to lose 2, but 2.5 will be dearly missed as it hit the sweet spot in terms of cost-performance :/

  11. comment
    Comment #48745562

    At least in economics it can easily be 1-5 years until you go from draft to journal. In the meantime, you want a way for others to easily cite your paper, to make different revisio…

  12. comment
    Comment #48702460

    I tested Fable through Cursor; asked for ideas on how to make a data website I have less "Claude-like" (IYKYK what are the usual tells), and it spun out the most useless, Claude-li…

  13. comment
    Comment #48660817

    Would you have used something else without that constraint?

  14. comment
    Comment #48625741

    I really wonder what's up with that. Also remember the crazy Stanford guys.. did something flip in their brain or were they just always like that?

  15. comment
    Comment #48558647

    Same path as you. Went from $60 cursor plan (often exceeding it which costed more in API) to a limitless $100 codex plan where I basically say "read the markdown and implement the …

  16. comment
    Comment #48521177

    I'm pretty sure only a small fraction of grants gave this issue, and the cuts have meanwhile being very wide, without any sort of intelligent approach (I know ppl doing stuff like …

  17. comment
    Comment #48512873

    I asked it to tweak the fonts/colors of a very very simple static page and it blew through $35 (which is a lot for me lol; it's 10 days of my monthly codex plan).

  18. comment
    Comment #48498088

    I managed to write one that at least didnt had the font and colors (using 4.5) Yesterday, I prompted Fable to improve the frontend to make it look different from Claude style, gave…

  19. comment
    Comment #48475732

    I created pages with Claude before and it's very very obvious when you see one. From the font choice to the color palette, and the style of the boxes. In fact if anyone has an effe…

  20. comment
    Comment #48468851

    It's increasingly obvious that the only safeguard we got is open models and semi open ones like from China. Crazy world

  21. comment
    Comment #48468004

    Exactly.. a bit of a red flag for me..

  22. comment
    Comment #48464279

    How credible is this benchmark? does it correlated with others real world experience?

  23. comment
    Comment #48340946

    Holy F.. $3 .. once I'm done with my base cursor allocation, each nontrivial question costs $5 . And yes, I'm now switching to a mix of codex and ds4pro

  24. comment
    Comment #48325808

    Sorry that's confusing cash flow with profits, where things get amortized

  25. comment
    Comment #48325681

    I'm looking at Schwab (and saw a few others) and couldn't find anything: https://www.schwab.com/learn/story/primer-on-wash-sales I would assume this is not an ETF but sth else?