Live data from Hacker News

Viewing profile — msp26

msp26

HN member
Joined
Mon, Jun 12, 2023, 7:03 PM UTC
HN karma
1,054
Public activity
263 items

About msp26

No profile information was provided.

Recent public activity

  1. comment
    Comment #49203129

    Not sure how to fully fix this but I remember a session last week where I got so fed up mid way though reading a response that I used the following: "give me this again without jar…

  2. comment
    Comment #49202738

    No the models are just ass at communication without being directed. Try asking them to make useful diagrams for some stuff in a codebase, out of the box without excessive hand hold…

  3. comment
    Comment #49202695

    yep matches my experience completely But even fable has the annoying tendency to invent new jargon and produce an incomprehensible soup of text.

  4. comment
    Comment #49129460

    the hand drawn diagrams and highlights are charming

  5. comment
    Comment #49042984

    Asking fable to read it's own model card triggers this btw. Or asking if mitochondria is the powerhouse of the cell.

  6. comment
    Comment #49042850

    Last I heard, the feature was in beta so avoided it. But I'll definitely give it a go if it's mature now! I have been using the --watch flag to let my agent play with the notebook …

  7. comment
    Comment #49042834

    I do that a decent chunk of the time yeah especially for learning. I also have a bunch of marimo notebooks that double as clis and they're lovely. But sometimes I want to do someth…

  8. comment
    Comment #49042093

    I fucking love marimo for exploring data. However my use of it has decreased a little with how easily I can conjure disposable frontends with agents to explore one off things.

  9. comment
    Comment #48949790

    hello, please fix needing to reauth every day (sometimes with email verification). This started happening this month. It's tedious and makes switching very tempting. I'm using the …

  10. comment
    Comment #48883206

    Deeply unserious company, flip flopping on policy every week with ludicrous, sometimes invisible, guard rails on their top model. Imagine trying to make business decisions about AI…

  11. comment
    Comment #48632253

    > Summarized thinking provides the full intelligence benefits of extended thinking, while preventing misuse. > preventing misuse. Imagine not being able to read the tokens you are …

  12. comment
    Comment #48559146

    Yep agreed completely. I couldn't imagine torturing myself with a small model for local coding. But Gemma 4 31B is so fucking good for a variety of language modelling tasks.

  13. comment
    Comment #48465592

    It triggered for me when I asked "Web search for your own model card (released today) and pick out your favourite highlights from the pdf"

  14. comment
    Comment #48463978

    >Pricing for both models is $10 per million input tokens and $50 per million output tokens.

  15. comment
    Comment #48195767

    hell will freeze over before anthropic release anything meaningful to the public

  16. comment
    Comment #48033628

    Interesting, I might try that, thanks!

  17. comment
    Comment #48026231

    Google is singlehandedly carrying western open source models. Gemma 4 31B is fantastic. However, it is a little painful to try to fit the best possible version into 24GB vram with …

  18. comment
    Comment #47987662

    I like starting most of my projects on marimo notebooks now and slowly moving parts of it to the main codebase + db. By the end of it I might remove the notebook entirely but usual…

  19. comment
    Comment #47938340

    session usage limits this week feel like ass. Even when being careful to not break prefix caching.

  20. comment
    Comment #47832277

    Not necessarily with speculative decoding. Whitespace would be trivial to predict and they would petty much keep using the same amount of compute as before. I don't think that's th…

  21. comment
    Comment #47794101

    They don't have the compute to make Mythos generally available: that's all there is to it. The exclusivity is also nice from a marketing pov.

  22. comment
    Comment #47794054

    > First, Opus 4.7 uses an updated tokenizer that improves how the model processes text wow can I see it and run it locally please? Making API calls to check token counts is retarde…

  23. comment
    Comment #47491546

    > Data extraction tasks are amongst the easiest to evaluate because there’s a known “right” answer. Wrong. There can be a lot of subjectivity and pretending that some golden answer…

  24. comment
    Comment #47419255

    Man the lowest end pricing has been thoroughly hiked. It was convenient while it lasted.

  25. comment
    Comment #47349069

    I got claude to reverse engineer the extension and compare to changedetection and here's what it came up with. Apologies for clanker slop but I think its in poor taste to not attri…