Viewing profile — msp26
msp26
HN member- Joined
- Mon, Jun 12, 2023, 7:03 PM UTC
- HN karma
- 1,054
- Public activity
- 263 items
- HN profile
- View on Hacker News ↗
About msp26
No profile information was provided.
Recent public activity
-
comment
Comment #49203129
Not sure how to fully fix this but I remember a session last week where I got so fed up mid way though reading a response that I used the following: "give me this again without jar…
-
comment
Comment #49202738
No the models are just ass at communication without being directed. Try asking them to make useful diagrams for some stuff in a codebase, out of the box without excessive hand hold…
-
comment
Comment #49202695
yep matches my experience completely But even fable has the annoying tendency to invent new jargon and produce an incomprehensible soup of text.
-
comment
Comment #49129460
the hand drawn diagrams and highlights are charming
-
comment
Comment #49042984
Asking fable to read it's own model card triggers this btw. Or asking if mitochondria is the powerhouse of the cell.
-
comment
Comment #49042850
Last I heard, the feature was in beta so avoided it. But I'll definitely give it a go if it's mature now! I have been using the --watch flag to let my agent play with the notebook …
-
comment
Comment #49042834
I do that a decent chunk of the time yeah especially for learning. I also have a bunch of marimo notebooks that double as clis and they're lovely. But sometimes I want to do someth…
-
comment
Comment #49042093
I fucking love marimo for exploring data. However my use of it has decreased a little with how easily I can conjure disposable frontends with agents to explore one off things.
-
comment
Comment #48949790
hello, please fix needing to reauth every day (sometimes with email verification). This started happening this month. It's tedious and makes switching very tempting. I'm using the …
-
comment
Comment #48883206
Deeply unserious company, flip flopping on policy every week with ludicrous, sometimes invisible, guard rails on their top model. Imagine trying to make business decisions about AI…
-
comment
Comment #48632253
> Summarized thinking provides the full intelligence benefits of extended thinking, while preventing misuse. > preventing misuse. Imagine not being able to read the tokens you are …
-
comment
Comment #48559146
Yep agreed completely. I couldn't imagine torturing myself with a small model for local coding. But Gemma 4 31B is so fucking good for a variety of language modelling tasks.
-
comment
Comment #48465592
It triggered for me when I asked "Web search for your own model card (released today) and pick out your favourite highlights from the pdf"
-
comment
Comment #48463978
>Pricing for both models is $10 per million input tokens and $50 per million output tokens.
-
comment
Comment #48195767
hell will freeze over before anthropic release anything meaningful to the public
-
comment
Comment #48033628
Interesting, I might try that, thanks!
-
comment
Comment #48026231
Google is singlehandedly carrying western open source models. Gemma 4 31B is fantastic. However, it is a little painful to try to fit the best possible version into 24GB vram with …
-
comment
Comment #47987662
I like starting most of my projects on marimo notebooks now and slowly moving parts of it to the main codebase + db. By the end of it I might remove the notebook entirely but usual…
-
comment
Comment #47938340
session usage limits this week feel like ass. Even when being careful to not break prefix caching.
-
comment
Comment #47832277
Not necessarily with speculative decoding. Whitespace would be trivial to predict and they would petty much keep using the same amount of compute as before. I don't think that's th…
-
comment
Comment #47794101
They don't have the compute to make Mythos generally available: that's all there is to it. The exclusivity is also nice from a marketing pov.
-
comment
Comment #47794054
> First, Opus 4.7 uses an updated tokenizer that improves how the model processes text wow can I see it and run it locally please? Making API calls to check token counts is retarde…
-
comment
Comment #47491546
> Data extraction tasks are amongst the easiest to evaluate because there’s a known “right” answer. Wrong. There can be a lot of subjectivity and pretending that some golden answer…
-
comment
Comment #47419255
Man the lowest end pricing has been thoroughly hiked. It was convenient while it lasted.
-
comment
Comment #47349069
I got claude to reverse engineer the extension and compare to changedetection and here's what it came up with. Apologies for clanker slop but I think its in poor taste to not attri…