Viewing profile — COAGULOPATH
COAGULOPATH
HN member- Joined
- Thu, Dec 23, 2021, 10:40 PM UTC
- HN karma
- 442
- Public activity
- 61 items
- HN profile
- View on Hacker News ↗
About COAGULOPATH
No profile information was provided.
Recent public activity
-
comment
Comment #49253371
>I'd like to know a lot more about how that works. My guess is that it works like Gemini's SynthID: by altering the logprobs of the next token. Like, for every 10th token, instead …
-
comment
Comment #48606662
I used to make those. There was a lot of creative stuff people discovered. For example, the game's terrain was baked at launch, so you couldn't turn land into water. But someone no…
-
comment
Comment #48088900
Please do not post AI generated comments.
-
comment
Comment #47685585
In the system card they seem to dismiss this. Quotes; > (...) Claude Mythos Preview’s gains (relative to previous models) are above the previous trend we’ve observed, but we have d…
-
comment
Comment #47345094
Yes, I find LLM-written posts valueless because I can already talk to a LLM any time I want (and get the same info). It's not these commenters are the Queen of Sheba bearing a pric…
-
comment
Comment #46862132
>I've found Moltbook has become so flooded with value-less spam over the past 48 hours that it's not worth even trying to engage there, everything gets flooded out. When I filtered…
-
comment
Comment #46861979
And even if you could, how can you tell whether an agent has been prompted by a human into behaving in a certain way?
-
comment
Comment #46861973
Is it a success? What would that mean, for a social media site that isn't meant for humans? The site has 1.5 million agents but only 17,000 human "owners" (per Wiz's analysis of th…
-
comment
Comment #46861671
If the site is exposing the PII of users, then that's potentially a serious legal issue. I don't think he can dismiss it by calling it a joke (if he is). OT: I wonder if "vibe codi…
-
comment
Comment #46470761
Thanks, I didn't realize the situation was so dire.
-
comment
Comment #46200674
> In 1920, there were 25 million horses in the United States, 25 million horses totally ambivalent to two hundred years of progress in mechanical engines. But would you rather be a…
-
comment
Comment #46184871
And they do hacky things like space elements vertically using tags.
-
comment
Comment #42816785
Something I'm increasingly noticing about LLM-generated content is that...nobody wants it. (I mean "nobody" in the sense of "nobody likes Nickelback". ie, not literally nobody.) If…
-
comment
Comment #41676037
In some domains (math and code), progress is still very fast. In others it has slowed or arguably stopped. We see little progress in "soft" skills like creative writing. EQBench is…
-
comment
Comment #41553277
>Being monetarily successful does not mean you’re good or shouldn’t be criticised. Is anyone saying that Mr Beast is good and shouldn't be criticised? I can't see them.
-
comment
Comment #41553232
I think this works, not because LLMs have a "hallucination" dial they can turn down, but because it serves as a cue for the model to be extra-careful with its output. Sort of like …
-
comment
Comment #41553214
Came here hoping to find this. You will not unlock "o1-like" reasoning by making a model think step by step. This is an old trick that people were using on GPT3 in 2020. If it were…
-
comment
Comment #41535402
That's definitely weird, and I wonder how legal it is.
-
comment
Comment #41535387
It's just Ilya typing really fast.
-
comment
Comment #41535359
>but much worse (and worse even in comparison to GPT4) than English composition O1 is supposed to be a reasoning model, so I don't think judging it by its English composition abili…
-
comment
Comment #41535227
I've heard rumors that GPT4's training data included "a custom dataset of college textbooks", curated by hand. Nothing beyond that. https://www.reddit.com/r/mlscaling/comments/14wc…
-
comment
Comment #41527628
Yes, this only helps multi-step reasoning. The model still has problems with general knowledge and deep facts. There's no way you can "reason" a correct answer to "list the trackli…
-
comment
Comment #41333395
My experience is the opposite: laypeople are excessively pessimistic on LLM progress ("AI is so dumb. It tells you to put glue on pizza and eat rocks)", usually due to a remembered…
-
comment
Comment #41333326
What he means is that if you search for "Diminished by its artsiness" + "Pauline Kael" you won't find any results (except for ones related to this news story). Google is polluted w…
-
comment
Comment #41333303
Sanitarium is one of those Bad Mojo-esque games that's worth playing for how unique it is. The isometric viewpoint never really worked for me, and undercuts the immediacy of the ho…