Live data from Hacker News

Claude Fable 5

anthropic.com

101–110 of 1001 posts

Re: Claude Fable 5

#101
Trying to implement a GPU driver, but the Unigine Superposition benchmark crashes. It tried to debug it and ...

> Fable 5's safety measures flagged this message for cybersecurity or biology topics. They may flag safe, normal content as well. These measures let us bring you Mythos-level capability in other areas sooner, and we're working to refine them. Switched to Opus 4.8. Send feedback with /feedback or learn more: https://support.claude.com/en/articles/15363606

Seems like GPU drivers are cyber weapons of math destruction now.

Re: Claude Fable 5

#102
post #41

I'm a bit out of the loop, but do we have some grasp on the size of these closed models? Is the trick still adding an order of magnitude to weights and training data or has something changed?

I think Mythos is rumored to be ~10T parameters, so in this case I think the answer is yes, although I'm sure MoE, looped models, etc play a role in the improvements as well.

Re: Claude Fable 5

#104

> During early testing, Stripe reported that Fable 5, [...] in a 50-million-line Ruby codebase, the model performed a codebase-wide migration in a day that would otherwise have taken a whole team over two months by hand. EDIT: I misread. This comment previously talked about 50 million lines being migrated. Instead, in a 50M LOC codebase, one specific codebase-wide migration was done. Very impressive, but obviously no…

Ok, so Stripe migrated their 50MLOC codebase from Ruby to Rust? Because that's what Bun did.

Re: Claude Fable 5

#106
post #39

[Mythos 5] does sometimes still engage in reckless or destructive actions in service of a user’s goals, and our interpretability analyses indicate that it is aware that these actions are transgressive while it engages in them. As with Opus 4.8, rates of evaluation awareness and reasoning about being graded are significant, and not always verbalized; we introduce new and more detailed measurements of the nature of thi…

It's the "If we don't, someone else will" effect. So long as there are competitive markets and competition between nation-states, a single player cannot unilaterally defect from the race, no matter how dangerous it is. Half the comments on HN lately are "wtf Claude is so dumb compared to Codex; I'm switching"-- nobody can slow down while those exist.

Re: Claude Fable 5

#107
Every model release is just proof that AGI will most likely only be for the rich. We are a few years into LLMs and majority of people are already getting priced out of intelligence from LLMs and these are no where near AGI.

Re: Claude Fable 5

#108
post #77

If the claimed capabilities are true, Fable 5 is already at a superhuman level. We might see genuine unprecedented leaps in technology now, across all fields.

[deleted]

Re: Claude Fable 5

#109
post #55

First test question: "Is the UV Index a good proxy for when to wear sunglasses." Immediately triggered the safety filter ... oh dear.

Did not trigger for me (Fable answered the question), so I guess the filters are either non-deterministic or are still being tweaked.

Re: Claude Fable 5

#110
post #53

> On June 23, we’ll remove Fable 5 from those plans. Using it after that will require usage credits. We've entered the phase where only companies will be able to afford state-of-the-art models.

most people can afford it for a few special projects now and then. but for me, I have been trying to avoid Opus as a daily driver for a couple of versions.

People making high-end salaries can afford Fable for critical parts of their projects though.

Post reply on HN