Live data from Hacker News

Hy4 preview

tencent.com

51–60 of 256 posts

Re: Hy4 preview

#51
> Notably, Hy4 preview also contributed to its own development process, participating for the first time in the automated optimization of training methods, data strategies, evaluation frameworks, and low-level operators. The model proposed approaches, ran experiments, and iterated based on the results, with the resulting code, logs, and feedback feeding into subsequent rounds of exploration. This established an early-stage recursive self-improvement loop.

This reminds me of one of the predictions from https://ai-2027.com/ . Only that there it's "OpenBrain" doing this, not the Chinese. And the authors of that paper were also slightly wrong about "Mid 2026: China Wakes Up": China woke up already a while ago. And:

> But China is falling behind on AI algorithms due to their weaker models. The Chinese intelligence agencies—among the best in the world—double down on their plans to steal OpenBrain’s weights.

No need to steal anything, they have already caught up.

And then there's this prediction for February 2027:

> Officials are most interested in its cyberwarfare capabilities: Agent-2 is “only” a little worse than the best human hackers

I think we're past that point now, too…

Re: Hy4 preview

#52
post #46
post #32

Earlier quoted context omitted.

That's because Deepseek invented the paradigm of prompt caching, they are the SOTA when it comes these techniques. Despite them open sourcing all their research, nobody beats them. edit: I do wish openrouter would let you sort providers by Cache Hit % and Cache cost. These are the only things that matter to me at this point when choosing a provider.

Cache hit % on openrouter is not a good metric, it's mainly driven by openrouter's own provider juggling than the providers themselves

This is not true, there isn't even a way to see a cache hit % model for a specific model, that wouldn't make any sense. You are confusing what I'm saying with cache cost, that has nothing to do with effective cache hit %. I'm talking about when you click on a specific provider for a specific model, you can scroll down on the view and see their cache hit % for that model [0].

These cache Hit % are accurate, I've done a ton of testing of this myself. The cache hit % is one of the most important metrics as far as estimating cost. There are many providers with cheap cache reads, but have an effective cache hit % of 30%, making their cheaper cache pricing meaningless compared to another provider who charges more but has a 85% cache hit percentage.

[0]: https://openrouter.ai/deepseek/deepseek-v4-flash-0731?endpoi...

scroll down on the provider/model card and you'll see a field called cache hit %, its different for every provider/model.

I don't use routing on openrouter, I strictly use models with a single provider and no fallback, at least for use with harnesses its pretty dumb to route requests to multiple providers you are busting your cache every other request and increasing costs by 20-50%.

Re: Hy4 preview

#54
post #48
post #44

> [...] Let's maybe add a helmet? It could improve riding theme, but may obscure head. Maybe a small cycling cap or helmet? The user didn't ask; can add red helmet? Might be cute. But pelican with big beak; a helmet might obscure. Better maybe no. > Maybe add sunglasses? no. > Maybe add water? no. https://tools.simonwillison.net/markdown-svg-renderer#url=ht...

[flagged]

Comments that the HN community find interesting are surfaced higher.

Just tap on the [-], and upvote what you find more interesting :)

Re: Hy4 preview

#55
post #17

Is anyone here working on a problem for which current generation LLMs are inadequate, but that could possibly be solved by the next release of a first tier LLM? Or is it like bicycles? Unless your problem is named Tadej, you don't need a $13,000 bike.

This is the exact same type of comment I heard about computer hardware upgrades for three decades in a row.

“Very few people actually require a Pentium workstation, a 486 is perfectly adequate for the majority”

The logical fallacy is taking an extant distribution of “product capability” that is priced to fit what the market will bear and assuming the “next upgrade” simply tacks on a little bit more to the right hand rail of that curve.

No!

It shifts the entire curve!

Everything for everyone gets better and the top 1% of the most demanding users will continue to pay the same-ish premium.

“Nothing” will change.

Look at it this way: you can buy a $200 laptop for your kid or a $20,000 Mac with an M5 Ultra processor.

BOTH are vastly more powerful than either a $200 PC or a $20,000 “workstation” from 20+ years ago.

Look at: https://arena.ai/leaderboard/text?q=openai&utm_source=chatgp...

The “budget” 5.5 Instant model beats o1 and o3 which were “pro” models at the time of their release!

Re: Hy4 preview

#56
I'm liking where LLMs are headed:

They can do the difficult small level optimization, the boring but tedious code but cannot be tasteful.

That means I'm more valuable and more productive. Good stuff

Re: Hy4 preview

#57
post #25

[flagged]

Seriously, without China we'd just have two parasitic companies hoarding this tech and deciding whom and how is allowed to use it.

Claude and OpenAi are not allowed in Venezuela, so I thank China too and I swear to god I'll never use them and will be rooting for chinese models forever

Re: Hy4 preview

#59
post #44

> [...] Let's maybe add a helmet? It could improve riding theme, but may obscure head. Maybe a small cycling cap or helmet? The user didn't ask; can add red helmet? Might be cute. But pelican with big beak; a helmet might obscure. Better maybe no. > Maybe add sunglasses? no. > Maybe add water? no. https://tools.simonwillison.net/markdown-svg-renderer#url=ht...

Is the broken English an optimization or a byproduct of the model being developed in China?

Re: Hy4 preview

#60
post #16
post #11

Earlier quoted context omitted.

>dont do Capitalism like the rest of the AI field Like lobbying the US president to harm their competitors?

I would suggest "lobbying" is not the correct word to describe all the corruption going on in the current USA administration cesspool.

Well, it sure as hell isn't capitalism.
Post reply on HN