Live data from Hacker News

Hy3

hy.tencent.com

91–100 of 125 posts

Re: Hy3

#91
post #66

Earlier quoted context omitted.

Writes pretty engaging prose, finetunes well, now MIT licensed... what's not to like? Oh and very good world knowledge for the size: better than than DS4 Flash

Do people really use 100B+ models for writing? I am no writer but to me it seems like writing is one of the easiest tasks with barely any logic or reasoning and as long as its not longer than a handful of pages I expect even 8B models to perform great.

It's pretty clear you've never experimented with it. Creative writing demands everything the model can do and more, and most problems are still unsolved. It's extremely heavy reasoning-wise, more so than coding (check e.g. Engram paper for some insights), but also needs good scattered retrieval, careful subjective training for prose quality, character, and human likeness, a ton of facts baked in, and much much more. Mode collapse is not solved. No LLM does creative writing well but historically only the absolute largest models were able to do write anything complex more or less convincingly and were creative enough.

Re: Hy3

#92
post #33

Pelican from a few days ago: https://simonwillison.net/2026/Jul/6/hy3/ - I was using the free tier on OpenRouter, which expires on July 21st. I tried the preview model 41 days ago and got a pelican with a "change pelican color" button: https://static.simonwillison.net/static/2026/hy3-preview-pel...

Curious why TFA calls out "Tencent in China". tencent/Hy3. New Apache 2.0 licensed model from Tencent in China Is there a Tencent AI lab elsewhere (MiniMax have some association with Tencent, for example)?

I think it's just their version of "Designed by Apple in California"

Re: Hy3

#93
post #33

Pelican from a few days ago: https://simonwillison.net/2026/Jul/6/hy3/ - I was using the free tier on OpenRouter, which expires on July 21st. I tried the preview model 41 days ago and got a pelican with a "change pelican color" button: https://static.simonwillison.net/static/2026/hy3-preview-pel...

Curious why TFA calls out "Tencent in China". tencent/Hy3. New Apache 2.0 licensed model from Tencent in China Is there a Tencent AI lab elsewhere (MiniMax have some association with Tencent, for example)?

Tencent seems to have subsidiaries: https://en.wikipedia.org/wiki/Tencent#Subsidiaries

Also they have large European / South African shareholders.

Re: Hy3

#94

Earlier quoted context omitted.

I did read the the full comment and I did in fact mean exactly what I wrote when I used the term "unsupervised". I think the condescension does nothing but get in the way. Try extending the benefit of the doubt. > enough to measure creativity a ton of different ways ... The things you listed seem more like temperature than creativity to me. At this point it occurs to me that this is likely yet another case of highly…

The first thing I read from you was a sardonic browbeating in response to the exact comment I gave an earnest response. And even in domains that lean heavily on "usual phrasing", like technical writing, human writing has notably higher perplexity compared to another LLM's outputs: https://www.sciencedirect.com/science/article/abs/pii/S10766... With such a low baseline for what's unusual, you do need to get the LLM wr…

> Getting the model to break out of that baseline without disrupting the model's ability to follow technical rules, maintain logic and reasoning, etc. is the difficult part.

Sure, that is also somewhat challenging and is necessary to get human sounding prose. However doing so is not sufficient to produce "creative" literature by any reasonable metric.

> you're again saying unsupervised then following up with descriptions that sure sound like you're referring to RL and supervised learning respectively this time.

Are you sure it isn't you who is confused about the usage of those terms? I merely suggested that both preparing and making use of labeled data (ie supervised learning) seemed like it would prove quite difficult here. Quoting from wikipedia (https://en.wikipedia.org/wiki/Unsupervised_learning):

> Unsupervised learning is a framework in machine learning where, in contrast to supervised learning, algorithms learn patterns exclusively from unlabeled data.

Re: Hy3

#95
post #55

Novita is offering free Hy3 on OpenRouter until July 21st https://openrouter.ai/tencent/hy3:free https://x.com/novita_labs/status/2074158304159510819

Nice

Re: Hy3

#96
There are now 3 tiers of competition:

-Fable + Gpt 5.6 Sol

-Opus + Gpt 5.6 Terra + Grok 4.5 + Muse Spark 1.1

-Open Chinese models: GLM + et family

The economics is on the Fable tier people are willing to spend a lot on it and on the Open tier you have to give it away to drive usage. The bottom tiers are also getting more and more competitive.

Re: Hy3

#97
post #49
post #33

Pelican from a few days ago: https://simonwillison.net/2026/Jul/6/hy3/ - I was using the free tier on OpenRouter, which expires on July 21st. I tried the preview model 41 days ago and got a pelican with a "change pelican color" button: https://static.simonwillison.net/static/2026/hy3-preview-pel...

Recently tried the pelican test on GPT-OSS which was probably one of the best local models of 2025. So cool to see how models have improved in the SVG pelican!

I'm skeptical, as tests go, I think that's burned out now. They could easily be training specifically to get a better pelican...

Re: Hy3

#98
post #38

Curious how people feel about this compared to DS4 Flash, given they are pretty close in size. Also curious how well it holds up to heavy quantization. DS4 Flash can currently run reasonably well on systems with ~96gb+ RAM, I wonder if Hy3 can compete there.

> given they are pretty close in size One thing that might not be obvious about about DSV4 is how much innovation the Deepseek team implemented in its architecture. When llama.cpp fully supports its lightning indexer, the full 1M context will only require about 6G of RAM. So even though they are similar in size, I believe Deepseek will be much more efficient in that regard. > I wonder if Hy3 can compete there Highly…

I have been telling people exactly this the last few months.

We have not seen the full power of deepseek v4 yet.

Re: Hy3

#99

I feel like I'm taking crazy pills with hy3, it's either benchmaxxed to hell and back or skill issue on my part but I'd rather use dense gemma. I don't think there's a single model that's wasted more of my time in recent memory.

Gemma 4 31B is underrated. It surprises me a lot.
Post reply on HN