Earlier quoted context omitted.
Writes pretty engaging prose, finetunes well, now MIT licensed... what's not to like? Oh and very good world knowledge for the size: better than than DS4 Flash
Do people really use 100B+ models for writing? I am no writer but to me it seems like writing is one of the easiest tasks with barely any logic or reasoning and as long as its not longer than a handful of pages I expect even 8B models to perform great.
Hy3
91–100 of 125 posts
Re: Hy3
#92Pelican from a few days ago: https://simonwillison.net/2026/Jul/6/hy3/ - I was using the free tier on OpenRouter, which expires on July 21st. I tried the preview model 41 days ago and got a pelican with a "change pelican color" button: https://static.simonwillison.net/static/2026/hy3-preview-pel...
Curious why TFA calls out "Tencent in China". tencent/Hy3. New Apache 2.0 licensed model from Tencent in China Is there a Tencent AI lab elsewhere (MiniMax have some association with Tencent, for example)?
Re: Hy3
#93Pelican from a few days ago: https://simonwillison.net/2026/Jul/6/hy3/ - I was using the free tier on OpenRouter, which expires on July 21st. I tried the preview model 41 days ago and got a pelican with a "change pelican color" button: https://static.simonwillison.net/static/2026/hy3-preview-pel...
Curious why TFA calls out "Tencent in China". tencent/Hy3. New Apache 2.0 licensed model from Tencent in China Is there a Tencent AI lab elsewhere (MiniMax have some association with Tencent, for example)?
Also they have large European / South African shareholders.
Re: Hy3
#94Earlier quoted context omitted.
I did read the the full comment and I did in fact mean exactly what I wrote when I used the term "unsupervised". I think the condescension does nothing but get in the way. Try extending the benefit of the doubt. > enough to measure creativity a ton of different ways ... The things you listed seem more like temperature than creativity to me. At this point it occurs to me that this is likely yet another case of highly…
The first thing I read from you was a sardonic browbeating in response to the exact comment I gave an earnest response. And even in domains that lean heavily on "usual phrasing", like technical writing, human writing has notably higher perplexity compared to another LLM's outputs: https://www.sciencedirect.com/science/article/abs/pii/S10766... With such a low baseline for what's unusual, you do need to get the LLM wr…
Sure, that is also somewhat challenging and is necessary to get human sounding prose. However doing so is not sufficient to produce "creative" literature by any reasonable metric.
> you're again saying unsupervised then following up with descriptions that sure sound like you're referring to RL and supervised learning respectively this time.
Are you sure it isn't you who is confused about the usage of those terms? I merely suggested that both preparing and making use of labeled data (ie supervised learning) seemed like it would prove quite difficult here. Quoting from wikipedia (https://en.wikipedia.org/wiki/Unsupervised_learning):
> Unsupervised learning is a framework in machine learning where, in contrast to supervised learning, algorithms learn patterns exclusively from unlabeled data.
Re: Hy3
#95Novita is offering free Hy3 on OpenRouter until July 21st https://openrouter.ai/tencent/hy3:free https://x.com/novita_labs/status/2074158304159510819
Re: Hy3
#96-Fable + Gpt 5.6 Sol
-Opus + Gpt 5.6 Terra + Grok 4.5 + Muse Spark 1.1
-Open Chinese models: GLM + et family
The economics is on the Fable tier people are willing to spend a lot on it and on the Open tier you have to give it away to drive usage. The bottom tiers are also getting more and more competitive.
Re: Hy3
#97Pelican from a few days ago: https://simonwillison.net/2026/Jul/6/hy3/ - I was using the free tier on OpenRouter, which expires on July 21st. I tried the preview model 41 days ago and got a pelican with a "change pelican color" button: https://static.simonwillison.net/static/2026/hy3-preview-pel...
Recently tried the pelican test on GPT-OSS which was probably one of the best local models of 2025. So cool to see how models have improved in the SVG pelican!
Re: Hy3
#98Curious how people feel about this compared to DS4 Flash, given they are pretty close in size. Also curious how well it holds up to heavy quantization. DS4 Flash can currently run reasonably well on systems with ~96gb+ RAM, I wonder if Hy3 can compete there.
> given they are pretty close in size One thing that might not be obvious about about DSV4 is how much innovation the Deepseek team implemented in its architecture. When llama.cpp fully supports its lightning indexer, the full 1M context will only require about 6G of RAM. So even though they are similar in size, I believe Deepseek will be much more efficient in that regard. > I wonder if Hy3 can compete there Highly…
We have not seen the full power of deepseek v4 yet.
Re: Hy3
#99I feel like I'm taking crazy pills with hy3, it's either benchmaxxed to hell and back or skill issue on my part but I'd rather use dense gemma. I don't think there's a single model that's wasted more of my time in recent memory.