Live data from Hacker News

Mercury 2.5

inceptionlabs.ai

41–50 of 58 posts

Re: Mercury 2.5

#41

Oh great a new model annoncemzzzzzz ZZZZZZZZZ

Not just a model - It's a diffusion model. Instead of Next Token prediction it builds the entire page at once and then denoises it. Kind of mind blowing when you watch it happen.

Where can you watch it happen? Is there a video visualizing the diffusion process on text?

Re: Mercury 2.5

#44
Ridiculously fast on OpenRouter, just subjectively it's a really strange experience because I've never seen a model respond or execute that quickly.

Re: Mercury 2.5

#46

Ridiculously fast on OpenRouter, just subjectively it's a really strange experience because I've never seen a model respond or execute that quickly.

Try Cerebras. When I think about how speed of generation is another variable to tweak for "intelligence", it seems like this speed is best used for searching for solutions in a problem space and then validating and discarding and keeping what is best. Being intelligent at the Fable level, but what if the Fable level machine could think at 100x? What does that mean: perhaps it means more parallel "experiments" for solutions in the token/generation/hyper-dimensions of the latent space.

Re: Mercury 2.5

#47

I like the model. FYI: "If you do not want us to use your User Submissions to train our models, you can opt-out by setting the ‘Improve the model for everyone’ option under User Settings in the API Platform to OFF."

I have ZDR enabled globally on openrouter, and was able to use the model.

It likely depends on how you access it. With AI and "Free usage allowance", the price tends to be your soul.

Re: Mercury 2.5

#48

Ridiculously fast on OpenRouter, just subjectively it's a really strange experience because I've never seen a model respond or execute that quickly.

Try https://chatjimmy.ai/ from Taalas. There is an emergent space for super-fast-models esp finetuned or guardrailed to solve very specific latency sensitive tasks.

Re: Mercury 2.5

#49

Earlier quoted context omitted.

Not just a model - It's a diffusion model. Instead of Next Token prediction it builds the entire page at once and then denoises it. Kind of mind blowing when you watch it happen.

Where can you watch it happen? Is there a video visualizing the diffusion process on text?

This is an older post from June about a Google model, but the same general process applies and they have a couple demo videos:

https://blog.google/innovation-and-ai/technology/developers-...

Post reply on HN