Live data from Hacker News

DeepSeek V4 Pro 0813

openrouter.ai

331–340 of 493 posts

Re: DeepSeek V4 Pro 0813

#331
post #289
post #173

Why does this link to OpenRouter, which has no useful information on its own? Linking to the official API or the benchmarks would make more sense: - https://api-docs.deepseek.com/ - https://x.com/ChrisGPT/status/2087572834650407024/photo/1 (officially posted on WeChat, this is just one of many reposts)

There's no new page for this model. Hackernews didn't allow the same link be posted twice.

Is this a new rule? I've seen some popular blog posts reposted here several times and still highly upvoted.

Re: DeepSeek V4 Pro 0813

#332
post #320

Earlier quoted context omitted.

Can you explain how you used 12 billion tokens to do useful work?

(Not the original commenter.) You can rack up quite a lot of tokens if you ask it to try out a lot of things, eg for performance investigations and trying out optimisation ideas.

is there a standard pattern for this? Like spawn an agent for each technique to try?

Re: DeepSeek V4 Pro 0813

#333
post #153

Earlier quoted context omitted.

Default effort for OpenRouter. I'll try a grid of efforts... Wow, the low, medium, and high pelicans came out in surprisingly different styles: https://tools.simonwillison.net/markdown-svg-renderer#url=ht...

It's interesting that all three of those used roughly the same amount of tokens, and almost entirely output. Feels like the thinking level lever didn't alter cost at all for this specific task, even though it did change the output.

I never trust OpenRouter to forward parameters correctly and would only ever conduct benchmarks with the official api, personally.

Re: DeepSeek V4 Pro 0813

#336
post #173

Why does this link to OpenRouter, which has no useful information on its own? Linking to the official API or the benchmarks would make more sense: - https://api-docs.deepseek.com/ - https://x.com/ChrisGPT/status/2087572834650407024/photo/1 (officially posted on WeChat, this is just one of many reposts)

I don’t know about you but I find the information about prices, effective price (weighted average), providers and performance, benchmarks (down bottom) very useful. With openrouter I can even test it right away and compare with other models (use the chat functions).

Re: DeepSeek V4 Pro 0813

#338

Earlier quoted context omitted.

Curious about what you found. I agree harness for browser automation is vital - I work on https://libretto.sh

Hand-crafted adapters that sit between playwright primitives and the agent loop are the secret sauce. The goal is to insulate the agent from the raw DOM without any loss in fidelity regarding the logical business information and available actions.

+1 I have found extremely reliable systems require a mix of deterministic "adapters" is a good word for agents to actually get through the workflows I've created. I'm still amazed that it can work with both those and some "intuition" to bend the rules around the adapters if prompted.

Re: DeepSeek V4 Pro 0813

#340

Earlier quoted context omitted.

I thought it was impossible to downvote posts?

I thought it was impossible to downvote posts? User Posts can be downvoted but you need over 500 karma to have access to the downvote button. A Submission can not be downvoted.

comments can be downvoted, posts can't
Post reply on HN