Live data from Hacker News

Ask HN: Which large model works best?

news.ycombinator.com

1–3 of 3 posts

Re: Ask HN: Which large model works best?

#2
It depends heavily on what you are trying to do.

For pure coding tasks and refactoring, Claude 3.5 Sonnet is currently the strongest performer. It tends to hallucinate less on specific library syntax compared to the others.

However, for creative writing or "reasoning" through complex logic puzzles, I've found Gemini (specifically the Advanced/Ultra tiers) to have a more natural "voice" and better instruction following for long contexts.

GPT-4o is still the best generalist, but it feels like it has softened slightly while the others have specialized.

Re: Ask HN: Which large model works best?

#3
post #2

It depends heavily on what you are trying to do. For pure coding tasks and refactoring, Claude 3.5 Sonnet is currently the strongest performer. It tends to hallucinate less on specific library syntax compared to the others. However, for creative writing or "reasoning" through complex logic puzzles, I've found Gemini (specifically the Advanced/Ultra tiers) to have a more natural "voice" and better instruction followin…

might want to give claude-opus-4.5 a spin if you haven't