Wondering about other people's experiences.
Gemini 2.5 Pro vs. Claude 3.7 Sonnet: Coding Comparison
121–130 of 336 posts
Re: Gemini 2.5 Pro vs. Claude 3.7 Sonnet: Coding Comparison
#122Here is a real coding problem that I might be willing to make a cash-prize contest for. We'd need to nail down some rules. I'd be shocked if any LLM can do this: https://github.com/solvespace/solvespace/issues/1414 Make a GTK 4 version of Solvespace. We have a single C++ file for each platform - Windows, Mac, and Linux-GTK3. There is also a QT version on an unmerged branch for reference. The GTK3 file is under 2KLOC.…
Why not modularize the backend and build a better UI with tech that’s actually relevant in 2025?
Re: Gemini 2.5 Pro vs. Claude 3.7 Sonnet: Coding Comparison
#123These one-shot prompts aren't at all how most engineers use these models for coding. In my experience so far, Gemini 2.5 Pro is great at generating code but not so great at instruction following or tool usage, which are key for any iterative coding tasks. Claude is still king for that reason.
Re: Gemini 2.5 Pro vs. Claude 3.7 Sonnet: Coding Comparison
#124Earlier quoted context omitted.
For me I had to upload the library's current documentation to it because it was using outdated references and changing everything that was working in the code to broken and not focusing on the parts I was trying to build upon.
If you don't mind me asking how do you go about this? I hear people commonly mention doing this but I can't imagine people are manually adding every page of the docs for libraries or frameworks they're using since unfortunately most are not in one single tidy page easy to copy paste.
Re: Gemini 2.5 Pro vs. Claude 3.7 Sonnet: Coding Comparison
#125I'd like to see an honest attempt by someone to use one of these SOTA models to code an entire non-trivial app. Not a "vibe coding" flappy bird clone or minimal ioS app (call API to count calories in photo), but something real - say 10K LOC type of complexity, using best practices to give the AI all the context and guidance necessary. I'm not expecting the AI to replace the programmer - just to be a useful productivi…
Re: Gemini 2.5 Pro vs. Claude 3.7 Sonnet: Coding Comparison
#126Is there a less biased discussion? The OP link is a thinly veiled and biased advert for something called composio and really a biased and overly flowery view of Gemini 2.5 pro. Example: “Everyone’s talking about this model on Twitter (X) and YouTube. It’s trending everywhere, like seriously. The first model from Google to receive such fanfare. And it is #1 in the LMArena just like that. But what does this mean? It me…
If it's not astroturfing, the people who are so vocal about it act in a way that's nearly indistinguishable from it. I keep looking for concrete examples of use cases that show it's better, and everything seems to point back to "everyone is talking about it" or anecdotal examples that don't even provide any details about the problem that Gemini did well on and that other models all failed at.
Re: Gemini 2.5 Pro vs. Claude 3.7 Sonnet: Coding Comparison
#127Earlier quoted context omitted.
I mean responses like this one: I understand the desire for a simple or unconventional solution, however there are problems with those solutions. There is likely no further explanation that will be provided. It is best that you perform testing on your own. Good luck, and there will be no more assistance offered. You are likely on your own. This was about a SOCKS proxy which was leaking when the OpenVPN provider was d…
You like that? This junk is why I don't use Gemini. This isn't a feature. It's a fatal bug. It decides how things should go, if its way is right, and if I disagree it tells me to go away. No thanks. I know what's happening. I want it to do things on my terms. It can suggest things, provide alternatives, but this refusal is extremely unhelpful.
Also, don't forget that I can then continue the chat.
Re: Gemini 2.5 Pro vs. Claude 3.7 Sonnet: Coding Comparison
#128Gemini is the only model which tells me when it's a good time to stop chatting because either it can't find a solution or because it dislikes my solution (when I actively want to neglect security). And the context length is just amazing. When ChatGPT's context is full, it totally forgets what we were chatting about, as if it would start an entirely new chat. Gemini lacks the tooling, there ChatGPT is far ahead, but a…
Re: Gemini 2.5 Pro vs. Claude 3.7 Sonnet: Coding Comparison
#129 consistently 1-shots entire tickets
Uhh no? First of that's a huge exaggeration even on human coders, second, I think for this to be true your project is probably a blog.Re: Gemini 2.5 Pro vs. Claude 3.7 Sonnet: Coding Comparison
#130Earlier quoted context omitted.
Absolute golden age YouTube brain rot. I had to disable the youtube sidebar with a custom style because just seeing these thumbnails and knowing some stupid schmuck is clicking on them like an ape when they do touchscreen experiments really lowers my mood.
If you find youtubers talking about it, they all fully agree that making these thumbnails is soul draining and they are totally aware how stupid they are. But they are also aware that click-through rates fall off a cliff when you don't use them. Humans are mostly dumb, it's up to you if you want to use it to your advantage or to your detriment.
Is that true? I like to think it’s mostly kids. Honestly the world is a dark place if it’s adults doing the clicking.