Earlier quoted context omitted.
> on a GB10 cluster I'd love to hear more about your setup. I have a single GB10 and am thinking about adding an additional one.
I'm coming up on a month with my 2x GB10 cluster. I was at 1x for a week before I pulled the trigger for the second. Initial setup was a tad annoying because I had to update their firmwares and then power cycle them to get the 200GbE link working at full speed. After setting that up, it has been pretty smooth. I don't directly deal with the cluster, usually I just have the LLM itself handle updates/stopping to load d…
Ask HN: What default model do you use and why?
101–110 of 115 posts
Re: Ask HN: What default model do you use and why?
#102I try the ChatGPT free tier as a backup second opinion sometimes, but find their newest model to be painfully rambling.
I want to like Claude and might consider subscribing to it instead, but am put off by reports of its tight usage limits. With Gemini, though, I never hit a limit unless I conduct a few Deep Research queries at the highest level.
Re: Ask HN: What default model do you use and why?
#103Re: Ask HN: What default model do you use and why?
#104Re: Ask HN: What default model do you use and why?
#105Re: Ask HN: What default model do you use and why?
#106Re: Ask HN: What default model do you use and why?
#107Re: Ask HN: What default model do you use and why?
#108Earlier quoted context omitted.
I'm coming up on a month with my 2x GB10 cluster. I was at 1x for a week before I pulled the trigger for the second. Initial setup was a tad annoying because I had to update their firmwares and then power cycle them to get the 200GbE link working at full speed. After setting that up, it has been pretty smooth. I don't directly deal with the cluster, usually I just have the LLM itself handle updates/stopping to load d…
How much does it cost per month, which provider and what do you get out of it?
What I get out of it is the ability to hand login credentials to my other computers to manage their updates, bug fixes etc. Eg. After updating my proxmox server, the nvme drive kept dying. Was able to let my local AI in to figure out and fix what was wrong (known issue). A cloud-based AI could've done it too, but I don't want to be sending internal passwords out of my network like that.
Plus, the ability to freely delegate tasks or exploration of things cloud models generally avoid. For example, I draw as a hobby, and when I'm struggling with a pose but can't quite figure out what I'm missing, I pass it into a VLM for advice, but Claude etc get unnecessarily cautious because they interpret an anatomical sketch as a naked person.