Live data from Hacker News

I put a datacenter GPU in my gaming PC

blog.tymscar.com

161–170 of 199 posts

Re: I put a datacenter GPU in my gaming PC

#161
Very interesting, but slight note on the fact that writer sort of forgets to calculate in the price of the 4080 it works together with.

He spend 200 to upgrade his existing setup, great nonetheless. But not "32gb gram for only 200 bucks"

Re: I put a datacenter GPU in my gaming PC

#162
This lack of driver support for hardware that’s still functional and available in reasonable quantities is annoying.

I was very close to buying a retired POWER7+ server with an ungodly amount of memory, but decided being unable to run a modern Linux kernel would be more work than I wanted to have. Modern kernels need POWER8 and above.

Re: I put a datacenter GPU in my gaming PC

#163
This lack of driver support for hardware that’s still functional and available in reasonable quantities is annoying.

I was very close to buying a retired POWER7+ server with an ungodly amount of memory, but decided being unable to run a modern Linux kernel would be more work than I wanted to have. Modern kernels need POWER8 and above.

OTOH, if these chips were fully supported, they wouldn’t hit the second hand market at the prices they do.

Re: I put a datacenter GPU in my gaming PC

#164
Very nice to see this older hardware getting repurposed. I have been running 2x Tesla V100s in a dual-core supermicro X10DRU-i server. With qwen3.6-27B-mtp I get about 35-40tok/s for inference for moderate context sizes ($100s if I had to pay claude API costs). However, the main purpose that I have to for these cards is for scientific compute, the FP64 performance (7+ TFLOPS!) is fantastic given their age, and not something you can get on even the latest consumer grade cards since Nvidia nerfed their performance after Kepler. The server lives in the basement though...it is freaking loud!

Re: I put a datacenter GPU in my gaming PC

#165
post #122

Earlier quoted context omitted.

> We might have just shot our most valuable non-AI tech products in the foot. Counterpoint: the fiber buildout during the dotcom boost. That crashed the economy pretty hard when the bubble burst, but we are still benefitting from all the dark fiber that was arranged for and built out back in that era. A lot of today's ISPs were able to grab up that fiber after the bust for cents on the dollar. Assume that OpenAI and…

You're expecting that there's going to be a supply collapse only, but there's a real risk the collapse hits both supply and demand. A lot of the current AI business is FOMO and vanity metrics. Nobody really wants to acknowledge the support tickets where the first three responses are the customer cursing because they didn't appreciate being handed off to a chatbot, or the reworks, or the compliance/policy/privacy conc…

> You're expecting that there's going to be a supply collapse only, but there's a real risk the collapse hits both supply and demand.

That was what I alluded to in the last paragraph. Semiconductor industry and everything associated with it will get screwed hard.

But in case you mean a demand collapse from the entire economy because even more people get laid off... yes, agreed. Dotcom bubble bust, here we come, full steam ahead.

> We probably don't need AI-powered support experiences that manage to be worse than actually keyword-searching your company's Confluence.

I'd pay good money for an AI that could actually ingest Confluence. In literally every organization that does not have a dedicated team to manage it on all aspects, it inevitably devolves into a tire fire. Unfortunately there's no easy way (yet?) to "after-train" a model - what I'd envision here is a nightly batch job that adds another layer to the AI model from all the information in the Confluence so searches don't incur a giant cost for the AI agent to process everything.

> We probably don't need to be spinning up coding agents to spend 15 minutes discombobulating and bibblewabbling and re-reading 82 billion tokens of context before making a two-line change that an actual developer with learned experience in the code would make in 15 seconds.

Oh we do. The stonk markets don't like it when companies employ people. People need office space, they need associated services (say, IT, fruit baskets and other amenities), they need wages, and in everywhere but the US you can't just go and fire them on a whim. The less people an organization has, the better the company looks on an investor relations press release. That is why the large AI organizations are investing untold billions of dollars... the race to be the first one that can fully replace a class of human employment. Say an SWE makes 130k/y on average - fire 100 of the 150 you have, that's 13 million dollars. That can buy you a looooot of tokens or hardware.

Re: I put a datacenter GPU in my gaming PC

#166
post #32

Earlier quoted context omitted.

Where do you think llms learned to write that way?

Because their custom training data contains an emphasis on such verbiage. It doesn't come from the God-knows-how-many TB of web content the model is pre-trained on. There, such phrasing is only a drop in the sea. But the "yes, you're right" phrases, the em dash, etc., come from the later stage, for which content is created according to some (probably overprecise) guidelines.

It's a very specific style of condescending journalism that US media has been nurturing and recycling for decades now. I was going to write this this whole comment as a parody of it, starting with some literary hook like 'Call it Ouroboros syndrome:' but I can't bring myself to add to the pile.

I have not done the textual and statistical analysis to verify this, but I feel like it's something you could trace back to east coast journalism schools and publishers mediated via television, which long predates mass adoption of AI. Think how many news articles you've read with titles like 'Anatomy of a murder' os 'Inside the meeting that changed everything.' The hooky, slightly pompous tone is something you can find back as far as the 1960s or 1970s; browsing through old issues of Readers Digest and you'll find tons of it. When I say it's mediated through television, I'm talking about both the dramatic and heavily conclusory style of fictional prosecutors and narrators, and the extremely shallow style of TV news reports (often transcribed to the web) which are only one or two sentences per paragraph. And this is before we consider the stylistic impact of ad copywriting on communication in general.

And there's something else.

The one sentence paragraph interjection, designed to refocus your attention in a surprising new direction after two paragraphs of stuff you already know. 'I never thought I'd end upere,' said Sally Nocontext, hooking you in for another paragraph or two where you try to figure out who this woman is, where she ended up, and what it has to do with the article you are already halfway through reading. After all, I've come this far, the reader through. I might as well see it through to the end.

And that's just what publishers wanted.

One sentence can also validate a truism that the reader already suspects, flattering their beliefs in their own analytical powers....

...well you get the idea. When I'm using LLMs for any sort of extended session, I find myself reaching for the same few prompts to break it of such clicheed expression; I'm especially averse to the habit of adding zippy-sounding nicknames to complex or potentially dull concepts. I don't have a favorite starting prompt, but I generally find that asking for 'a concise, academic tone' does wonders to de-fluff its output. Remember, it defaults toward being as widely accessible as possible, and much journalism is aimed at consumers with only a high school education and maybe middle-school reading comprehension, math ability, and appetite for depth over sensation.

Re: I put a datacenter GPU in my gaming PC

#167

This is great! I've been trying to get into local models for a while as I share the sentiment that local models will eventually be so good that there won't be a need to use frontier models for most coding tasks (perhaps that's already true today?). I have zero experience building computers - where would I even start? I mean, aside from the things already well documented and mentioned in the blog post.

Building computers is very easy. I would suggest watching a YouTube video to get the general gist, and then once you buy the parts, just Google for whatever doesn’t go well.

I built my first Pentium 4 one when I was like six, so I’m sure someone much older that’s into tech can do it without an issue.

There are also tons of Discord communities that are willing to help you live if you encounter any issues.

Re: I put a datacenter GPU in my gaming PC

#168
post #152

The V100 and the 4090 are based on vastly different architectures, the former uses the older Volta while the latter uses Ada. Last I checked you cannot meaningfully combine them. The 3090 is better than the V100, just get two 3090 and a NVLink.

Well I did in fact meaningfully combined them without an issue, that was the whole point of the blogpost.

Re: I put a datacenter GPU in my gaming PC

#169

Very interesting, but slight note on the fact that writer sort of forgets to calculate in the price of the 4080 it works together with. He spend 200 to upgrade his existing setup, great nonetheless. But not "32gb gram for only 200 bucks"

I did mention that in my case I had the 4080, but you do not need the 4080 whatsoever. You can run the same model with less context on a single £200 V100 or with the same amount of context on two of these.
Post reply on HN