Live data from Hacker News

The tiny corp raised $5.1M

geohot.github.io

301–310 of 331 posts

Re: The tiny corp raised $5.1M

#302

Earlier quoted context omitted.

It does not. Logical outcomes that use infinity as an intermediary are inherently not reliable. An example of this is the Ramanujan summation where 1+2+3+... results in -1/12, an outcome which is disputed among mathematicians due to the fact that we have not defined the concept of infinity properly.

> Ramanujan summation That's not a sum in the traditional sense so don't think about it this way. Infinities are used quite often in mathematics for rather mundane things. Calculus doesn't work without it. It is also quite important to the foundation of many other areas but this is often hidden unless you get into advanced works (in this sentence we're not considering a typical undergraduate Multivariate Calculus, PD…

Calculus works fine without infinity. Finitism is basically a philosophical position without practical consequences. Plenty of serious people have planted their flag there. I don't find it particularly surprising that someone who works with computers, especially at a low level, would be drawn to it.

Re: The tiny corp raised $5.1M

#303

Earlier quoted context omitted.

AMD can't be bothered. They've had literally years to fix the SW issue that allows Nvidia to stomp them in the space quarter after quarter with no fix in sight.

Being able to make stable software isn't just a matter of wanting it though. It's possible to be both bothered about something and still fail.

You know, I agree with you in a way. There are plenty of clueless teams that will produce broken software for years. I’ve been working since the early nineties and I’ve seen just incredible garbage and especially vendors who can’t stop themselves from just producing pathological orgs that produce bad software.

Look, the reality of software is that it’s actually dozens of markets, from random bloated web crap to building safety-focused critical systems. Each of those areas values different skills and has different knowledge.

But I can attest that AMDs problems - and for that matter, Intel’s since I am familiar with both - is that the companies see software outside of a very few niches as an afterthought and relatively a cost of doing business rather than a value.

It should in fact amaze you that Nvidia has pulled off being dominant. Jensen is pretty well known to have relatively poor understanding of software and Nvidia for a long time had a reputation exactly as I describe in one of my other comments. But Nvidia stock appreciation made it possible for them to become an attractive destination despite their core corporate mentality and accidentally created a company that ended up with a strong software culture.

AMD could solve this problem tomorrow with the right level of investment and a willingness to bulldoze from above the corporate politics that prevent them from having a markwt-leading software team.

Yes, it is hard. I have worked in companies like that before. I am not just saying “hey, go hire some rockstar coders” because that answer is ALWAYS bullshit. Software people are especially prone to thinking that’s the answer and that isn’t what I am saying.

But there are ways for companies in their situation to structure things to make it work out. The specifics are not appropriate for a public forum as it would identify a number of former employers who were successful in fixing their issues.

Re: The tiny corp raised $5.1M

#304
post #226

When you click on the strip link to preorder the tinybox, it is advertised as a box running LLaMA 65B FP16 for $15000. To be fair, the previous page has a bit more details on the hardware. I can run LLaMA 65B GPTQ4b on my $2300 PC (built from used parts, 128GB RAM, Dual RTX 3090 @ PCIe 4.0x8 + NVLink), and according to the GPTQ paper(§) the quality of the model will not suffer much at all by the quantization. Just sa…

What case/MB/GPUs do you use for your dual 3090 build? Liquid cooled cards?

Re: The tiny corp raised $5.1M

#305

Earlier quoted context omitted.

This. Modular and OctoML are building on top of MLIR and TVM respectively. > It's pretty hard to really beat NVIDIA for developer support though, they've invested a lot of work into the CUDA ecosystem over the years and it shows. Yup, strong CUDA community and dev support. That said, more ergonomic domain specific languages like Mojo might finally give CUDA some competition though - it's still a very high bar for sur…

There's also OpenAI Triton. People seem to miss that OpenAI is not using CUDA...

Triton outputs PTX which still requires CUDA to be installed.

Re: The tiny corp raised $5.1M

#306
post #226

When you click on the strip link to preorder the tinybox, it is advertised as a box running LLaMA 65B FP16 for $15000. To be fair, the previous page has a bit more details on the hardware. I can run LLaMA 65B GPTQ4b on my $2300 PC (built from used parts, 128GB RAM, Dual RTX 3090 @ PCIe 4.0x8 + NVLink), and according to the GPTQ paper(§) the quality of the model will not suffer much at all by the quantization. Just sa…

Are you able to memory pool two 3090s for 48gb and if so what's your setup?

I looked into this previously[1] but wasn't super confident it's possible or what hardware is required (2x x8 pcie and official SLI support?). AFAICT still would look like two GPUs to the system.

[1] https://discuss.pytorch.org/t/is-there-will-have-total-48g-m...

Re: The tiny corp raised $5.1M

#307
post #305

Earlier quoted context omitted.

There's also OpenAI Triton. People seem to miss that OpenAI is not using CUDA...

Triton outputs PTX which still requires CUDA to be installed.

Sure, but the point is that Triton is not dependent on CUDA language or frontend. Triton also outputs PTX using LLVM's NVPTX backend. Devils are in the details, but at a very high level, Triton could be ported to AMD by doing s/NVPTX/AMDGPU/. Given this, people should think again when they say NVIDIA has CUDA moat.

Re: The tiny corp raised $5.1M

#308
post #279

Earlier quoted context omitted.

I think the more interesting question is why this symptom mostly happens to "engineers". I've seen enough engineers presume they can easily become experts in law; I haven't seen many lawyers presume they can easily become experts in engineering. Why?

It's certainly not _just_ engineers; you see it in the hard sciences and medicine to an extent, as well. Someone recently posted a study purporting to show harm caused by masks to HN, say; while its authors didn't appear to include anyone with expertise in the relevant medical specialties, they did include a chemist and a veterinarian. And, if you're a fan of Matt Levine, you'll know that dentists stereotypically ten…

I've noticed it as a pretty widespread phenomenon for anyone who has the subjective experience of being competent and thinking that's enough to translate to other fields.

Super common in hot takes on politics, medical contrarianism, etc.

Though it's probably true that certain fields are more predisposed to it than others.

Re: The tiny corp raised $5.1M

#309

I was always surprised at how AMD hasn't already thrown a bunch of money at this problem. Maybe they have and are just incompetent in this area. My prediction is AMD is already working on this internally, except more oriented around PyTorch not Hotz's Tinygrad, which I doubt will get much traction.

I think AMD is going down a different path, ie. ROCm then partnering with ML frameworks further up the stack for first class support. https://pytorch.org/blog/pytorch-for-amd-rocm-platform-now-a...

AMD is not going down the path of ROCm; perhaps they claim to do so, but as evidenced by the lack of both effort and results, they clearly are not.

The parent post is surprised that they still aren't making the appropriate investments to make it work. They kind of started to do that a few years ago, but then it fell on the wayside without reaching even table stakes, which in my opinion would require providing a ROCm distribution that works out of the box for most of their recent consumer cards (i.e. those cards which the enthusiasts/students/advocates/researchers might use while choosing which software stack to learn, and afterward base corporate compute cluster purchasing decisions on whether they support the software they wrote for e.g. CUDA+Pytorch), and they seem to be failing at that.

Re: The tiny corp raised $5.1M

#310
post #33

I respect Geohot's reputation and this company looks amazing. I might be in the market to work there... except "No Remote." For such a smart guy, locking yourself out of a ton of talent by requiring software developers to be on-site in 2023 seems...out of character, to put it politely. (Rephrased, my original post was a bit too ad hominem and accumulating downvotes rapidly. I wanted to delete this entire comment but…

From Twitter Remote work is available to everyone on GitHub. If you submit a bunch of good PRs and show me you are easy to work with, I'm down to pay per project. Source https://twitter.com/realGeorgeHotz/status/166153013618397184... I'm not anti remote, I'm anti full time remote. It's hard to build a culture

The big benefit of remote is being able to live 500 miles away and/or in a different country; and that requires being full-time remote.

If he's anti full-time remote, then his pool of candidates is still limited to those who live in San Diego or very close to it.

Post reply on HN