Viewing profile — brucethemoose2
brucethemoose2
HN member- Joined
- Mon, Feb 13, 2023, 5:50 AM UTC
- HN karma
- 7,874
- Public activity
- 3,307 items
- HN profile
- View on Hacker News ↗
About brucethemoose2
No profile information was provided.
Recent public activity
- comment
-
comment
Comment #39876619
Some phones have limiters to keep the battery at 60%-80%. I believe most can do this with the right software. It's not a fix, but it should extend the life considerably.
-
comment
Comment #39875777
Another benefit: modern smartphones have large GPUs, large media blocks, and fast RAM. With the right software , they can be a surprisingly powerful AI host or transcoding server.
-
comment
Comment #39870001
Real world GPU performance is hugely influenced by hand optimization of the CUDA kernels.
-
comment
Comment #39869240
Power/Weight is extremely high. A tiny wankel will do the job, and weight is everything on cars. It does prefer a narrow RPM band, which is fine. Reliability is the biggest concern…
-
comment
Comment #39857290
Yeah, its an unspoken but rampant thing in the llm community. Basically no one respects licenses for training data. I'd say the majority of instruct tunes, for instance, use OpenAI…
-
comment
Comment #39855004
Yeah I know, hence its odd I found it kind of dumb for personal use. Moreso with the smaller models, which lost an objective benchmark I have to some Mistral finetunes. And I don't…
-
comment
Comment #39844574
I would note the actual leading models right now (IMO) are: - Miqu 70B (General Chat) - Deepseed 33B (Coding) - Yi 34B (for chat over 32K context) And of course, there are finetune…
-
comment
Comment #39829582
The conspiracy theorist in me says thats a low priority due to perverse incentives (namely selling more storage at a huge markup). Another rationale is that the Apple ecosystems te…
-
comment
Comment #39785642
Being a "hero" open source dev for a project like that can require a lot of neuroticism. Sometimes it works, but sometimes the project is just too big, I think.
-
comment
Comment #39771211
It's not either or, you can use different vendors for different tasks. tinygrad isn't in the realm of production ready though, AFAIK.
-
comment
Comment #39771190
The MI300 is the best accelerator you can buy, for many current workloads. It's technically way more advanced. Not as outrageously priced as an H100 either.
-
comment
Comment #39771125
I think you are preaching to the choir, and AMD is not listening. AMD would be selling 48GB 7900s or AI-only W7900s if they really wanted a consumer card ramp. They don't. Not beca…
-
comment
Comment #39771035
I never followed Hotz, so perhaps I missed something cool. But I never understood the hype myself.
-
comment
Comment #39771007
Well, personally, SDXL just blows 1.5 out of the water for me. I haven't had a reason to even touch 1.5 in months. But note that SDXL is really awful in automatic1111 or vanilla HF…
-
comment
Comment #39770963
Used 3090 prices are absolutely outrageous. And the 4090 MSRP was outrageous to begin with.
-
comment
Comment #39770558
I was talking about renting! There are some boutique hosts like Hot Aisle serving MI300s (who I really should reach out to), but for the immediate future our little startup is stuc…
-
comment
Comment #39770324
But is this going to blow over in a few days? Again? I can certainly appreciate frustration with the AMD stack, but be blunt, I was not impressed with Hotz's YouTube rant from befo…
-
comment
Comment #39770170
SDXL is amazing. The community is entrechend in 1.5 because that's what everyone is now familiar with, IMO
-
comment
Comment #39756241
Its more like the store being a literal hedge maze, with an entrance fee, and once you get to the actual products, they are outrageously priced junk. And the store is price fixing …
-
comment
Comment #39756021
My "oh no" moment was a vision model reading this perfectly : https://abadguide.files.wordpress.com/2012/01/jh66.jpg?w=640 Not an OCR program or anything specialized, just some gen…
-
comment
Comment #39739673
Tests are not out yet, but: - It's very large, yes. - It's a base model, so its not really practical to use without further finetuning. - Based on Grok-1 API performance (which its…
-
comment
Comment #39736280
Going to leave this gem here: https://www.vttoth.com/CMS/physics-notes/311-hawking-radiati... Black holes are weird because they are essentially macroscopic particles with only one…
-
comment
Comment #39735626
Groq's inference strategy appears to be "SRAM only." There is no external memory, like GGDR or HBM. Instead, large models are split between networked cards, and the inputs/outputs …
-
comment
Comment #39735223
> Proper moderation I would point to oldschool forums (and HN!), where the communities were just large enough to moderate themselves and stop nasty off topic junk like that from ap…