Live data from Hacker News

Viewing profile — brucethemoose2

brucethemoose2

HN member
Joined
Mon, Feb 13, 2023, 5:50 AM UTC
HN karma
7,874
Public activity
3,307 items

About brucethemoose2

No profile information was provided.

Recent public activity

  1. comment
  2. comment
    Comment #39876619

    Some phones have limiters to keep the battery at 60%-80%. I believe most can do this with the right software. It's not a fix, but it should extend the life considerably.

  3. comment
    Comment #39875777

    Another benefit: modern smartphones have large GPUs, large media blocks, and fast RAM. With the right software , they can be a surprisingly powerful AI host or transcoding server.

  4. comment
    Comment #39870001

    Real world GPU performance is hugely influenced by hand optimization of the CUDA kernels.

  5. comment
    Comment #39869240

    Power/Weight is extremely high. A tiny wankel will do the job, and weight is everything on cars. It does prefer a narrow RPM band, which is fine. Reliability is the biggest concern…

  6. comment
    Comment #39857290

    Yeah, its an unspoken but rampant thing in the llm community. Basically no one respects licenses for training data. I'd say the majority of instruct tunes, for instance, use OpenAI…

  7. comment
    Comment #39855004

    Yeah I know, hence its odd I found it kind of dumb for personal use. Moreso with the smaller models, which lost an objective benchmark I have to some Mistral finetunes. And I don't…

  8. comment
    Comment #39844574

    I would note the actual leading models right now (IMO) are: - Miqu 70B (General Chat) - Deepseed 33B (Coding) - Yi 34B (for chat over 32K context) And of course, there are finetune…

  9. comment
    Comment #39829582

    The conspiracy theorist in me says thats a low priority due to perverse incentives (namely selling more storage at a huge markup). Another rationale is that the Apple ecosystems te…

  10. comment
    Comment #39785642

    Being a "hero" open source dev for a project like that can require a lot of neuroticism. Sometimes it works, but sometimes the project is just too big, I think.

  11. comment
    Comment #39771211

    It's not either or, you can use different vendors for different tasks. tinygrad isn't in the realm of production ready though, AFAIK.

  12. comment
    Comment #39771190

    The MI300 is the best accelerator you can buy, for many current workloads. It's technically way more advanced. Not as outrageously priced as an H100 either.

  13. comment
    Comment #39771125

    I think you are preaching to the choir, and AMD is not listening. AMD would be selling 48GB 7900s or AI-only W7900s if they really wanted a consumer card ramp. They don't. Not beca…

  14. comment
    Comment #39771035

    I never followed Hotz, so perhaps I missed something cool. But I never understood the hype myself.

  15. comment
    Comment #39771007

    Well, personally, SDXL just blows 1.5 out of the water for me. I haven't had a reason to even touch 1.5 in months. But note that SDXL is really awful in automatic1111 or vanilla HF…

  16. comment
    Comment #39770963

    Used 3090 prices are absolutely outrageous. And the 4090 MSRP was outrageous to begin with.

  17. comment
    Comment #39770558

    I was talking about renting! There are some boutique hosts like Hot Aisle serving MI300s (who I really should reach out to), but for the immediate future our little startup is stuc…

  18. comment
    Comment #39770324

    But is this going to blow over in a few days? Again? I can certainly appreciate frustration with the AMD stack, but be blunt, I was not impressed with Hotz's YouTube rant from befo…

  19. comment
    Comment #39770170

    SDXL is amazing. The community is entrechend in 1.5 because that's what everyone is now familiar with, IMO

  20. comment
    Comment #39756241

    Its more like the store being a literal hedge maze, with an entrance fee, and once you get to the actual products, they are outrageously priced junk. And the store is price fixing …

  21. comment
    Comment #39756021

    My "oh no" moment was a vision model reading this perfectly : https://abadguide.files.wordpress.com/2012/01/jh66.jpg?w=640 Not an OCR program or anything specialized, just some gen…

  22. comment
    Comment #39739673

    Tests are not out yet, but: - It's very large, yes. - It's a base model, so its not really practical to use without further finetuning. - Based on Grok-1 API performance (which its…

  23. comment
    Comment #39736280

    Going to leave this gem here: https://www.vttoth.com/CMS/physics-notes/311-hawking-radiati... Black holes are weird because they are essentially macroscopic particles with only one…

  24. comment
    Comment #39735626

    Groq's inference strategy appears to be "SRAM only." There is no external memory, like GGDR or HBM. Instead, large models are split between networked cards, and the inputs/outputs …

  25. comment
    Comment #39735223

    > Proper moderation I would point to oldschool forums (and HN!), where the communities were just large enough to moderate themselves and stop nasty off topic junk like that from ap…