Live data from Hacker News

Viewing profile — yolo-auto

yolo-auto

HN member
Joined
Tue, May 19, 2026, 2:15 AM UTC
HN karma
8
Public activity
12 items

About yolo-auto

No profile information was provided.

Recent public activity

  1. story
  2. story
  3. comment
    Comment #48818020

    lmao dont hurt yourself

  4. comment
    Comment #48817008

    Hi danger dingo, Sure, sign up and ping me in discord

  5. comment
    Comment #48817001

    Well, are you a robot? Ok well I need a way to securely give you an API key for you to try it out but I don't really want to post it on HN so I'm open to ideas lol

  6. comment
    Comment #48812895

    1x MI300x , rocm/sglang , highly optimized image, FP8, FP16 Cache (this fixes looping issues with qwen3.6-35b ~100tok/s average

  7. comment
    Comment #48812883

    actually the math is more like $1500 a month on a MI300x on a highly optimized rocm/sglang image and everyone's getting 100+ tok/sec and we have 100 active people and plenty of roo…

  8. comment
    Comment #48812865

    i know, but im scared to own our own auth, when there's money involved... too cheap to pay for a service. So we lean on the back of the giants for now. If it helps, there's really …

  9. story
    Show HN: An unmetered LLM API–$6/month, no token tracking, no limits

    Hi HN, I was once given the advice: Don't waste expensive frontier model credits (GPT/Claude/etc.) on bulk work. Send the boring, repetitive, high-volume jobs to a smaller model, a…

  10. comment
    Comment #48525804

    So on mi 300x we run FP8 and that was original plan. We are doing some weird q4 on 3090s now that is surprisingly good (for qwen) . will people pay for it? Yeah, sometimes. Some du…

  11. story
    Story of How Im Running an Unlimited $6/Month AI Provider on 4x RTX 3090s

    This submission is a tale about how I launched an unlimited LLM provider to about 60 hyped people on the waitlist, then immediately served them a fully dysfunctional death-loop mod…

  12. story