Live data from Hacker News

Viewing profile — aliljet

aliljet

HN member
Joined
Mon, Feb 22, 2016, 6:56 AM UTC
HN karma
1,407
Public activity
258 items

About aliljet

contact me here: pav.gup@gmail.com

Recent public activity

  1. comment
    Comment #49201063

    Is there a path to distill this model to do very specific things? Like a RAG strategy for a small (or even large) corpus?

  2. comment
    Comment #49187553

    There is a more serious question in here that's not being answered. How effective is the retrieval in finding buried needles in larger and larger haystacks. And there's a correlary…

  3. comment
    Comment #49151236

    I'm trying and failing to find value running a potential Qwen 3.8 27b dense model on a 16 core, 128 GB of ram, 2080ti box. Yes, the GPU yells for help, but the problem is that no m…

  4. comment
    Comment #49079207

    Honestly, I have a 2080ti that I use to play and I can tell you the math isn't there to upgrade it. It's much easier to just find a 3090/4090/5090 and keep pace with the software a…

  5. comment
    Comment #48854424

    I'd be curious about an.option that would allow glm use with a low end GPU like a 2080 ti...

  6. comment
    Comment #48656399

    The benchmarks here are confusing at best. Am I reading correctly that this model is essentially as good or better than all frontier models right now?

  7. comment
    Comment #48646583

    I was just using infinity parser 2 (flash, to be fair) for pennies self-hosted to run through thousands of pages of documents with remarkable confidence. I decided to use https://h…

  8. comment
    Comment #48646269

    I'm curious about this. What models/tools have you been using?

  9. comment
    Comment #48646125

    How does this compare with infinty parser 2 which seemed to be running the table on every other OCR tool ( https://huggingface.co/datasets/allenai/olmOCR-bench ). To be fair, there…

  10. comment
    Comment #48561258

    This sounds incredible. Have these models effectively solved the problem of trying to use a fast-processing network to predict the world's state ahead? For example, to catch a ball…

  11. comment
    Comment #48557062

    The problem here is always the cost-benefit. For $200/mo, you're receiving subsidized best of breed access. There's no model competing for that price anywhere. If a 27B param model…

  12. comment
    Comment #48374314

    Is this just one giant marketing plot?

  13. comment
    Comment #48210181

    Where can a user reasonably host this in an affordable way to access the local LLM revolution?

  14. comment
    Comment #48197288

    I'm really running into this deep at the edges of content creation. Take, for example, a need to general some kind of legal work. The cost of painstakingly checking and rechecking …

  15. comment
    Comment #48197168

    Is there a good benchmark tracking hallucinations? The models are all incredibly good now, even the open ones, and my hope is that the rate of hallucinations is something that's fa…

  16. comment
  17. comment
    Comment #48017319

    This is so cool. I would love to revitalize a generation of great, but perhaps boring older cars with FSD. Just so much work...

  18. comment
  19. comment
    Comment #47984224

    Why did Spirit die? Was there any last of this that had to do with their abysmal customer service?

  20. comment
    Comment #47979031

    What systems are you actively using? And what systems have you tried? It seems like law, generally, may be hitting a tipping point on LLM use...

  21. comment
    Comment #47957014

    This is a tough moment. Claude is simultaneously becoming substantially more expensive, substantially less reliable (single 9 of reliability), and substantially less performant. It…

  22. comment
    Comment #47954050

    I wonder how this kind of response from Anthropic is actually being read by the community at large. If you consider the rough sentiment of the r/ClaudeCode subreddit against the r/…

  23. comment
    Comment #47921499

    Why is this being made public?

  24. comment
    Comment #47885572

    How can you reasonably try to get near frontier (even at all tps) on hardware you own? Maybe under 5k in cost?

  25. comment
    Comment #47880762

    Mythos is only real when it's actually available. If you're using Opus 4.7 right now, you know how incredibly nerfed the Opus autonomy is in service of perceived safety. I'm not so…