Viewing profile — aliljet
aliljet
HN member- Joined
- Mon, Feb 22, 2016, 6:56 AM UTC
- HN karma
- 1,407
- Public activity
- 258 items
- HN profile
- View on Hacker News ↗
About aliljet
Recent public activity
-
comment
Comment #49201063
Is there a path to distill this model to do very specific things? Like a RAG strategy for a small (or even large) corpus?
-
comment
Comment #49187553
There is a more serious question in here that's not being answered. How effective is the retrieval in finding buried needles in larger and larger haystacks. And there's a correlary…
-
comment
Comment #49151236
I'm trying and failing to find value running a potential Qwen 3.8 27b dense model on a 16 core, 128 GB of ram, 2080ti box. Yes, the GPU yells for help, but the problem is that no m…
-
comment
Comment #49079207
Honestly, I have a 2080ti that I use to play and I can tell you the math isn't there to upgrade it. It's much easier to just find a 3090/4090/5090 and keep pace with the software a…
-
comment
Comment #48854424
I'd be curious about an.option that would allow glm use with a low end GPU like a 2080 ti...
-
comment
Comment #48656399
The benchmarks here are confusing at best. Am I reading correctly that this model is essentially as good or better than all frontier models right now?
-
comment
Comment #48646583
I was just using infinity parser 2 (flash, to be fair) for pennies self-hosted to run through thousands of pages of documents with remarkable confidence. I decided to use https://h…
-
comment
Comment #48646269
I'm curious about this. What models/tools have you been using?
-
comment
Comment #48646125
How does this compare with infinty parser 2 which seemed to be running the table on every other OCR tool ( https://huggingface.co/datasets/allenai/olmOCR-bench ). To be fair, there…
-
comment
Comment #48561258
This sounds incredible. Have these models effectively solved the problem of trying to use a fast-processing network to predict the world's state ahead? For example, to catch a ball…
-
comment
Comment #48557062
The problem here is always the cost-benefit. For $200/mo, you're receiving subsidized best of breed access. There's no model competing for that price anywhere. If a 27B param model…
-
comment
Comment #48374314
Is this just one giant marketing plot?
-
comment
Comment #48210181
Where can a user reasonably host this in an affordable way to access the local LLM revolution?
-
comment
Comment #48197288
I'm really running into this deep at the edges of content creation. Take, for example, a need to general some kind of legal work. The cost of painstakingly checking and rechecking …
-
comment
Comment #48197168
Is there a good benchmark tracking hallucinations? The models are all incredibly good now, even the open ones, and my hope is that the rate of hallucinations is something that's fa…
-
comment
Comment #48094794
[dead]
-
comment
Comment #48017319
This is so cool. I would love to revitalize a generation of great, but perhaps boring older cars with FSD. Just so much work...
-
comment
Comment #48004629
[dead]
-
comment
Comment #47984224
Why did Spirit die? Was there any last of this that had to do with their abysmal customer service?
-
comment
Comment #47979031
What systems are you actively using? And what systems have you tried? It seems like law, generally, may be hitting a tipping point on LLM use...
-
comment
Comment #47957014
This is a tough moment. Claude is simultaneously becoming substantially more expensive, substantially less reliable (single 9 of reliability), and substantially less performant. It…
-
comment
Comment #47954050
I wonder how this kind of response from Anthropic is actually being read by the community at large. If you consider the rough sentiment of the r/ClaudeCode subreddit against the r/…
-
comment
Comment #47921499
Why is this being made public?
-
comment
Comment #47885572
How can you reasonably try to get near frontier (even at all tps) on hardware you own? Maybe under 5k in cost?
-
comment
Comment #47880762
Mythos is only real when it's actually available. If you're using Opus 4.7 right now, you know how incredibly nerfed the Opus autonomy is in service of perceived safety. I'm not so…