Viewing profile — XCSme
XCSme
HN member- Joined
- Sun, Dec 28, 2014, 6:47 PM UTC
- HN karma
- 2,940
- Public activity
- 3,022 items
- HN profile
- View on Hacker News ↗
About XCSme
Self-Hosted Analytics with Heatmaps and Session Recordings: https://www.uxwizz.com
WordPress Analytics: https://www.wplytic.com
X (Twitter): @XCSme
Recent public activity
-
comment
Comment #49217887
As a solo dev, this gives me hope. I feel like I have an advantage over big companies, if I can use the best models on a subscription and not worry about costs much, when they can'…
-
comment
Comment #49208480
My question was more about more complex problems, which no seem to be multi-turn somehow, or maybe just the harnesses make it look that way. I am curious what the drop in thoughput…
-
comment
Comment #49208440
Yeah, makes sense, if it's good for very small models only, then there's no point, as those van already run on cheap consumer hardware. Yet, maybe it can work well enough, so that …
-
comment
Comment #49205779
I asked a LLM after posting my comment, to see if I had a genius idea or not,just for it to tell me the same as you, that's now they work already...
-
comment
Comment #49205731
My concern is that reasoning could involve some sequential steps that instant models don't. Not sure if modern models "think" only by outputting blocks, or there is a more complex …
-
comment
Comment #49204573
It was just a random example, you could think of it as being a lot more complex (detect which type of food it is, what detergent to use, how much water, remember patterns, learn ov…
-
comment
Comment #49204472
So local personalized ads? Not sure if that's better or worse than online personalizaed ads...
-
comment
Comment #49203940
Why not have some a device/hardware that programs itself on-boot. Sort of a FPGA, that (electrically) arranges the connections on-boot, and then it's like a static inference chip.
-
comment
Comment #49203917
I think this would make sense for consumer hardware, not for AI companies. AI companies constantly update/change stuff, new models come out, new requirements, etc. But if you ship …
-
comment
Comment #49203902
Wait, is it even thinking? Or is it an instant model?
-
comment
Comment #49203887
Wow, that's instant, crazy.
-
comment
Comment #49154084
I am surprised that they keep going with it, seeing how fast it improves and basically soon running themselves too out of business. What's even their end goal? Open source models m…
-
comment
Comment #49151019
If they trained it well, and can do computer use, it will be a new era. Companies can keep PCs, put Qwen 3.8 27b on it and get rid of the employees, lol...
-
comment
Comment #49121972
Can't really use it now, without giving away your data: > Trains: this provider may use prompts for training and may retain prompt data.
-
comment
Comment #49121946
Or maybe not: https://news.ycombinator.com/item?id=49119559
-
comment
Comment #49119888
I think that with this change, DeepSeek v4 Flash has finally been dethroned.
-
comment
Comment #49119369
Wasn't one of the main original points of LLMs to be creative? To create stories, creative writing?
-
comment
Comment #49080820
Is this like Runescape free armour trimming? You send your gpu, and get it back with 2x memory?
-
comment
Comment #49069986
Depends on whether the models report the correct amount of tokens. 5.5 Sol repors 10x fewer reasoning tokens than Kimi k3. If it is correct, than it unlikely has those doubt issues…
-
comment
Comment #49069194
Yeah, they thing forever and doubt everything "wait but" for 200k tokens for almost any question.
-
comment
Comment #49069187
So nowadays the hardware and hosting providers must be in an optimization race, whoever can make the model just a bit smaller or more efficient (to fit on fewer/less powerful cards…
-
comment
Comment #49044428
One of the best hamsters [0]. Again, their "none" version costs more than "low", and says zero reasoning tokens, makes no sense[1]. As always, the "low" version seems to be the bes…
-
comment
Comment #49043813
Twice the cost for 4% more intelligence, is it worth it?
-
comment
Comment #49031287
I have a spare 3090 that I want to use to off-load some tasks from Claude to a local model (probably Qwen 3.6 27b), any success with that? Is it good enough to follow some tasks, c…
-
comment
Comment #49031282
I thought login-protected apps are not allowed on Show HN.