Viewing profile — benxh
benxh
HN member- Joined
- Tue, May 02, 2023, 2:02 PM UTC
- HN karma
- 46
- Public activity
- 28 items
- HN profile
- View on Hacker News ↗
About benxh
@nipple_nip on twitter
Recent public activity
-
comment
Comment #48570442
benchmark where gemini flash is better than fable btw.
-
comment
Comment #46624840
Crazy calling sovereign states "US Puppets".
-
comment
Comment #46548082
Minimax has been great for super high speed web/js/ts related work. It compares in my experience to Claude Sonnet, and at times gets stuff similar to Opus. Design wise it produces …
-
comment
Comment #46545912
It is arguable that the new Minimax M2.1 and GLM4.7 are drastically above Sonnet 3.7 in capabilities.
-
comment
Comment #45650461
cline is used by a lot of devs
-
comment
Comment #44241482
The longer "it" reasons, the more attention sinks are used to come to a "better" final output.
-
comment
Comment #42935055
To prove you right, you can read up on the incredible giga-brained countrywide experiments by Kardelj in Socialist Yugoslavia [0]. The result being a country where no-one wanted to…
-
comment
Comment #42887047
My biggest gripe with Ollama is the badly named models, e.g. under deepseek-r1, it defaults to the distill models.
-
comment
Comment #41720032
I'm pretty sure that Neosync[0] does this to a pretty good degree, it is open source and YC funded too. [0] https://www.neosync.dev/
-
comment
Comment #41015429
I am assuming this will be solved this year.
-
comment
Comment #39484032
If GPT4 is 220B/8 experts, that would be in-line with 3.5 Turbo being a 20B model, and GPT4 being a 55B activation out of a total 220B parameters. It is ultimately all speculation,…
-
comment
Comment #39042509
I personally was affected by this fire, although I've always kept 3 month backups of production data, encrypted, on-site, just in case of emergencies like this. Haven't touched the…
-
comment
Comment #38547706
It's buried deep in the Gemini report, but goddamn are these incredible stats.
-
comment
Comment #38310201
The Albanian takeover of AI continues. It's incredibly exciting!
-
comment
Comment #38089290
I can't wait to see this open sourced, there's a lot of sampling strategies that help coding. And I also can't wait to see how much Phind will improve further if the Glaive dataset…
-
comment
Comment #37983721
It was known by Polynesians for at least 1000 years before Columbus. See sweet potatoes.
-
comment
Comment #37843955
It's missing a lot of crucial details. Nothing on the dataset used, nothing on the data mix, nothing on their data cleaning procedures, nothing on the tokens trained.
-
comment
Comment #37813886
Not all of them per se, take a look at something like Mistral. It's a 7B model displaying incredible performance. IMO, we still haven't even scratched the surface of what is possib…
-
comment
Comment #37803605
Added, and reached out on Twitter.
-
comment
Comment #37782516
I would like to get in touch with you related to books4. Do you happen to have discord? or would twitter be ok? There's currently multiple attempts at creating what you describe as…
-
comment
Comment #37438479
I've had some success using vast.ai[0] with the Oobabooga LLM WebUI (LLaMA2) instances. One click to start up, minimal editing in the interface settings to enable OpenAI compatible…
-
comment
Comment #37360987
Yeah wildly inaccurate.
-
comment
Comment #36843563
This reads like a Serbian owned business from the North of Kosovo. But anybody following the local politics would know that: 1) Corruption as an issue is disappearing in Kosovo, es…
-
comment
Comment #36163076
Yes, but the acquisition of that data itself is illegal in almost all jurisdictions, since libgen is treated as a piracy website. Now if there were a pipeline to access books from …
-
comment
Comment #36152367
How does one go about becoming a distributor of Quest products in countries which arent served at all by Meta? Considering I can leverage existing infrastructure and network connec…