Viewing profile — hislaziness
hislaziness
HN member- Joined
- Mon, Jun 29, 2009, 7:40 AM UTC
- HN karma
- 478
- Public activity
- 121 items
- HN profile
- View on Hacker News ↗
About hislaziness
No profile information was provided.
Recent public activity
-
comment
Comment #48105407
Am I missing something here. Not finding a major bug/vulnerability just means that maybe the code is really good, not that the model is not what is claimed?
- story
- story
-
comment
Comment #44871577
Would it be more appropriate to compare LLMs to Autotunes rather than pianos?
-
comment
Comment #44641750
Akamai CTO Robert Blumofe offers four tips for business leaders striving to foster AI fluency by empowering employees with the right tools and best use cases. useful insights but d…
-
comment
Comment #44621230
Great. I enjoy these kind of articles. My all time favorite book for 'C' is Expert C Programming: Deep C Secrets.
-
comment
Comment #44621180
Terence Tao on the matter - https://imgur.com/a/terence-tao-on-supposed-gold-imo-sMKP0bm
-
comment
Comment #41921983
As I understand, the LLM uses the techniques of searchformer - https://arxiv.org/abs/2402.14083 . To do "slow thinking" doing a A* search using a transofrmer.
- story
-
comment
Comment #41805991
It is not just MSRP, management and operations cost too. The article goes into the details of this.
-
comment
Comment #41805949
The details are in the article. They have done the math.
-
comment
Comment #41805550
TLDR: Don’t buy H100s. The market has flipped from shortage ($8/hr) to oversupplied ($2/hr), because of reserved compute resales, open model finetuning, and decline in new foundati…
-
comment
Comment #41764325
The screen seems to be stuck at Wait a second... for me.
-
comment
Comment #41470901
I tried a few local LLMs. None of them could give me the right answer for "How many 'r's in straberry. All LLMs were 8-27B.
-
comment
Comment #41343607
I know you mean this in jest, but we are much closer to this than we would imagine, the use of LLMs to process communication / translation is becoming ubiquitous. We are 1 bad tran…
-
comment
Comment #41253755
I also use some email providers ability to have +xyz at the end of the username. So for a registration I would for user.name+sitea@domain.com. has helped me track spam and leaks in…
-
comment
Comment #41053839
Cool. I will try it out. I tried the same with ollama, the non english part needs a lot more polish. Do you see the outcome being any different?
-
comment
Comment #41053081
This is pretty awesome. How did you do it? Any blog on the detials?
-
comment
Comment #41000896
The model description on huggingface says - Model size - 12.2B params, Tensor type - BF16. Is the Tensor type different from the training param size?
-
comment
Comment #41000885
I just checked huggingface and the model files download is about 25GB but in a comment below someone mentioned it is 8fp quantized model. Trying to understand how the quantization …
-
comment
Comment #40997003
isn't it 2 bytes (fp16) per param. so 7b = 14 GB+some for inference?
-
comment
Comment #40968186
The title does not do justice to the article. It talks about a new classification system that OpenAI has introduced for LLMs
- story
-
comment
Comment #40799403
[flagged]
-
comment
Comment #40605266
Same here. Coincidentally I joined in June 2009 and this is the only social media I indulge in.