Viewing profile — bguberfain
bguberfain
HN member- Joined
- Mon, Nov 20, 2017, 9:24 PM UTC
- HN karma
- 138
- Public activity
- 52 items
- HN profile
- View on Hacker News ↗
About bguberfain
Recent public activity
-
comment
Comment #48374939
It is good to se big companies like Microsoft launching LLMs. They have large amount of compute power and good scientists to create useful models.
-
comment
Comment #48213326
Any plans to port to sglang or vLLM?
-
comment
Comment #48151901
It seems to be something related the moving average calculation. So it is just a glitch on the chart.
-
comment
Comment #47892694
This guy seems to be talking seriously.
-
comment
Comment #47768482
Not to demerit the recording, but I felt more nostalgic for the last sentence of the article "Sometimes, the internet is good" than for the musics itself.
-
comment
Comment #47693379
We all know it... but I think they were very bold in this warning about using your private messages to train public models. _Your messages with AIs will be used to improve AI at Me…
-
comment
Comment #47439231
"A watchdog kernel thread monitors RAM and NVMe pressure and signals userspace before things get dangerous." - which kind of danger this type of solution can have?
-
comment
Comment #46346194
We can finally search for playlists with a giving song! A basic feature that Spotify is missing!
-
story
Ask HN: Is it possible to implement this button on a browser?
I saw this video [0] of a virtual button, created with Houdini, and was wondering if it is possible to implement this behavior on a browser. Any suggestions? [0] https://www.reddit…
-
comment
Comment #44159077
So they used a LLM with knowledge cut in mid 2023 to evaluate 2023? Seems like a classic leakage problem. From paper: "testing set: January 1, 2023, to December 31, 2023" From the …
-
comment
Comment #44056109
I think that there may be another solution for this, that is the LLM write a valid code that calls the MCP's as functions. See it like a Python script, where each MCP is mapped to …
- story
-
comment
Comment #44032128
https://blogs.windows.com/windowsdeveloper/2025/05/19/the-wi...
-
comment
Comment #44004524
Not available in my country :(
-
comment
Comment #43680860
Unfortunately, it uses Miniconda, which does not allow usage in companies with more than 200 employees. I think it conflicts with AGPL license. I created a PR to fix that.
- story
-
comment
Comment #43345092
Can you provide more information about this “bigger teacher” model?
-
comment
Comment #43198203
Until GPT-4.5, GPT-4 32K was certainly the most heavy model available at OpenAI. I can imagine the dilemma between to keep it running or stop it to free GPU for training new models…
-
comment
Comment #43017563
Any chance you could release the dataset to the public? I imagine NewsCatcher and Polymarket might not agree..
-
comment
Comment #42952741
It remembers me Theano [0]. [0] https://en.wikipedia.org/wiki/Theano_(software)
-
comment
Comment #42658818
I agree. One image of what it is doing would improve the comprehension of the algorithm.
-
comment
Comment #42268025
The file I mentioned is just the begining... there is a folder full of .dll files, renamed to .pyd. I understand that this is the proprietary part, that limits usage for 30 minutes…
-
comment
Comment #42267139
Thanks for sharing this! But I have some doubts about hidden installation procedures. It imports all functions from one_click (from one_click import *), which points to a compiled …
- story
-
comment
Comment #40683737
"Nemotron-4-340B-Instruct is a chat model intended for use for the English language" - frustrating