Live data from Hacker News

Viewing profile — bguberfain

bguberfain

HN member
Joined
Mon, Nov 20, 2017, 9:24 PM UTC
HN karma
138
Public activity
52 items

About bguberfain

meet.hn/city/br-Rio-de-Janeiro Data Scientist @ Petrobras

Recent public activity

  1. comment
    Comment #48374939

    It is good to se big companies like Microsoft launching LLMs. They have large amount of compute power and good scientists to create useful models.

  2. comment
    Comment #48213326

    Any plans to port to sglang or vLLM?

  3. comment
    Comment #48151901

    It seems to be something related the moving average calculation. So it is just a glitch on the chart.

  4. comment
    Comment #47892694

    This guy seems to be talking seriously.

  5. comment
    Comment #47768482

    Not to demerit the recording, but I felt more nostalgic for the last sentence of the article "Sometimes, the internet is good" than for the musics itself.

  6. comment
    Comment #47693379

    We all know it... but I think they were very bold in this warning about using your private messages to train public models. _Your messages with AIs will be used to improve AI at Me…

  7. comment
    Comment #47439231

    "A watchdog kernel thread monitors RAM and NVMe pressure and signals userspace before things get dangerous." - which kind of danger this type of solution can have?

  8. comment
    Comment #46346194

    We can finally search for playlists with a giving song! A basic feature that Spotify is missing!

  9. story
    Ask HN: Is it possible to implement this button on a browser?

    I saw this video [0] of a virtual button, created with Houdini, and was wondering if it is possible to implement this behavior on a browser. Any suggestions? [0] https://www.reddit…

  10. comment
    Comment #44159077

    So they used a LLM with knowledge cut in mid 2023 to evaluate 2023? Seems like a classic leakage problem. From paper: "testing set: January 1, 2023, to December 31, 2023" From the …

  11. comment
    Comment #44056109

    I think that there may be another solution for this, that is the LLM write a valid code that calls the MCP's as functions. See it like a Python script, where each MCP is mapped to …

  12. story
  13. comment
    Comment #44032128

    https://blogs.windows.com/windowsdeveloper/2025/05/19/the-wi...

  14. comment
    Comment #44004524

    Not available in my country :(

  15. comment
    Comment #43680860

    Unfortunately, it uses Miniconda, which does not allow usage in companies with more than 200 employees. I think it conflicts with AGPL license. I created a PR to fix that.

  16. story
  17. comment
    Comment #43345092

    Can you provide more information about this “bigger teacher” model?

  18. comment
    Comment #43198203

    Until GPT-4.5, GPT-4 32K was certainly the most heavy model available at OpenAI. I can imagine the dilemma between to keep it running or stop it to free GPU for training new models…

  19. comment
    Comment #43017563

    Any chance you could release the dataset to the public? I imagine NewsCatcher and Polymarket might not agree..

  20. comment
    Comment #42952741

    It remembers me Theano [0]. [0] https://en.wikipedia.org/wiki/Theano_(software)

  21. comment
    Comment #42658818

    I agree. One image of what it is doing would improve the comprehension of the algorithm.

  22. comment
    Comment #42268025

    The file I mentioned is just the begining... there is a folder full of .dll files, renamed to .pyd. I understand that this is the proprietary part, that limits usage for 30 minutes…

  23. comment
    Comment #42267139

    Thanks for sharing this! But I have some doubts about hidden installation procedures. It imports all functions from one_click (from one_click import *), which points to a compiled …

  24. story
  25. comment
    Comment #40683737

    "Nemotron-4-340B-Instruct is a chat model intended for use for the English language" - frustrating