Live data from Hacker News

Viewing profile — lewtun

lewtun

HN member
Joined
Tue, Apr 17, 2018, 3:34 PM UTC
HN karma
22
Public activity
31 items

About lewtun

No profile information was provided.

Recent public activity

  1. comment
    Comment #48051191

    Shameless plug: https://huggingface.co/spaces/smolagents/ml-intern It’s a simple harness around Opus, but with tight integration to Hugging Face infra, so the agent can read papers…

  2. comment
    Comment #47758377

    Hugging Face Buckets are pretty simple: https://huggingface.co/docs/huggingface_hub/en/guides/bucket... Disclaimer: I work at HF

  3. comment
    Comment #45797861

    The analogy stems from the notion that neural nets are "grown" rather than "engineered". Chris Olah has an old, but good post with some specific examples: https://colah.github.io/n…

  4. comment
    Comment #45790288

    Thanks! I expect the book will remain relevant as long as the Transformers architecture does. That’s why we mostly focus on topics we think will stand the test of time, but let’s s…

  5. comment
    Comment #45788176

    In the specific case of SmolLM, it originates from the meme in this dataset https://huggingface.co/datasets/bigcode/the-stack-smol

  6. comment
    Comment #45785734

    Hi, Lewis here (one of the co-authors). Happy to answer any questions people have about the book :)

  7. story
  8. story
  9. comment
    Comment #45476663

    For those interested in playing with an implementation of these ideas, my colleagues at HF made some recipes here: https://github.com/huggingface/trl/blob/main/docs/source/lor...

  10. comment
    Comment #45147923

    “QED and the Men Who Made It” [1] might be close to what you’re after for quantum theory at least. Unlike other popular accounts, it gets quite technical and covers a lot of the hi…

  11. comment
    Comment #45096080

    > We instantiate this idea through Preference-prior Informed Linucb fOr adaptive rouTing (PILOT), a novel extension of LinUCB Academics are pretty creative at naming their creation…

  12. comment
    Comment #44502761

    Indeed we opted for offline methods like Anchored Preference Optimization as we found in the Open R1 project that doing multi-task RL on small models is quite a hassle to get right…

  13. comment
    Comment #43982288

    > The absolute best way of doing this is these days is likely through a vision based machine learning model, but that is an approach that is very far away from scaling to processin…

  14. story
  15. story
  16. story
  17. comment
    Comment #41717642

    I gave the demo a spin and it’s pretty nice! One thing I noticed is that the avatar doesn’t seem to be aware of it’s surroundings- for example, I asked it why it was wearing a cowb…

  18. comment
    Comment #41190984

    > I expect language models to also get crazy good at mathematical theorem proving Indeed, systems like AlphaProof / AlphaGeometry are already able to win a silver medal at the IMO,…

  19. story
  20. story
  21. story
  22. comment
    Comment #40007770

    Hello everyone, we just did a speed run with Argilla and KAIST AI to fine-tune the beefy new Mixtral model with some new techniques that came out recently. More details in the mode…

  23. story
  24. story
  25. story