Live data from Hacker News

Viewing profile — gliched_robot

gliched_robot

HN member
Joined
Sun, Feb 18, 2024, 9:24 PM UTC
HN karma
36
Public activity
15 items

About gliched_robot

No profile information was provided.

Recent public activity

  1. comment
    Comment #40371315

    If anyone wants to try this out, here is the hugging-face spaces demo: https://huggingface.co/spaces/google/paligemma

  2. comment
    Comment #40359808

    This is far more superior than SORA, there is no comparison.

  3. comment
    Comment #40205988

    Lmsys devs have all the answers, I am not sure how this has not leaked yet. They must a strong NDAs.

  4. comment
    Comment #40199781

    I do not understand the taught process here. They are regulating it so fast. It's almost like regulating car before even engine is invented.

  5. comment
    Comment #40083582

    GPU server locations, maybe?

  6. comment
    Comment #40083578

    Inference speed is not a great metric given the horizontal scalability of LLMs.

  7. comment
    Comment #40083570

    Disagree on Nvidia, most folks fine-tune model. Proof: there are about 20k models in huggingface derived from llama 2, all of them trained on Nvidia GPUs.

  8. comment
    Comment #40083556

    Maybe a typo?

  9. comment
    Comment #40083540

    This llama model some made it run on an iphone. https://x.com/1littlecoder/status/1781076849335861637?s=46

  10. comment
    Comment #40080313

    I see what you did here carrying the "torch" . LOL

  11. comment
    Comment #40078232

    The code it writes is getting worse eg. lazy and not updating the function, not following prompts etc. So we can objectively say its getting worse.

  12. comment
    Comment #40078206

    Wild considering, GPT-4 is 1.8T.

  13. comment
    Comment #40077871

    If any one is interesting in seeing how 400B model compares with other opensource models, here is a useful chart: https://x.com/natolambert/status/1780993655274414123

  14. comment
    Comment #40076985

    Is this real? Seems suspicious given that its just one model not a family of models like llama and llama-2.

  15. comment
    Comment #39423534

    This is very cool and will change the way we do lora now.