Live data from Hacker News

Viewing profile — zingelshuher

zingelshuher

HN member
Joined
Tue, Feb 06, 2024, 11:25 PM UTC
HN karma
1
Public activity
49 items

About zingelshuher

No profile information was provided.

Recent public activity

  1. comment
    Comment #40413234

    [flagged]

  2. comment
  3. comment
    Comment #40413141

    > Google knowingly made their search results shittier and shittier for years unfortunately this extends to youtube too. now they have a new shitty trick. you click on the link and …

  4. comment
    Comment #40413110

    > It is unlikely people are going to switch en mass to open source models It depends on the task at hands. For complex tasks no way personal computer can compete with giants data c…

  5. comment
    Comment #40194357

    > poorly designed government intervention due to misunderstanding of the dynamics behind the process (homelessness) Major drive is easy to understand, just cross the border and you…

  6. comment
    Comment #40099641

    had to upvote this

  7. comment
    Comment #40099606

    Only if it does nothing. In fact Google is one of the major players in LLM field. The winner is hard to predict, chip makers likely ;) Everybody jumped on bandwagon, Amazon is jump…

  8. comment
    Comment #40099550

    I often use ChatGPT4 for technical info. It's easier then scrolling through pages whet it works. But.. the accuracy is inconsistent, to put it mildly. Sometimes it gets stuck on wr…

  9. comment
    Comment #40099371

    It's impossible. Meta itself cannot reproduce the model. Because training is randomized and that info is lost. First samples a coming at random. Second there are often drop-out lay…

  10. comment
    Comment #40059350

    If we can keep unlimited memory, but use only a selected relevant subset in each chat session. This should help. Of course the key is 'selected', it's another big problem. Like sho…

  11. comment
    Comment #40059213

    It's a different animal. In general you cannot reproduce the model even having all the training data. There are too many random factors and nobody keeps track of them. Just pushing…

  12. comment
    Comment #40034565

    It's a matter of opinion how much open model should be to be called 'open source'. Looks like some believe they have the right to define it for everybody else to use. Like for soft…

  13. comment
    Comment #40034428

    > I'll bet Adobe or similiar will buy this 0,5 mil May be that was the business plan, or 'plan B'

  14. comment
    Comment #40034374

    Yes. And there are many forgotten accounts on youtube, their owners now have another way to monetarize. Not sure why Adobe didn't talk to Youtube directly.

  15. comment
    Comment #40024499

    As you mentioned models need _random_ videos. Just 'walk outside' will produce more or less the same. My guess Adobe is more interested in 'family' sort of videos, with humans. Thi…

  16. comment
    Comment #40024450

    It doesn't affect creators life as they guarantee (likely) it will not be made public. So it's just side income. Likely creators make a lot of fragments which don't get into the fi…

  17. comment
    Comment #40009357

    It reflects the fact that Amazon in serious about AI. If fact they are well positioned with their datacenters and a lot of ways to apply, starting with smarter Alexa.

  18. comment
    Comment #40009343

    Expect sh*t load of AI hallucinations. As if Wiki isn't bad enough with BS some intentionally posting.

  19. comment
    Comment #39966478

    "the bigger you make epsilon "... " thus slower the training progress will be" Sounds like variable epsilon is optimal, that's instead of learning rate, or both together. Would be …

  20. comment
    Comment #39957659

    Intuitively looks like models should be close enough, or sparse enough for merge to work. I wonder if MoE experts can be merged(?)

  21. comment
    Comment #39957590

    Good luck with that ;) Actually you _can_ learn juggling this way, just couple of minutes at a time.

  22. comment
    Comment #39957569

    There are unstable cases when static learning rate doesn't work. Solution starts wobbling too much after some time and explodes. Using too small LR from the beginning leads to loca…

  23. comment
    Comment #39937546

    Have to disagree with this. The major limiting factor is the software and lack of applications. If there was a killer application it would be much easier to sell. Then the numbers …

  24. comment
    Comment #39874197

    Why isn't he identified personally? Very likely he is 'contributing' to other projects under different accounts.

  25. comment
    Comment #39795397

    Question, have you seen the improvement after adding the noise? I mean in practice. Asking because intuition sometimes doesn't work.