Live data from Hacker News

Viewing profile — peaslock

peaslock

HN member
Joined
Tue, Jun 21, 2022, 8:16 AM UTC
HN karma
37
Public activity
31 items

About peaslock

No profile information was provided.

Recent public activity

  1. comment
    Comment #34727609

    Maybe not a good idea to link a page that runs source code created by random people. Well CSS is very safe, but still.

  2. comment
    Comment #34625947

    Not if you ask first.

  3. comment
    Comment #34536351

    Neural nets often fail with (repetitive) gibberish output when the input is too different from the training data. This model appears to take in the entire text input at once or loo…

  4. comment
    Comment #34529707

    Can DRIZZLE help to achieve higher resolution? Though with hundreds of photos this will imply a lot of work: https://en.wikipedia.org/wiki/Drizzle_(image_processing)

  5. comment
  6. comment
    Comment #34153494

    Geoffrey Hinton has recently been talking about how analog and "imperfect" computing with specialized hardware/circuitry may yield much cheaper neural nets, that could easily be as…

  7. comment
    Comment #34104348

    But they have highly likely internal prototypes with higher bandwidth and latency. Also, with distilled latent diffusion one can probably generate text(-images) much faster anyhow …

  8. comment
    Comment #34095627

    Yes, but the losses in Figure 3 increase because the larger models see fewer data to keep the FLOP budget constant, not because of overfitting. Large models do not overfit very muc…

  9. comment
    Comment #34051435

    Though isn't it highly likely that core devs working at the big tech giants have access to 10x-100x faster compute, e.g. some secret TPU successor at Google?

  10. comment
    Comment #34049644

    > if you want improved performance, you still need more data Not true. See figure 2: https://arxiv.org/pdf/2203.15556.pdf#page=5 The loss decreases with greater model size at the s…

  11. comment
    Comment #34049152

    Not necessarily: https://arxiv.org/abs/2206.14486 Also, even with "Chinchilla laws", you still gain performance in a larger model, you just need a lot more data (if just as noisy) …

  12. comment
    Comment #34042539

    Yeah, continuous online learning by fine-tuning seems like an obvious way of making these models recall information from outside the perceptible context. One could also prompt the …

  13. comment
    Comment #34041831

    The model with the most similar name in this list is code-cushman-001 which is described as "Codex model that is a stronger, multilingual version of the Codex (12B) model in the pa…

  14. comment
    Comment #34041551

    Amazing if this is only a 12B model. If this already increases coding productivity by up to 50% (depending on kind of work), imagine what a 1T model will be capable of! I do wonder…

  15. comment
    Comment #33972862

    Can you use a different Wi-Fi network at the same time for internet access or does Wi-Fi Direct block any other Wi-Fi access?

  16. comment
    Comment #33942165

    > Plus you don't need to be on the same network You mean in case of the Wireless Display Adapter? For Miracast you do need to be on the same LAN, right?

  17. comment
    Comment #33941830

    Google will not disappear. They already have much larger neural nets almost ready for deployment, plus they will be able to afford even larger ones in future. And size is all that …

  18. comment
    Comment #33636666

    How will we justify our existence unable to contribute meaningfully to the economy?

  19. comment
    Comment #33309959

    My Firefox is now stuck with a note at the top of the screen saying something like "voxelchain.app is controlling your mouse cursor. Please press ESC to take over."

  20. comment
    Comment #33240143

    It is measurable, but not harmful to a meaningful extent. There are lots of sources of low-dose radiation in the natural human/primate/.../mammalian environment.

  21. comment
    Comment #32723142

    Speaking of which, is anyone aware of example code using the LSTM? I've been trying to get this to work, but there seems to be information missing e.g how to setup the input/output…

  22. comment
  23. comment
    Comment #32609410

    How long does the battery last for you browsing like that?

  24. comment
    Comment #32605370

    So 0.005% at 560 TWh total consumption.

  25. comment
    Comment #32605089

    > About third of what was used before. How much is that as a fraction of the total electricity consumption?