Live data from Hacker News

Viewing profile — mlb_hn

mlb_hn

HN member
Joined
Fri, Dec 29, 2017, 7:58 PM UTC
HN karma
726
Public activity
162 items

About mlb_hn

playing with little space aliens at julius.ai

Recent public activity

  1. comment
    Comment #44261004

    Our GCP is down

  2. comment
  3. comment
    Comment #40182302

    nice overview of progress over time. are there quant metrics for the sim capabilities or is it mostly vibes?

  4. comment
    Comment #39868691

    he was good people

  5. story
  6. comment
    Comment #30308662

    yeah, that's a very big caveat - haven't checked neo 20b yet. I've had a hard time getting the AI21 models to use it and those are also pretty big so it's interesting why sometimes…

  7. comment
    Comment #30306157

    definitely. also works on text translation/comprehension like emojis! https://aidungeon.medium.com/introducing-ai-dungeon-translat... . For actual benchmarks, scratchpad improves G…

  8. comment
    Comment #30305380

    I get the tokenization argument and it may influence it a bit, but I suspect the n-digit math issue has to do more with search the way it samples (in the bpe link gwern references …

  9. comment
    Comment #30301893

    Couple things there where you can see if it improves with the prompt/formatting. E.g. with Davinci (and J a bit but didn't test too much) you can get bette results by: - Using few-…

  10. comment
    Comment #27498740

    An alternative of that is as Oppenheimer once said, sometimes things are secret because a man doesn't like to know what he's up to if he can avoid it

  11. story
  12. comment
    Comment #24604290

    This has been a systemic issue reported on for years; e.g. reported by the Intercept in 2017 [1] and Atlantic in 2019 [2]. Not really made clear from the story considering the Econ…

  13. comment
    Comment #24311749

    I think that goes back to a Karpathy quote [1], don't know where he got it from [1] https://twitter.com/karpathy/status/961307010246492160

  14. comment
    Comment #24246035

    My thinking there wasn't because of BPEs, I think it's a graph traversal issue.

  15. comment
  16. comment
    Comment #24245712

    Also, better prompt design if you have make implicit meaning explicit can improve the WiC score ( http://gptprompts.wikidot.com/linguistics:word-in-context ) and ANLI score ( http:…

  17. comment
    Comment #24245637

    Just priming an immediate availability response is likely going to get poor results. On the other hand, this does bring up an important point, which is that few people have been sy…

  18. comment
  19. comment
    Comment #24103747

    Great article. It's worth considering the Japanese reaction to the bomb - people thought it was cluster munitions or other conventional explosives leading to them taking sub-optima…

  20. comment
    Comment #24012586

    I don't think it's going to be about a single prompt; reverse engineering multiple prompts interacting with themselves is hard. There's a lot of cool things to be done with: (a) cr…

  21. comment
  22. comment
  23. comment
    Comment #23434503

    I think it would be jumping to conclusions to say because the military has vulnerabilities that the entire military industrial complex is a fraud. Pieces of it certainly are sub-op…

  24. comment
    Comment #23127048

    That's the dataset used in the paper.

  25. comment
    Comment #22624961

    Leadership in the US Government wasted a lot of time trying to play down the pandemic [1] as did large (conservative) news organizations [2]. Both are now trying to blame China, ar…