Viewing profile — mlb_hn
mlb_hn
HN member- Joined
- Fri, Dec 29, 2017, 7:58 PM UTC
- HN karma
- 726
- Public activity
- 162 items
- HN profile
- View on Hacker News ↗
About mlb_hn
Recent public activity
-
comment
Comment #44261004
Our GCP is down
- comment
-
comment
Comment #40182302
nice overview of progress over time. are there quant metrics for the sim capabilities or is it mostly vibes?
-
comment
Comment #39868691
he was good people
- story
-
comment
Comment #30308662
yeah, that's a very big caveat - haven't checked neo 20b yet. I've had a hard time getting the AI21 models to use it and those are also pretty big so it's interesting why sometimes…
-
comment
Comment #30306157
definitely. also works on text translation/comprehension like emojis! https://aidungeon.medium.com/introducing-ai-dungeon-translat... . For actual benchmarks, scratchpad improves G…
-
comment
Comment #30305380
I get the tokenization argument and it may influence it a bit, but I suspect the n-digit math issue has to do more with search the way it samples (in the bpe link gwern references …
-
comment
Comment #30301893
Couple things there where you can see if it improves with the prompt/formatting. E.g. with Davinci (and J a bit but didn't test too much) you can get bette results by: - Using few-…
-
comment
Comment #27498740
An alternative of that is as Oppenheimer once said, sometimes things are secret because a man doesn't like to know what he's up to if he can avoid it
- story
-
comment
Comment #24604290
This has been a systemic issue reported on for years; e.g. reported by the Intercept in 2017 [1] and Atlantic in 2019 [2]. Not really made clear from the story considering the Econ…
-
comment
Comment #24311749
I think that goes back to a Karpathy quote [1], don't know where he got it from [1] https://twitter.com/karpathy/status/961307010246492160
-
comment
Comment #24246035
My thinking there wasn't because of BPEs, I think it's a graph traversal issue.
- comment
-
comment
Comment #24245712
Also, better prompt design if you have make implicit meaning explicit can improve the WiC score ( http://gptprompts.wikidot.com/linguistics:word-in-context ) and ANLI score ( http:…
-
comment
Comment #24245637
Just priming an immediate availability response is likely going to get poor results. On the other hand, this does bring up an important point, which is that few people have been sy…
- comment
-
comment
Comment #24103747
Great article. It's worth considering the Japanese reaction to the bomb - people thought it was cluster munitions or other conventional explosives leading to them taking sub-optima…
-
comment
Comment #24012586
I don't think it's going to be about a single prompt; reverse engineering multiple prompts interacting with themselves is hard. There's a lot of cool things to be done with: (a) cr…
- comment
- comment
-
comment
Comment #23434503
I think it would be jumping to conclusions to say because the military has vulnerabilities that the entire military industrial complex is a fraud. Pieces of it certainly are sub-op…
-
comment
Comment #23127048
That's the dataset used in the paper.
-
comment
Comment #22624961
Leadership in the US Government wasted a lot of time trying to play down the pandemic [1] as did large (conservative) news organizations [2]. Both are now trying to blame China, ar…