Viewing profile — ziaowang
ziaowang
HN member- Joined
- Sun, Sep 22, 2024, 12:52 AM UTC
- HN karma
- 7
- Public activity
- 5 items
- HN profile
- View on Hacker News ↗
About ziaowang
Recent public activity
-
comment
Comment #43327332
Texts in the wild used during pre-training contain lots of biases, such as racial and sexual biases, which are picked-up by the model. During RLHF, the human evaluators are aware o…
-
comment
Comment #43326545
This understanding is incomplete in my opinion. LLMs are more than emulating observed behavior. In the pre-training phase tasks like masked language model indeed train the model to…
-
comment
Comment #42875111
Can you provide a link to the comment? R1's technical report ( https://github.com/deepseek-ai/DeepSeek-R1/blob/main/DeepSee... ) says the prompt used for training is " reasoning pr…
-
comment
Comment #42219916
Agreed. If it's useful, why not scale up the electricity and water supply, and make the latter sustainable.
-
comment
Comment #42132878
Though FAANG offers are usually more attractive than startups (considering pay level and stability), some startups could be more selective since they couldn't afford to hire the wr…