Viewing profile — mlin4589
mlin4589
HN member- Joined
- Mon, Oct 10, 2022, 2:48 AM UTC
- HN karma
- 12
- Public activity
- 7 items
- HN profile
- View on Hacker News ↗
About mlin4589
No profile information was provided.
Recent public activity
-
comment
Comment #43919127
The reality, I suspect is that internally models are likely modeling these alignment features such as refusals as a secondary filter. In fact, for many models you can remove refusa…
-
comment
Comment #43919062
Calibration (in a binary context) basically means that the confidence of a model/score matches the probability that a particular label is positive or not. For instance, a calibrate…
-
comment
Comment #43912267
Good question! We do know from OpenAI's system card from GPT-4 that the post-trained RLHF model is significantly less calibrated compared to the pre-trained model, so it's a matter…
-
job
Intrinsic (YC W23) Is Hiring
We're building real-time abuse prevention systems. Hiring Full-stack, frontend, and MLEs across product and PoCs. https://withintrinsic.com/careers
- comment
- story
- comment